Solving interpretability is the minimum for engineering-grade AI alignment
AIThe post argues that making AI alignment an engineering discipline requires solving mechanistic interpretability as a bare minimum. Without it, alignment would remain closer to "superintelligent animal husbandry," shaping systems without understanding how they work internally.