Skip to content
View original post on X: Dwarkesh Patel· 22/100AI score22/100

How will we know AI models are aligned as RSI begins?

AISummary

Dwarkesh Patel asks how researchers could verify that AI models are aligned once recursive self-improvement, or RSI, starts and accelerates AI progress. The post raises the question without offering a method or answer.

Post on XView on X
@dwarkesh_sp

When we're at the foothills of RSI, and we're about to kick off a period of accelerated AI progress, how will we actually know that the models are aligned?

Source: Dwarkesh Patel · x.comPublished · added here