Jun 16, 2026 | Reading Notes
AI safety has spent years formalizing corrigibility: the property of accepting corrective intervention, including shutdown. A 2026 synthesis defines it as remaining "amenable to correction, shutdown, or modification if it goes astray." The literature is...