
← LessWrong (30+ Karma)4 days ago · 3 min
“Anthropic and OpenAI haven’t published a plan for aligning superintelligence” by Zephaniah Roe
<p> While OpenAI and Anthropic pursue different lines of safety research, they have yet to produce a public-facing document describing concretely how their companies plan to align superintelligence. I think it is underappreciated how this points to general negligence or a lack of openness to third-party feedback. </p><p> By “plan,” I mean a document describing a proposal for technical alignment with at least the level of detail and research effort of AI 2040. Any such plan for technical alignment would likely be flawed in non-obvious ways. But having a proposal that's sensible enough to consider and detailed enough to critique is a good starting point for wiser proposals. Making such a plan public would also create feedback loops for accountability. </p><p> The closest thing to a plan came in 2023, when OpenAI announced their superalignment strategy (also relevant). I do not find this approach particularly convincing, though I do find it laudable that OpenAI explained what they planned to do, who would lead the effort, and what resources would be allocated, at a level of detail which made critique possible. This team no longer exists, and nowadays, as far as I am aware, the research community doesn’t have precise answers [...]</p> <p><i>The original text contained 3 footnotes which were omitted from this narration.</i> </p><p>---</p>
<p><b>First published:</b><br/>
September 13th, 2026 </p>
<p><b>Source:</b><br/>
<a href="https://www.lesswrong.com/posts/QrrEtYpwiHpes3rHd/anthropic-and-openai-haven-t-published-a-plan-for-aligning?utm_source=TYPE_III_AUDIO&utm_medium=Podcast&utm_content=Source+URL+in+episode+description&utm_campaign=ai_narration" rel="noopener noreferrer" target="_blank">https://www.lesswrong.com/posts/QrrEtYpwiHpes3rHd/anthropic-and-openai-haven-t-published-a-plan-for-aligning</a> </p>
<p>---</p>
<p>Narrated by <a href="https://type3.audio/?utm_source=TYPE_III_AUDIO&utm_medium=Podcast&utm_content=Narrated+by+TYPE+III+AUDIO&utm_term=lesswrong&utm_campaign=ai_narration" rel="noopener noreferrer" target="_blank">TYPE III AUDIO</a>.</p>