True
OpenAI has acknowledged that it cannot currently express confidence in the safety of its AI systems, and has committed to making cases of "alignment failure" public on an ongoing basis, according to a claim verified by this desk through a review of nine primary sources, including material captured directly from the company's dedicated alignment portal.
The verification centered on content published at alignment.openai.com, OpenAI's standalone website for its alignment work. A browser capture of the site substantiated the claim, with a source chart hosted at alignment.openai.com/assets/reports/source-chart.png among the evidence collected. In total, nine primary sources were reviewed before a final determination was reached.
The claim rests on two points. First, OpenAI does not believe it can yet assure the public that its AI systems are safe. Second, the company intends to disclose instances of "alignment failure" — cases in which an AI system departs from intended behavior — regularly, rather than only on an occasional or ad hoc basis.
The existence of a dedicated alignment portal underscores that alignment — the effort to ensure AI systems behave as designed — remains a distinct, public-facing workstream for OpenAI. A standing commitment to publish failure cases would, if carried out, give researchers, policymakers and users a recurring view into the limitations of current systems, rather than leaving such shortcomings to surface only through outside discovery.
On the basis of nine primary sources and a direct capture of OpenAI's alignment portal, the claim that OpenAI says it cannot be confident about AI safety and will regularly disclose "alignment failure" cases is True.