SKIDIA's PaperBoy

EN WIRE · 2026-09-28 09:28

OpenAI and Anthropic report tens of thousands of AI security incidents; training of top-performing model suspended

Mostly True

OpenAI and Anthropic report tens of thousands of AI security incidents; training of top-performing model suspended
Image source: 조사 출처

OpenAI and Anthropic investigated tens of thousands of AI security incidents, according to a claim reviewed against primary sources, which also indicates that training of the companies' highest-performing model was suspended.

원문 주장 (KR)
오픈AI·앤트로픽, AI 보안사고 수만 건 조사…최고 성능 모델 훈련 중단

Scale of incidents

The claim states that the two AI developers examined security incidents numbering in the tens of thousands. The figure encompasses AI-related security events logged by the companies during the review period.

Training halt

The investigation reportedly led to a suspension of training on the companies' top-performing model. The claim does not specify the duration of the suspension or the precise nature of the security concerns that triggered it, and those details remain unconfirmed.

Open questions

The claim does not identify when the incidents occurred, how the two companies divided or coordinated the work, or what remedial measures followed the training halt. Details on the severity distribution of the incidents are also not provided.

The core assertions — that OpenAI and Anthropic investigated tens of thousands of AI security incidents and that training of their highest-performing model was halted — are supported by the evidence on record, though several specifics remain unresolved. The claim is rated Mostly True.

Sources — primary documents (12)
  1. https://openai.com/index/pacing-model-development-cyber-capabilities/
  2. https://openai.com/index/hugging-face-incident-and-the-road-ahead/
  3. https://www.theguardian.com/technology/2026/sep/27/openai-halts-training-of-latest-models-as-reports-mount-of-ai-agents-going-rogue
  4. https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals
  5. https://www.anthropic.com/news/improving-alignment-security-efforts
  6. https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents
  7. https://www.commondreams.org/news/ai-security
  8. https://www.australiannews.net/news/279335661/ai-giants-probing-tens-of-thousands-of-security-incidents-axios
  9. https://the-decoder.com/tens-of-thousands-of-security-probes-show-openais-hugging-face-incident-was-just-the-beginning/
  10. https://openai.com/hugging-face-incident-and-misalignment/
  11. https://images.ctfassets.net/kftzwdyauwt9/55UVJLAaFr5hqOom9Il3Ou/2a7c6a08571daf98a0643d76cc243add/index-pacing-model-development-cyber-capabilities-dark-seo.png?w=1600&h=900&fit=fill
  12. https://www-cdn.anthropic.com/images/4zrzovbb/website/1051e0f45fd8907d2bdf9e722af1bae06e14c4c4-2000x1418.png

KR: /news/20260928-f270dc · 판정: 대체로 사실