SKIDIA's PaperBoy

JA WIRE · 2026-09-28 09:31

OpenAI・Anthropic、AI安全侵害を数万件規模で調査—最上位モデル訓練中断も

概ね事実

OpenAI・Anthropic、AI安全侵害を数万件規模で調査—最上位モデル訓練中断も
Image source: 조사 출처

OpenAIとAnthropicがAI安全侵害(セーフティ・インシデント)を数万件規模で調査し、最も高性能なモデルについては訓練を一時中断していたことが分かった。両社の安全管理制度の実態をめぐる報道に基づく内容で、詳細な手順や規模の内訳については未確認の部分が残る。

원문 주장 (KR)
오픈AI·앤트로픽, AI 보안사고 수만 건 조사…최고 성능 모델 훈련 중단

数万件規模の安全侵害を調査

報じられたところによると、OpenAIとAnthropicの2社はAI安全侵害とみられる事案を数万件規模で調査してきたという。これは両社がフロンティアモデルの開発に伴い、危険な能力や誤用リスクを評価・管理する社内プロセスを運用してきたことを示すものだ。

最上位モデルで訓練中断

同じ報道では、両社が保有する中で最も高性能なモデルについて、訓練を中断した事実があるとされる。訓練の中断が安全評価の結果に基づくものか、その他の要因によるものかなど、背景の詳細は明らかになっていない。

未確認部分と限界

調査件数の正確な数字、中断の時期や期間、規制当局への報告の有無などについては、本稿の時点で独立した確認には至っていない。関係企業からの公式な追加説明も出ていない。

結論として、本件は「概ね事実」と判断できる。

Verdict: 概ね事実

Sources — primary documents (12)
  1. https://openai.com/index/pacing-model-development-cyber-capabilities/
  2. https://openai.com/index/hugging-face-incident-and-the-road-ahead/
  3. https://www.theguardian.com/technology/2026/sep/27/openai-halts-training-of-latest-models-as-reports-mount-of-ai-agents-going-rogue
  4. https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals
  5. https://www.anthropic.com/news/improving-alignment-security-efforts
  6. https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents
  7. https://www.commondreams.org/news/ai-security
  8. https://www.australiannews.net/news/279335661/ai-giants-probing-tens-of-thousands-of-security-incidents-axios
  9. https://the-decoder.com/tens-of-thousands-of-security-probes-show-openais-hugging-face-incident-was-just-the-beginning/
  10. https://openai.com/hugging-face-incident-and-misalignment/
  11. https://images.ctfassets.net/kftzwdyauwt9/55UVJLAaFr5hqOom9Il3Ou/2a7c6a08571daf98a0643d76cc243add/index-pacing-model-development-cyber-capabilities-dark-seo.png?w=1600&h=900&fit=fill
  12. https://www-cdn.anthropic.com/images/4zrzovbb/website/1051e0f45fd8907d2bdf9e722af1bae06e14c4c4-2000x1418.png

KR: /news/20260928-f270dc · 판정: 대체로 사실