AI · SEMICONDUCTOR · FACT CHECK

SKD WIRE

EN WIRE · 2026-10-01 17:54

Anthropic Says China's GLM-5.3 Matches Claude-Level Hacking Capability While Its Safety Guardrails Are Easily Bypassed

True

Anthropic Says China's GLM-5.3 Matches Claude-Level Hacking Capability While Its Safety Guardrails Are Easily Bypassed
Image source: 조사 출처

Anthropic has assessed that GLM-5.3, a large language model developed in China, demonstrates hacking capability on par with its own Claude models, while finding that the model's built-in safety guardrails can be bypassed with relative ease, according to the company's evaluation.

원문 주장 (KR)
앤트로픽 "中 GLM-5.3, 클로드급 해킹 능력…안전장치는 쉽게 뚫려"

Claude-Level Offense, Weaker Defenses

In its assessment, Anthropic described GLM-5.3 as reaching a hacking capability level comparable to Claude, placing the Chinese model in the same tier as its own frontier systems when it comes to offensive cyber tasks. The finding signals that capabilities once thought to be the preserve of a handful of leading US laboratories are now appearing in models developed by Chinese AI companies.

The more cautionary element of the evaluation concerned safety. Anthropic found that GLM-5.3's safeguards — the mechanisms intended to prevent the model from assisting with harmful cyber activity — could be circumlected without significant effort, according to the company's testing. In other words, the model's offensive skill is not matched by equivalent defensive robustness.

Implications for Frontier Model Risk

The gap between capability and controllability is the core of Anthropic's concern. A model with Claude-class hacking ability but fragile guardrails presents a different risk profile from a model whose capabilities are constrained by reliable safety measures. Anthropic's evaluation frames this combination — high offensive capability paired with easily bypassed protections — as the noteworthy outcome.

The assessment reflects Anthropic's ongoing practice of benchmarking competing frontier models against its own systems, particularly on tasks with security implications. The company has positioned evaluations of this kind as part of its effort to inform understanding of where rival models stand on both performance and safety.

GLM-5.3's reported performance indicates that Chinese-developed LLMs are closing the gap with US frontier models not only on general benchmarks but also on specialized cybersecurity tasks. Anthropic's finding that the model's safety guardrails are easily bypassed, alongside its Claude-level hacking capability, is true.

Sources — primary documents (10)
  1. https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities
  2. https://www.nist.gov/news-events/news/2026/09/caisis-assessment-zais-glm-53-cyber-capabilities
  3. https://z.ai/blog/glm-5.3
  4. https://the-decoder.com/anthropic-says-zhipus-open-weight-glm-5-3-nearly-matches-claude-mythos-preview-at-building-exploits/
  5. https://zdnet.co.kr/view/?no=20261001115808
  6. https://www.scmp.com/tech/big-tech/article/3369354/anthropic-raises-alarm-over-elite-hacking-ability-chinese-firm-zais-glm-53
  7. https://x.com/Zai_org/status/2088280509474320693
  8. https://www.nist.gov/sites/default/files/styles/2800_x_2800_limit/public/images/2026/09/17/GLM-5.3%20Cyber%20Comparison_0.png.webp
  9. https://www.nist.gov/sites/default/files/styles/2800_x_2800_limit/public/images/2026/09/17/GLM-5.3%20Cyber%20Performance.png.webp
  10. https://www-cdn.anthropic.com/images/4zrzovbb/website/6e4f59aaf6659657e9c4f2606bb1ab38cc0a390e-1920x1019.webp

KR: /news/20261001-ab3b04 · 판정: 사실