메뉴
BL
The Decoder 46일 전

미국 정부, 안전성 문제로 앤스로픽 AI 모델 전 세계 접근 차단

IMP
9/10
핵심 요약

미국 정부의 국가 안보 명목의 수출 통제 지침에 따라, 앤스로픽(Anthropic)의 핵심 AI 모델인 Fable 5와 Mythos 5의 전 세계 접근이 전면 중단되었습니다. 정부 측은 모델의 안전장치를 우회할 수 있는 탈옥(Jailbreak) 기법이 발견되었다고 주장하는 반면, 앤스로픽은 기존 다른 모델들과 비교해도 위험성이 낮다고 반박하며 이번 조치를 강하게 비판하고 있습니다.

번역된 본문

미국 정부, 앤스로픽에 전 세계 Fable 5 및 Mythos 5 접근 차단 명령 Matthias Bastian / 2026년 6월 13일

핵심 요약

  • 미국 정부는 국가 안보 우려를 이유로 앤스로픽(Anthropic)에 AI 모델 Fable 5와 Mythos 5의 전 세계 접근을 즉각 차단할 것을 명령했습니다.
  • 이 수출 금지 조치는 미국 내외의 모든 외국인을 대상으로 하며, 앤스로픽의 외국인 직원들도 포함되어 사실상 미국 외부에서의 접근이 차단되었습니다.
  • 이 명령은 사용자가 Fable 5의 내장된 안전 가드레일을 우회할 수 있게 해주는 탈옥(Jailbreak) 의심 사례에서 비롯되었으며, 앤스로픽은 이 주장을 공개적으로 반박하고 있습니다.

본문

미국 정부는 국가 안보 우려를 이유로, 전 세계를 대상으로 앤스로픽(Anthropic)의 가장 강력한 AI 모델인 Fable 5와 Mythos 5에 대한 접근을 차단하도록 지시했습니다. 앤스로픽은 명령에 따르고 있지만 공개적으로 반발하고 있습니다.

이 수출 통제 지침은 미국 내외를 불문하고 모든 외국인의 Fable 5 및 Mythos 5 접근을 금지합니다. 앤스로픽 소속 외국인 직원들도 예외 없이 영향을 받습니다. 규정을 준수하기 위해 앤스로픽은 전 세계 모든 고객의 접근을 차단해야만 했습니다. 회사의 성명에 따르면, 다른 모든 앤스로픽 모델은 여전히 사용 가능합니다.

앤스로픽은 이번 조치를 "오해"라고 부르며, 최대한 빨리 접근을 복구하기 위해 노력 중이라고 밝혔습니다. 또한 24시간 이내에 더 자세한 내용을 공유할 계획입니다.

정부의 탈옥 위험 주장과 앤스로픽의 반박

앤스로픽에 따르면, 정부는 Fable 5의 안전 조치를 우회하는 방법을 찾았다고 믿고 있습니다. 앤스로픽은 해당 기법의 데모를 검토해 보았다고 밝히며, 이는 다른 공개된 모델들도 탐지할 수 있는 "이전에 알려진 소수의 사소한 취약점"을 식별하는 데 불과하다고 전했습니다.

현재까지 정부의 구두로만 설명된 잠재적인 탈옥(Jailbreak) 기법은 특정 코드베이스를 읽고 소프트웨어 버그를 수정하도록 모델에 요청하는 방식입니다. 앤스로픽은 지침의 근거가 된 보고서를 검토한 결과, OpenAI의 GPT-5.5를 포함해 다른 모델들에서도 "널리 사용 가능한 수준의 능력"이라고 결론지었습니다. 보안 연구원들은 시스템을 보호하기 위해 매일 이러한 기능을 사용하고 있습니다.

앤스로픽의 사이버 보안 마케팅이 부메랑으로 돌아져

출시 전, 미국 정부와 영국 AI 안전 연구소(UK AISI), 민간 제3자 기관 및 내부 팀이 수천 시간 동안 모델을 테스트했습니다. 앤스로픽은 안전 조치가 "이전에 배포된 어떤 모델보다 훨씬 더 효과적"이라고 말하며, 사용자들은 심지어 제한이 너무 과하다고 불만을 제기하기도 했습니다.

테스트에 참여한 그 누구도 모델의 안전 조치를 광범위하게 우회하고 다양한 사이버 역량을 해제할 수 있는 보편적인 탈옥 기법을 찾지 못했습니다. 하지만 앤스로픽은 현재 어떤 모델 제공업체든 완벽한 탈옥 방어가 불가능하다고도 덧붙였습니다. 이는 대규모 언어 모델(LLM)이 제공하는 공격 벡터가 너무 많아 입증된 사실입니다.

앤스로픽은 특정 상황에서 일부 정보를 추출할 수 있는 비보편적 탈옥 기법에 업계 전반의 모든 안전장치가 취약하다고 밝혔습니다. 이를 잘 알고 있기에 회사는 "심층 방어"라는 전략을 추구해 왔습니다. 즉, 탈옥의 범위를 좁히거나 실행 비용을 높게 만들고, 광범위한 모니터링을 결합하여 성공적인 공격을 신속하게 탐지하고 차단하는 것입니다. 이러한 전략에는 고객 데이터를 30일간 보존하는 것이 포함됩니다. 앤스로픽은 이것이 고객에게 진짜 비용을 발생시키지만, 탈옥 연구 및 완화를 가능하게 한다고 설명했습니다.

이전에 앤스로픽의 공포 마케팅을 비판했던 사람이라면 이 아이러니를 알아차릴 것입니다. 앤스로픽은 몇 달 동안 Mythos급 모델의 사이버 보안 위험성에 대해 큰 소리로 경고하며 자사 모델이 얼마나 우수한지 보여주기 위해 노력했습니다. 그러나 지금은 이미 시장에 출시된 다른 모델들도 유사한 역량을 가지고 있다고 변명해야 하는 상황에 직면했습니다.

앤스로픽, 업계 전체에 위험한 선례가 될 것이라 경고

앤스로픽은 명령에 순응하고 있지만, 향후 산업 전반에 미칠 영향을 우려하며 자사의 입장을 지속적으로 알리고 있습니다.

원문 보기
원문 보기 (영어)
US government forces Anthropic to disable Claude Fable 5 and Mythos 5 for all customers worldwide Matthias Bastian View the LinkedIn Profile of Matthias Bastian Jun 13, 2026 GPT-Image-2 prompted by THE DECODER Key Points The U.S. government has ordered Anthropic to immediately cut off global access to its AI models Fable 5 and Mythos 5, citing national security concerns. The export ban covers all foreign nationals, including Anthropic's own international employees, effectively shutting down access outside the country. The order stems from a suspected jailbreak that the government believes could let users get around Fable 5's built-in safety guardrails, a claim Anthropic has publicly disputed. Ask about this article… Search The US government has directed Anthropic to shut down access to its most powerful AI models, Fable 5 and Mythos 5, worldwide, citing national security concerns. Anthropic is complying but publicly pushing back. The export control directive bans all access to Fable 5 and Mythos 5 by foreign nationals, whether they're inside or outside the US. Even Anthropic's own foreign employees are affected. To comply, Anthropic has to cut off access for all customers worldwide. All other Anthropic models remain available, according to the company's statement . Anthropic calls the move a "misunderstanding" and says it's working to restore access as quickly as possible. The company plans to share more details within 24 hours. Ad Government claims jailbreak risk, Anthropic disagrees According to Anthropic, the government believes it has found a method to bypass Fable 5's safety measures. The company says it reviewed a demo of the technique and found it identifies only "a small number of previously known, minor vulnerabilities" that other publicly available models could also detect. Ad DEC_D_Incontent-1 The potential jailbreak—so far only described verbally by the government—boils down to asking the model to read a specific codebase and fix software bugs. Anthropic says it reviewed the report behind the directive and concluded that the capabilities shown are "widely available from other models," including OpenAI's GPT-5.5 . Security researchers already use these capabilities daily to protect systems. Anthropic's own cybersecurity marketing comes back to bite it Before launch, the US government, the UK AI Safety Institute (UK AISI) , private third-party organizations, and internal teams tested the model for thousands of hours combined. The safety measures are "substantially more effective than those of any previously deployed model," Anthropic says. Users even complained they were too restrictive. Ad No tester has found a universal jailbreak, a method that could broadly bypass the model's safety measures and unlock a wide range of cyber capabilities. But Anthropic also says that perfect jailbreak resistance isn't possible for any model provider right now, a fact well-documented given the sheer number of attack vectors LLMs offer . Every safeguard used across the industry is vulnerable to non-universal jailbreaks that can extract some information in specific cases, Anthropic says. Knowing this, the company pursued a strategy it calls "defense in depth": keep jailbreaks either narrowly scoped or expensive to pull off, combined with broad monitoring to quickly detect and shut down successful attacks. Part of this strategy includes 30-day data retention for customer data, which Anthropic says creates "real costs for us with customers" but enables jailbreak research and mitigation. Ad DEC_D_Incontent-2 Anyone who previously criticized Anthropic for fear-based marketing can see the irony here. The company spent months loudly warning about the cybersecurity risks of Mythos-class models, working hard to show how superior the model is. Now it has to argue that models already on the market have similar capabilities . Ad Anthropic warns of a dangerous precedent for the entire industry Anthropic is complying with the order but making its objections clear. "We disagree that the finding of a narrow potential jailbreak should be cause for recalling a commercial model deployed to hundreds of millions of people." If this standard were applied across the industry, it would effectively halt all new model deployments from every frontier model provider, the company says. In earlier public statements , Anthropic argued that the government should have the power to block unsafe deployments, but through a legal process that is "transparent, fair, clear, and grounded in technical facts." The current action doesn't meet those principles, the company says, hinting that this could become another chapter in the ongoing clash between Anthropic and the US government. The US government recently issued a new executive order that lets AI developers submit their models for government safety review before release. Anthropic welcomed that approach, but the process apparently wasn't in place yet when the directive came down. LLMs remain a weak spot in every cybersecurity setup Jailbreaks and the related problem of prompt injections have been an unsolved security problem since the early days of large language models. No LLM maker is immune. The vulnerability has been known since at least GPT-3 and affects all LLM-based systems. ChatGPT and Claude can still be attacked through prompt injection under certain conditions, even though their makers have added countermeasures . Even targeted security efforts have fallen short. About a year ago, Anthropic built a specialized defense against manipulation attempts and put it through a public jailbreaking challenge . After five days, over 300,000 messages, and roughly 3,700 collective work hours, the system was completely cracked, including a universal jailbreak. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: Anthropic