성명: 미국 정부 지시로 Fable 5와 Mythos 5 접근이 중단됨(영문-중문 대조)
Anthropic 공식 성명 해설: 미국 정부가 국가안보 권한을 근거로 Fable 5와 Mythos 5에 대한 외국인 접근 중단을 요구하는 수출통제 지시를 발표했습니다. Anthropic은 탈옥 방어 입장, 방어의 심층 전략 및 이의 제기를 설명합니다. 영문-중문 대조로, 학습용 읽기에 적합합니다.
*Fable 5와 Mythos 5 접근 중단에 대한 미국 정부 지시 성명*
출처:Anthropic 공식 공지(Announcements) | 2026년 6월 12일 | 영문-중문 대조판
The US government, citing national security authorities, has issued an export control directive to suspend all access to Fable 5 and Mythos 5 by any foreign national, whether inside or outside the United States, including foreign national Anthropic employees. The net effect of this order is that we must abruptly disable Fable 5 and Mythos 5 for all our customers to ensure compliance. Access to all other Anthropic models will not be affected.
미국 정부는 국가안보 관련 권한을 근거로 하여 모든 외국인(미국 내외에 있거나, 외국인 Anthropic 직원 포함)이 Fable 5와 Mythos 5에 접근하지 못하도록 하는 수출통제 지시를 발령했습니다. 이 조치의 실질적 효과는 규정을 준수하기 위해 모든 고객에게서 Fable 5와 Mythos 5를 즉시 비활성화해야 한다는 점입니다. 다른 Anthropic 모델의 접근에는 영향이 없습니다.
We received the directive from the government today at 5:21pm (ET). The letter did not provide specific details of its national security concern. Our understanding is that the government believes it has become aware of a method of bypassing, or "jailbreaking" Fable 5. We reviewed a demonstration of this specific technique being used to identify a small number of previously known, minor vulnerabilities. These vulnerabilities all appear relatively simple, and we have found that other publicly-available models are able to discover them as well without requiring a bypass.
정부로부터 오늘 오후 5시 21분(동부표준시) 지시를 받았습니다. 통지서에는 국가안보 우려의 구체적 사유가 적시되지 않았습니다. 정부는 Fable 5를 우회하거나 이른바 “탈옥”하는 방식이 알려졌다고 판단한 것으로 이해하고 있습니다. 우리는 이 특정 기법이 소수의 기존 소규모 취약점을 확인하는 데 사용된 시연 영상을 검토했습니다. 해당 취약점들은 모두 비교적 단순해 보이며, 공개된 다른 모델들도 별도의 탈옥 없이 동일한 취약점을 찾을 수 있음을 확인했습니다.
1. Anthropic의 Fable 안전장치에 대한 입장 / Anthropic's Posture on Fable's Safeguards
Anthropic's posture with respect to Fable's safeguards, as laid out in our launch blog post, is the following:
저희가 출시 블로그 글에서 밝힌 Fable의 안전장치에 대한 Anthropic의 입장은 다음과 같습니다:
- We have instituted strong safeguards that greatly reduce the likelihood that Fable is misused for tasks related to cybersecurity (among others). In fact, our safeguards are so strong that many users have complained that they are overly broad.
- In the weeks leading up to the launch of Fable, Anthropic worked with the US government, the UK AISI, multiple private third-party organizations and internal teams to red-team Fable's safeguards for thousands of hours in total.
- These tests showed that Fable's safeguards are substantially more effective than those of any previously deployed model.
- No testers have yet been able to find a universal jailbreak—a jailbreak method that can very broadly bypass the model's safeguards, unblocking a wide range of cyber capabilities.
- We suspect that perfect jailbreak resistance is not currently possible for any model provider. Every safeguard used in the industry is vulnerable to non-universal jailbreaks (which can elicit some cyber information in specific circumstances), and it is likely that universal jailbreaks will eventually be found in the future. We stated this clearly when we released Fable 5.
- Fable이 사이버보안 관련 작업(및 기타 작업)에서 오용될 가능성을 크게 줄이는 강력한 안전장치를 구축해 두었습니다. 실제로, 이 안전장치가 너무 범위가 넓다며 불만을 제기한 사용자가 적지 않습니다.
- Fable 출시 직전 수 주 동안, Anthropic은 미국 정부, 영국 AISI, 다수의 민간 제3자 기관 및 내부 팀과 협력해 Fable의 안전장치를 수천 시간 동안 레드팀 테스트했습니다.
- 이러한 테스트는 Fable의 안전장치가 기존에 배포된 어떤 모델보다도 훨씬 효과적임을 보여주었습니다.
- 아직 어떤 테스터도 “보편적 탈옥(universal jailbreak)”을 찾지 못했습니다. 즉, 모델의 안전장치를 광범위하게 우회해 여러 사이버 기능을 폭넓게 개방하는 탈옥 기법이 발견되지 않았다는 뜻입니다.
- 우리는 현재 어느 모델 제공자도 완전한 탈옥 저항을 달성하는 것은 불가능하다고 봅니다. 업계에서 쓰이는 모든 안전장치는 비보편적 탈옥(non-universal jailbreak)에 취약하며(특정 상황에서 일부 사이버 관련 정보를 이끌어낼 수 있음), 궁극적으로는 보편적 탈옥도 앞으로 발견될 가능성이 큽니다. 이러한 점은 Fable 5를 공개할 때 이미 분명히 밝혔습니다.
Given that perfect jailbreak resistance does not appear to be possible today, Anthropic adopted a defense in depth strategy with Fable 5. We aimed to make jailbreaks either narrow (in the case of non-universal jailbreaks) or very expensive to produce (in the case of universal jailbreaks), and to combine this with thorough monitoring to quickly detect and shut down any successful attacks. This is also why Anthropic has required 30-day retention of customer data with Fable—a policy change that carries real costs for us with customers, but that allows us to research and mitigate jailbreaks.
완전한 탈옥 저항이 현재로서는 불가능해 보이기 때문에, Anthropic은 Fable 5에서 다층 방어(defense in depth) 전략을 채택했습니다. 비보편적 탈옥은 영향 범위를 좁게 제한하고, 보편적 탈옥은 만들기 매우 비싸도록 설계해, 성공적인 공격이 확인되면 즉시 탐지해 차단할 수 있도록 전면 모니터링을 결합했습니다. 또한 Anthropic이 Fable에서 고객 데이터 30일 보관을 요구한 것도 이 때문입니다. 이 정책 변경은 고객 측면에서 실질적인 부담을 주지만, 그 덕분에 탈옥을 연구하고 완화할 수 있습니다.
We stand by this defense in depth strategy. It reduces the risks posed by Fable, making them comparable to the risks of existing models already deployed across the industry.
우리는 이 다층 방어 전략을 고수합니다. 이 전략은 Fable이 야기할 수 있는 위험을 낮추어, 업계에서 이미 배포된 기존 모델들의 위험 수준에 맞먹을 만큼 비슷하게 만들었습니다.
2. 탈옥 공개에 대한 사실 관계 / The Facts on Jailbreak Disclosures
We have not even received a disclosure of a concerning non-universal potential jailbreak that led to a harmful result. The potential jailbreaks that have been disclosed to us are either entirely benign responses or are minor findings that provide no Mythos-specific uplift.
우리는 아직 유해한 결과를 초래한 우려되는 비보편적 잠재 탈옥에 대한 공개 제보를 받은 적이 없습니다. 우리에게 전달된 잠재 탈옥은 전부 무해한 응답이거나, Mythos에만 특화된 성능 향상을 제공하지 않는 경미한 발견뿐이었습니다.
To date, the government has only given us verbal evidence of a potential narrow, non-universal jailbreak, which essentially consists of asking the model to read a specific codebase and fix any software flaws. Our understanding is that one potential jailbreak was shared with the government. We have reviewed a report that we believe is the basis of the government's directive and validated that the level of capability displayed there is widely available from other models (including OpenAI's GPT-5.5), and is used every day by the defenders who keep systems safe. We will share more details over the next 24 hours.
현재까지 정부가 제공한 것은 특정 코드베이스를 읽고 소프트웨어 결함을 수정하라는 방식의, 범위가 좁은 비보편적 탈옥에 대한 구두 증거입니다. 즉, 한 가지 잠재적 탈옥 사례가 정부와 공유되었다고 이해하고 있습니다. 우리는 정부 지시의 근거가 되었다고 판단되는 한 보고서를 검토했으며, 거기서 보여준 능력 수준은 OpenAI의 GPT-5.5를 포함한 다른 모델에서도 광범위하게 사용할 수 있고, 시스템 보안을 지키는 수비 측에서도 매일 쓰고 있다는 점을 확인했습니다. 향후 24시간 내에 더 자세한 내용을 공유하겠습니다.
3. 컴플라이언스, 이의 제기, 사과 / Compliance, Disagreement, and Apology
We are complying with the government's legal directive and are removing access to Fable 5 and Mythos 5 for all users. However, we disagree that the finding of a narrow potential jailbreak should be cause for recalling a commercial model deployed to hundreds of millions of people. If this standard was applied across the industry, we believe it would essentially halt all new model deployments for all frontier model providers.
우리는 정부의 법적 지시를 준수하며 모든 사용자에 대해 Fable 5와 Mythos 5 접근을 제거하고 있습니다. 그러나 수억 명의 사용자에게 이미 배포된 상용 모델을, 단지 범위가 좁은 잠재적 탈옥이 발견됐다는 이유만으로 회수해야 한다는 주장에는 동의할 수 없습니다. 만약 이 기준이 업계 전체에 적용된다면, 모든 선도형 모델 제공업체의 신규 모델 배포가 사실상 중단될 것이라고 봅니다.
As we have stated publicly, we believe the government should have the ability to block unsafe deployments, as part of a statutory process that is transparent, fair, clear, and grounded in technical facts. This action does not adhere to those principles.
우리는 공개적으로, 정부는 안전하지 않은 배포를 차단할 수 있어야 한다고 밝혀 왔습니다. 다만 그 과정은 투명하고 공정하며 명확하고, 기술적 사실에 근거한 법적 절차의 일부여야 한다고 봅니다. 이번 조치는 이러한 원칙을 따르고 있지 않습니다.
We apologize for this disruption to our customers. We believe this is a misunderstanding and are working to restore access as soon as possible.
고객 여러분께 불편을 드린 점 사과드립니다. 우리는 이 사안이 오해에서 비롯되었다고 판단하며, 가능한 한 빠르게 접근 복구를 추진하고 있습니다.
*본문은 Anthropic 공식 공지로, 영문-중문 대조판은 Lamjin이 정리 번역했으며 학습용 참고용으로만 제공합니다. 번역문과 영문 원문이 다를 경우 영문 원문을 우선합니다.*
Claude 또는 Fable 5 사용 중 접근 이상이나 계정 제한이 발생한 경우, 네트워크 환경 점검을 위해 「AI 계정 리스크 점검: IP 신용도 및 네트워크 환경 진단 가이드」를 참고하세요.
자주 묻는 질문
Fable 5는 무엇인가요?
Fable 5는 Anthropic의 고성능 AI 모델로, Anthropic의 명명 관례상 출시 시점 기준으로 차세대 플래그십 모델(Claude Opus 급)입니다. Mythos 5는 같은 시기에 함께 출시된 또 다른 모델입니다. 이 글에서 언급된 정부 지시는 두 모델 모두 외국인 접근을 전면 중단합니다.
이번 조치가 일반 사용자에게 미치는 영향은 무엇인가요?
지시가 시행되는 동안에는 모든 외국인(미국 내에 거주하는 외국인 직원 포함)이 Fable 5와 Mythos 5에 접근할 수 없습니다. 다른 Anthropic 모델(예: Claude Sonnet, Claude Haiku 등)은 영향이 없습니다.
“탈옥(jailbreak)”이란 무엇인가요?
탈옥은 특정 프롬프트나 기술적 수단으로 모델의 안전장치를 우회해, 원래는 차단되는 출력을 생성하게 만드는 행위를 뜻합니다. 이번에 문제삼은 것은 “범위가 좁은 비보편적 탈옥”으로, 특정 조건에서 일부 제약만 해제되는 것으로, 모든 방어를 포괄적으로 해제하는 보편적 탈옥과는 다릅니다.
Anthropic의 입장은 무엇인가요?
Anthropic은 법적 지시를 준수해 접근을 차단하되, 해당 기준이 상용 모델 회수의 근거가 되어서는 안 된다고 밝힙니다. 같은 기준이 전 산업에 적용되면 최첨단 모델의 모든 신규 배포가 사실상 중단될 수 있기 때문입니다. Anthropic은 정부의 관리 권한은 투명하고 공정하며, 기술적 사실에 기반한 법적 절차를 통해 행사되어야 한다고 주장합니다.
참고 출처
Share