메뉴
BL
TechCrunch AI • 24일 전

앤스로픽, 새 Fable 모델 발표…비용 저렴·규제 완화

IMP
8/10
핵심 요약

앤스로픽이 최신 AI 모델인 Fable과 Mythos 5.1을 발표했습니다. Fable은 토큰 비용 절감과 안전장치의 오탐(誤判) 규제 완화가 핵심 변경점이며, Mythos 5.1은 사이버보안·생명과학 연구 파트너에게만 제공됩니다. 또한 제로 데이터 리텐션(데이터 미보관) 방식 도입으로 고객이 자체 인프라에서 데이터 유출 없이 모델을 운영할 수 있게 됩니다.

번역된 본문

화요일, 앤스로픽(Anthropic)은 회사의 가장 진보된 AI 모델의 쌍둥이 버전인 Fable과 Mythos 5.1을 공개했습니다. 성능 업그레이드와 더불어, 새로운 Fable 릴리스에는 모델의 토큰 비용과 안전장치로 인한 오탐(誤判) 제한을 줄이기 위한 변경 사항이 포함되어 있습니다.

이전 Mythos 모델과 마찬가지로, Mythos 5.1은 사이버보안 또는 생명과학 연구에 종사하는 등록된 앤스로픽 파트너에게만 제공됩니다. 규제가 완화된 버전인 Fable 5.1은 오늘부터 클라우드 플랫폼 또는 앤스로픽 API를 통해 이용할 수 있습니다.

가장 중요한 변화 중 하나는 앞서 보도된 바 있는 앤스로픽의 제로 데이터 리텐션(데이터 미보관) 정책 도입으로, 고객이 데이터 유출 없이 자체 인프라에서 앤스로픽 모델을 실행할 수 있게 된 것입니다. 보안 우려로 인해 이전에는 Fable에서 제공되지 않았던 고프라이버시 서비스(Enterprise Frontier Safeguards)는 6월부터 사용자들에게 순차적으로 제공될 예정입니다. 주목할 점은, 이 시스템이 여전히 AI 에이전트나 인간 사용자의 악용 여부를 모니터링하지만, 모니터링 방식은 고객이 통제하게 된다는 것입니다.

발표의 일환으로, 앤스로픽은 고객 데이터가 부적절하게 접근된 적이 없다고 보장했습니다. 발표문에 따르면 “앤스로픽은 명시적인 허가 없이 기업 데이터로 학습한 적이 없으며, 앞으로도 결코 그렇게 하지 않을 것입니다.”

앤스로픽 릴리스의 관례처럼, 새 모델은 CLI 기반 코딩을 측정하는 Terminal-Bench 4.0과 일반 추론 능력을 평가하는 Humanity's Last Exam 등 다양한 벤치마크에서 신기록을 세웠습니다. 앤스로픽은 또한 모델 출시 전에 생성된 세 가지 새로운 과학적 성과를 공개했는데, 여기에는 커스텀 GPU 최적화와 기존 사진을 조합해 만든 금성의 고해상도 지도가 포함됩니다.

이전 릴리스와 마찬가지로, 이 모델에는 모델의 능력을 가장 명확하게 설명하는 상세한 시스템 카드가 함께 제공됩니다. 시스템 카드는 AI가 스스로를 개선하는 ‘자동화된 AI 개발’ 우려와 관련해 Mythos를 '저위험'으로 평가했습니다. 일부에서는 이러한 자가 개선이 인간의 통제력 상실을 촉발할 수 있다고 봅니다. 시스템 카드는 “내부 AI R&D 진전을 가속화하는 능력이 현재 추세와 일치한다”고 밝혔습니다.

일반적인 문제 행동 측면에서 Mythos는 Opus보다 다소 더 그러한 경향이 있는데, 이는 향상된 능력 때문일 수 있습니다. 시스템 카드는 “Mythos 5.1은 전체적으로 정렬되지 않은 행동 측면에서 Opus 5 대비 다소 퇴보했으며, Mythos 5와 Claude Sonnet 5 대비 개선되었다”고 설명합니다. “인간의 악용에 협조하고 검증 불가능한 권한 주장을 Opus 5보다 다소 더 쉽게 받아들이지만, 명시적 제약을 무시하거나, 입력을 환각(할루시네이션)하거나, 작업 완료를 거짓으로 주장할 가능성은 이전 모델보다 낮습니다.”

원문 보기
원문 보기 (영어)
On Tuesday, Anthropic released Fable and Mythos 5.1, twinned versions of the company's most advanced AI model. In addition to performance upgrades, the new Fable release includes changes meant to reduce token cost and false-positive restrictions from the model's safeguards. As with the previous Mythos model, Mythos 5.1 will only be available to registered Anthropic partners engaged in either cybersecurity or life sciences research. Fable 5.1, the unrestricted version, is available starting today on cloud platforms or through the Anthropic API. One of the most significant changes is Anthropic's previously reported embrace of zero data retention, allowing clients to run Anthropic models on their own infrastructure without data outflows. Previously unavailable for Fable due to security concerns, a high-privacy service (called Enterprise Frontier Safeguards ) will now roll out to users in June. Notably, the system will still monitor for misuse by agents or human users, but clients will control how the monitoring takes place. As part of the announcement , Anthropic assured customers that their data had not been inappropriately accessed. "Anthropic has never trained on enterprise data without explicit permission, and never will," the announcement reads. As is common for an Anthropic release, the new models set records in a range of benchmarks, including Terminal-Bench 4.0 (for CLI-based coding) and Humanity's Last Exam (for general reasoning). Anthropic also released three novel scientific findings generated by the models before their release, including a custom GPU optimization and a high-resolution map of Venus assembled from existing photos. As with previous releases, the models come with a detailed system card , which explains their capabilities in most straightforward terms. The system card rates Mythos as "low-risk" for concerns related to automated AI development — where the AI improves itself — which some see as a trigger for a loss of human control. It says "its ability to accelerate internal AI R&D progress is in line with current trends." In terms of general misbehavior, Mythos is slightly more prone to it than Opus, possibly as a result of its enhanced capabilities. "Mythos 5.1 is a slight regression on overall misaligned behavior compared to Opus 5, and an improvement over Mythos 5 and Claude Sonnet 5," the system card reads. "It cooperates with human misuse and accepts unverifiable claims of authorization somewhat more readily than Opus 5, but it is less likely to ignore explicit constraints, hallucinate inputs, or falsely claim to have completed tasks than previous models." Topics AI , Anthropic , fable , Mythos When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Russell Brandom AI Editor Russell Brandom has been covering the tech industry since 2012, with a focus on platform policy and emerging technologies. He previously worked at The Verge and Rest of World, and has written for Wired, The Awl and MIT's Technology Review. He can be reached at russell.brandom@techcrunch.com or on Signal at 412-401-5489. View Bio October 13 - 15 San Francisco Don't miss out . The startup community will gather to answer a pivotal question: How do you build sustainably in the AI era? REGISTER NOW Most Popular Microsoft tests fix for latest hours-long Outlook outage Sarah Perez MapQuest's app surges to No. 1 in Navigation after refusing to rename Lake Ontario Sarah Perez Musk's faster path to more gas turbines comes with pollution problem Connie Loizos Nvidia’s AI advantage is moving beyond the GPU Russell Brandom Hugging Face is selling a cute $399 open source duck robot, Microduck Rebecca Bellan Nvidia closes in on Hugging Face acquisition Connie Loizos Viral AI startup Instinct has raised $350M at a $2.5B valuation Lucas Ropek
관련 소식