메뉴
BL
The Decoder • 48일 전

안스로픽, 클로드 코드 기본 '자동 모드' 도입

IMP
8/10
핵심 요약

안스로픽은 개발자의 승인 대기 시간을 줄이기 위해 AI 코딩 도구인 '클로드 코드'를 기본적으로 자동 모드로 실행하도록 변경했습니다. 내부적으로 위험도를 평가하는 분류기가 작동해 보안을 강화할 뿐만 아니라, 프롬프트 인젝션 공격에 대한 방어 능력도 우수한 것으로 입증되었습니다. 결과적으로 개발자는 직접 코드를 작성하기보다는 AI가 생성한 결과물을 검토하는 역할로 중심이 이동하게 되었습니다.

번역된 본문

클로드 코드(Clause Code)가 곧 기본적으로 자동 모드(Auto Mode)로 실행될 예정입니다. 엔터프라이즈(Enterprise) 고객만 여전히 별도로 옵트인해야 합니다. 이번 변화는 AI가 훨씬 더 많은 개발 과정을 독자적으로 처리하게 된다는 것을 의미합니다.

자동 모드는 AI 코딩 도구가 모든 단계에서 수동 승인을 기다리지 않고 스스로 작업할 수 있게 해줍니다. 내장된 분류기(Classifier)가 해당 작업이 위험하거나 되돌릴 수 없는지 확인하며, 이런 경우에만 사용자에게 확인을 요청합니다. 안스로픽의 블로그 게시물에 따르면, 8월 14일부터 클로드 코드는 Pro, Max, Team 요금제에 한해 자동 모드가 기본으로 활성화된 상태로 출시됩니다.

1,053명의 유료 테스터 및 내부 레드팀(Red-teaming) 테스트 결과, 자동 모드는 수동 승인과 최소한 동일한 수준의 안전성을 보였으며, 대부분의 경우 더 안전한 것으로 나타났습니다. 자동 모드를 사용한 팀은 약 25% 더 많은 풀 리퀘스트(Pull Request)를 생성하여 더 많은 작업을 완료했습니다.

광고 안스로픽은 또한 자동 모드가 프롬프트 인젝션 공격(Prompt Injection Attacks)에 대한 보호 계층을 추가한다고 밝혔습니다. 프롬프트 인젝션 공격은 주입된 악성 코드가 사용자의 원래 지시를 가로채 에이전트를 제어하려는 시도입니다. Trajectory Labs의 독립적인 감사에서는 72가지 공격 시나리오를 각각 10회씩 테스트했습니다. 자동 모드에서 클로드의 현재 모델인 Fable 5, Opus 5, Sonnet 5를 상대로 시도된 총 720회의 공격 중 성공한 것은 단 한 건도 없었습니다. 반면, OpenAI의 GPT-5.6 Sol을 Codex 자동 리뷰 모드에서 테스트했을 때는 5.83%의 공격이 뚫렸습니다.

광고 안스로픽 내부 테스트에서도 자동 모드는 클로드가 기밀 데이터를 공개된 웹페이지에 업로드하는 것을 차단했습니다. 또한 회사 측에 따르면, 한 긴 세션 동안 진행 중이던 GPU 훈련 작업을 방해할 뻔했던 약 2,000개의 프로세스를 종료하기도 했습니다.

안스로픽은 분류기 자체가 소비하는 토큰 비용을 부과하지 않습니다. 하지만 자동 모드를 기본값으로 설정하는 것은 회사 입장에서도 여전히 좋은 수익 모델입니다. 클로드가 더 오래 작동하고 더 많은 작업을 수행할수록 전체 토큰 사용량이 증가하며, 이는 곧 수익 증가로 이어집니다. 비록 수익 증가가 안스로픽의 주된 변화 동기는 아니었을지라도 말입니다.

광고 개발자는 코드를 작성하는 역할에서 AI가 코드를 작성하는 것을 지켜보는 역할로 전환됩니다. 클로드 코드는 현재 압도적인 차이로 가장 널리 사용되는 AI 코딩 도구입니다. 자동 모드를 기본값으로 채택함에 따라 개발자의 역할은 능동적인 코딩에서 AI가 생성한 결과물을 검토하는 방향으로 더욱 밀려나게 됩니다.

안스로픽 스스로도 주의를 당부하고 있습니다. 분류기가 위험을 줄이기는 하지만 완전히 제거하지는 못하기 때문입니다. 회사는 배포 인프라 등 중대한 변경 사항에 대해서는 여전히 개발자가 직접 클로드의 작업을 검토할 것을 권장합니다.

광고 이러한 조언은 하나의 역설을 만들어냅니다. 개발자가 개입하는 횟수가 줄어들수록 그들의 감시와 통제는 더욱 중요해집니다. 하지만 인간의 개입 없이 대부분 자동 모드로 구축된 프로젝트에 대해 깊은 이해를 갖는 것은 점점 더 어려워집니다. 게다가 사이버 보안은 그 어떤 인간도 현실적으로 따라잡기 힘들 정도로 빠르게 변화하고 복잡해지고 있습니다.

광고 과장 없는 진정한 AI 뉴스 - 전문가가 직접 큐레이션합니다. 광고 없는 읽기 환경, 주간 AI 뉴스레터, 연 6회 제공되는 독점 프론티어 보고서인 'AI Radar', 전체 아카이브 액세스, 그리고 댓글 섹션 이용을 원하신다면 THE DECODER를 구독하세요. 지금 구독하세요. 출처: Claude

원문 보기
원문 보기 (영어)
Anthropic sets Claude Code to Auto Mode by default to protect developers from bad approvals Matthias Bastian View the LinkedIn Profile of Matthias Bastian Aug 8, 2026 Nano Banana Pro prompted by THE DECODER Ask about this article… Search Claude Code will soon run in Auto Mode by default. Only Enterprise customers still need to opt in. The change means AI handles even more of the development process on its own. Auto Mode lets the AI coding tool work on its own without waiting for manual approval at every step. A classifier checks whether an action is dangerous or irreversible and only asks for confirmation in those cases. Starting August 14, Claude Code will ship with Auto Mode enabled by default for Pro, Max, and Team plans, as Anthropic announced in a blog post . In tests with 1,053 paid testers and internal red-teaming, Auto Mode performed at least as safely as manual approvals, and often better. Teams using Auto Mode also generated about 25 percent more pull requests, meaning they got more work done. Ad Anthropic also says Auto Mode adds a layer of protection against prompt injection attacks , where injected code tries to hijack the agent away from the user's instructions. An independent audit by Trajectory Labs tested 72 attack scenarios ten times each. None of the 720 attempts succeeded against Claude's current models, Fable 5, Opus 5, and Sonnet 5, in Auto Mode. With OpenAI's GPT-5.6 Sol in Codex Auto-Review mode, 5.83 percent of the attacks got through. Ad DEC_D_Incontent-1 Internally at Anthropic, Auto Mode stopped Claude from uploading confidential data to a public page. During one long session, it also killed roughly 2,000 processes that would have disrupted ongoing GPU training jobs, the company says. Anthropic doesn't charge for the tokens the classifier itself consumes. But making Auto Mode the default is likely still a good deal for the company. When Claude works longer and gets more done, total token usage goes up, and so does revenue, even if that wasn't Anthropic's main motivation for the change. Ad Developers shift from writing code to watching AI write it Claude Code is currently the most widely used AI coding tool by a wide margin, and making Auto Mode the default pushes the developer's role further from active coding toward reviewing AI-generated output. Anthropic itself urges caution. The classifier reduces risks but doesn't eliminate them. "For high-stakes changes to production infrastructure, we still recommend reviewing Claude's actions yourself," the company writes. Ad DEC_D_Incontent-2 That advice creates a paradox. The less often developers step in, the more their oversight matters. But it gets harder to build a deep understanding of projects that were largely built by Auto Mode without much human involvement. And cybersecurity is moving faster and growing more complex than any human can realistically keep up with. Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: Claude