메뉴
BL
The Decoder • 39일 전

안스로픽, 클로드 출력에 워터마크 삽입… 품질 저하 논란

IMP
7/10
핵심 요약

안스로픽이 EU 규제 대응 차원에서 클로드의 텍스트에 통계적 패턴 기반 워터마크를 삽입했으나, 블로거 존 그루버 등은 단어 선택이 의미가 아닌 워터마크 키에 따라 이뤄져 텍스트 품질이 떨어진다고 비판합니다. 법률 업계는 대부분 큰 문제가 없다고 보지만, AI 사용 금지 계약이나 수수료 협상에서 AI 기여도가 검증 가능해지는 상황은 주의가 필요합니다.

번역된 본문

안스로픽이 클로드 출력에 워터마크를 삽입하지만, 비판론자들은 그 대가를 의문한다

핵심 요점

  • 안스로픽은 EU 규제를 준수하기 위해 클로드에 워터마크를 삽입해 단어 선택의 통계적 패턴을 통해 AI 텍스트를 식별한다.
  • 비판론자들은 모델이 의미가 아닌 워터마크 키에 따라 단어를 선택하기 때문에 텍스트 품질이 저하된다고 지적한다.
  • 다만 Declaude 같은 도구로 이 표시를 제거할 수 있다.
  • 법률사무소에는 대부분 큰 문제가 아니다. 하지만 영구적으로 탐지 가능한 AI 텍스트는 수수료 협상이나 계약서에 AI 사용을 명시적으로 금지한 경우 곤란해질 수 있다.

안스로픽의 클로드 텍스트 워터마킹은 AI 생성 콘텐츠를 탐지 가능하게 만드는 것이 목적이다. 하지만 비판론자들은 단어 선택이 영향을 받지 않는다고 믿지 않으며, 변호사들은 새로운 투명성 문제에 직면하고 있다.

안스로픽은 클로드의 텍스트 워터마킹이 내용, 창의성, 가독성에 아무런 영향을 주지 않는다고 주장한다. 이 방식은 구글의 SynthID-Text 접근법에 기반하며, 텍스트 생성 시 단어 선택의 무작위성 원천을 조정해 통계적으로 탐지 가능한 패턴을 만들어낸다. 눈에 보이는 표시도, 숨겨진 문자도 삽입되지 않는다.

블로거 존 그루버는 동의하지 않는다. 그는 세련된 파이어볼(Daring Fireball)의 장문 글에서 두 동의어가 정확히 같은 의미를 지니는 경우는 없다고 주장한다. 클로드가 의미적 정확성이 아닌 비밀 워터마크 키에 따라 "overcast(흐림)"와 "grey(회색)" 중 하나를 선택할 때, 텍스트 품질은 필연적으로 저하된다는 것이다. 그루버는 이 시스템이 때때로 더 나쁜 단어의 확률을 높이고 최선의 단어 확률을 낮춘다고 썼다. 그는 이 차이가 "감지할 수 없다"는 안스로픽의 주장이 단순히 틀렸다고 말한다.

그루버는 2002년부터 영향력 있는 테크 블로그 데이어링 파이어볼을 운영해왔다. 또한 현재 챗봇에서 널리 쓰이는 포맷 언어인 마크다운(Markdown)의 공동 개발자이기도 하다.

안스로픽이 인용한 구글 딥마인드가 네이처(Nature)에 발표한 SynthID 연구조차 그루버를 설득하지 못했다. 그는 그 연구에서 측정한 좋아요/싫어요 비율은 텍스트 품질의 유효한 척도가 아니라고 주장한다. "파인애플" 대신 "바나나"라고 썼다고 챗봇에 싫어요를 누르는 사람은 없다는 것이다. 그루버는 제미나이가 클로드와 챗GPT보다 약하다는 평판이 부분적으로는 이미 SynthID가 활성화되어 있기 때문일 수 있다고 추측한다.

법률사무소, 투명성 질문에 대비해야 법률 전문 매체 아티피셜 로이어(Artificial Lawyer)는 안스로픽의 워터마크가 법률사무소에 어떤 의미인지 분석했다. 이 매체의 결론은 대체로 낙관적이다. 워터마크는 무해하며, 대부분의 경우 고객이나 법원도 AI 사용에 반대하지 않는다는 것이다. 오히려 일부 고객은 적극적으로 AI 사용을 요청하고 있다고 분석은 지적한다.

다만 아티피셜 로이어의 평가에 따르면 특정 상황에서는 문제가 까다로워진다. 고객이 자신의 사건에 AI 사용을 명시적으로 금지했거나 판사가 AI에 회의적인 경우, 제출된 문서가 사실상 완벽하더라도 AI 기여가 증명될 수 있다. 이는 소송 절차나 재판 진행에 영향을 미칠 수 있다.

아티피셜 로이어는 워터마크가 텍스트와 함께 이동한다는 점도 지적한다. 복잡한 계약서에 일부 조항은 AI 표시가 남아 있고 나머지는 인간이 작성했을 수 있다. AI 생성 섹션이 포함된 과거 템플릿을 기반으로 만든 미래의 계약서에는 그 표시가 그대로 이어진다. 여러 LLM을 사용하면 하나의 문서에 서로 다른 워터마크가 겹칠 수도 있다.

수수료 협상도 달라질 수 있다. 고객이 AI 덕분에 업무가 쉬워졌다며 할인을 요구할 경우, AI 사용 비중이 원칙적으로 검증 가능해진다고 매체는 지적한다.

안스로픽 자체는 사실 중심 구간에서는 단어 대안이 적어 워터마킹이 더 희소해진다고 밝혔다. 정확성이 중요한 법률 텍스트에는 이것이 관련 있는 유보 사항이지만, 아직 이 주제에 대한 실증 연구는 없다.

이 표시를 제거하고 싶은 사람은 Declaude 같은 패러프레이징 도구로 가능하다. 개발자 제임스 파돌시는 기반이 된 EU 규제가 자의적이라고 비판한다. 이 시스템은 주로 일반적인…(원문 누락)

원문 보기
원문 보기 (영어)
Anthropic watermarks Claude's output, but critics question the tradeoffs Maximilian Schreiner View the LinkedIn Profile of Maximilian Schreiner Aug 17, 2026 Nano Banana Pro prompted by THE DECODER Key Points Anthropic is embedding a watermark in Claude that tags AI text through statistical patterns in word choice to comply with EU regulations. Critics say text quality suffers because the model picks words based on the watermark key rather than meaning. Tools like Declaude can strip the marking, though. For law firms, the system is mostly a non-issue. But permanently detectable AI text could get awkward during fee negotiations or when contracts explicitly ban AI use. Ask about this article… Search Anthropic's text watermarking for Claude is supposed to make AI-generated content detectable. But critics doubt that word choice stays unaffected, and lawyers are facing new transparency headaches. Anthropic insists that text watermarking for Claude has no effect on content, creativity, or readability. The method is based on Google's SynthID-Text approach and tweaks the randomness source for word selection during text generation, creating a statistically detectable pattern. No visible marks are inserted. No hidden characters are added. Blogger John Gruber disagrees. In a lengthy post on Daring Fireball , he argues that no two synonyms carry the exact same meaning. When Claude picks between "overcast" and "grey" based on a secret watermark key rather than semantic precision, text quality inevitably suffers. The system sometimes boosts the probability of a worse word and lowers it for the best one, Gruber writes. Anthropic's claim that the difference is "imperceptible" is simply wrong, he says. Ad Gruber has run the influential tech blog Daring Fireball since 2002. He also co-created Markdown, the formatting language now widely used by chatbots. Ad Even the SynthID study that Google Deepmind published in Nature, which Anthropic cites, doesn't convince Gruber. The thumbs-up/down rates measured there aren't a valid gauge of text quality, he argues. Nobody gives a chatbot a thumbs-down because it wrote "bananas" instead of "pineapples." Gruber suggests that Gemini's reputation as weaker than Claude and ChatGPT could partly stem from SynthID already being active. Law firms need to prepare for transparency questions Legal trade publication Artificial Lawyer examined what Anthropic's watermarks mean for law firms. The publication's verdict is mostly relaxed. The markings are harmless, Artificial Lawyer argues, and in most cases neither clients nor courts would object to AI use. Some clients are even actively requesting it now, the analysis notes. Ad Things get tricky in specific situations, though, according to Artificial Lawyer's assessment. If clients have explicitly banned AI use for their cases, or if a judge is skeptical of it, the AI contribution could be proven even when the submitted document is factually flawless. That could have consequences for how proceedings or a trial play out. The watermarks also travel with the text, Artificial Lawyer points out. A complex contract might contain clauses that are AI-marked while a human wrote the rest. Future contracts built on older templates with AI-generated sections carry those markings forward. When multiple LLMs are used, different watermarks could even overlap in a single document. Ad Fee negotiations could shift, too. If clients demand discounts because AI made the work easier, the AI share becomes verifiable in principle, the publication notes. Ad Anthropic itself says that watermarking is sparser in fact-heavy passages because fewer word alternatives exist. For legal texts that depend on precision, that's a relevant caveat, even though there aren't any empirical studies on the topic yet. Anyone who wants to strip the markings can do so with paraphrasing tools like Declaude . Its developer James Padolsey criticizes the underlying EU regulation as arbitrary. The system mostly hits ordinary users while doing little against deliberate circumvention, he says. Anthropic is rolling out watermarking to comply with the EU AI Act but applies it globally because it can't limit the feature by region. All Claude models released after August 2 support the marking. Older models will be retrofitted in the coming months. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: Daring Fireball / Watermark criticism | Artificial Lawyer / Law firm impact