메뉴
HN
Hacker News • 46일 전

클로드(Claude), AI 생성 콘텐츠 표기 방식 공개

IMP
8/10
핵심 요약

Anthropic은 EU AI 법안의 투명성 의무를 준수하기 위해 클로드가 생성한 모든 텍스트에 워터마크를, 이미지 등 파일에는 출처 메타데이터를 포함하도록 설정할 예정입니다. 이러한 기계 판독용 표기는 2026년 8월 이후 출시되는 신규 모델부터 전 세계 모든 제품에 기본 적용되며, 이는 AI 생성물의 출처를 명확히하여 투명성을 높이는 중요한 정책적 변화입니다.

번역된 본문

Anthropic은 생성형 AI 모델 및 시스템 제공자로서 EU AI 법안(AI Act)의 'AI 생성 콘텐츠 투명성에 관한 실무 강령(Article 50(2))'에 서명했습니다. 이 글에서는 이러한 약속을 실제로 어떻게 이행할 계획인지, 콘텐츠 표기는 어떻게 작동하며 그 한계는 무엇인지 설명합니다. 상세한 기술 가이드가 마련되는 대로 이 문서를 업데이트할 것입니다.

EU AI 법안 실무 강령에 따른 Anthropic의 표기 의무가 Claude에 미치는 의미:

  • 새로운 모델은 출시 첫날부터 AI 생성 콘텐츠를 표기합니다. 2026년 8월 2일 이후 EU에 출시되는 Claude 모델은 출시와 동시에 기계 판독이 가능한 표기를 지원합니다. 생성된 텍스트에는 내장형 워터마크가 포함되고, 지원되는 경우 생성된 파일에는 디지털 서명된 출처 메타데이터가 포함됩니다.
  • Claude를 사용하는 모든 곳에서 표기가 작동합니다. 이 표시는 Claude 플랫폼(API), Claude, Claude Code, Claude Cowork, Claude Tag 등 지원되는 Claude 모델의 출력물에 전 세계적으로 적용됩니다. 단, 일부 플랫폼이나 기능은 특정 표기 유형을 지원하지 않을 수 있습니다.
  • Claude의 표기를 감지할 수 있도록 지원합니다. 실무 강령의 요구사항에 따라 사용자 및 제3자가 Claude의 표기를 감지할 수 있도록 지원하며, 향후 공개될 문서를 통해 세부 정보를 공유할 예정입니다.
  • 기존 모델에 대한 지원도 진행 중입니다. 법안에는 2026년 8월 2일 이전에 출시된 Anthropic 모델에 대한 유예 기간이 포함되어 있으며, 이러한 기존 모델에도 표기 기능을 추가하기 위해 노력하고 있습니다.

Claude가 생성한 콘텐츠의 기계 판독용 표기 AI 생성 콘텐츠가 일상화됨에 따라 콘텐츠의 출처에 대한 더 큰 투명성과 신호는 사람들이 소비하는 정보에 유용한 맥락을 제공할 수 있습니다. Anthropic은 투명성을 높이고 법적 의무를 준수하기 위해 Claude가 생성한 콘텐츠에 기계 판독용 표기를 포함하기 위해 노력하고 있습니다.

  • 적용 대상
    • 모델: 2026년 8월 2일 이후 출시되는 Claude 모델은 출시 시점부터 표기를 지원합니다. 그 이전에 출시된 모델에도 표기 지원을 추가하기 위해 작업 중이며, 준비되는 대로 이 문서를 업데이트할 것입니다.
    • 제품: Claude 플랫폼(API), Claude, Claude Code, Claude Cowork, Claude Tag 등 지원되는 모델의 출력물에 전 세계적으로 표기가 적용됩니다. 내장형 워터마크는 생성된 모든 텍스트에 적용됩니다. 출처 메타데이터는 Claude가 파일 처리를 지원하는 곳에 적용됩니다.
    • 클라우드 파트너: AWS, Google Cloud 또는 Microsoft Foundry를 통해 지원되는 Claude 모델에 액세스할 때 내장형 워터마크가 적용됩니다. 서명된 출처 메타데이터는 각 플랫폼이 제공하는 기능에 따라 모든 플랫폼에서 지원되지 않을 수 있습니다.
    • 지역: 표기는 전 세계적으로 Claude가 제공되는 모든 곳의 지원 모델 출력물에 적용됩니다.

Claude의 콘텐츠 표기 방식 Claude는 생성 및 처리된 콘텐츠를 표시하기 위해 두 가지 보완적 기술을 사용합니다: (1) 텍스트에 내장된 워터마크, (2) 파일에 첨부된 서명된 출처 메타데이터.

  1. 텍스트 내 내장형 워터마크 지원되는 Claude 모델이 텍스트를 생성할 때, 텍스트 자체에 사람의 눈에 보이지 않는 워터마크를 직접 위빙(weaving)합니다. 이는 의미, 품질 또는 가독성을 변경하지 않으며 눈으로 볼 수 없습니다. 워터마크는 텍스트의 일부이므로 다른 곳에 복사하여 붙여넣을 때 함께 전달되며, 일부 편집 과정을 거쳐도 유지될 수 있습니다. 워터마킹은 모델 수준에서 적용되므로 텍스트가 어떤 Claude 제품이나 인터페이스에서 왔는지에 관계없이 항상 존재합니다.

  2. 서명된 출처 메타데이터 Claude가 .svg, .png, .jpg와 같은 지원되는 파일 유형을 생성할 때, 서명된 출처 메타데이터를 첨부합니다. 이 메타데이터는 콘텐츠 출처 및 진위 확인을 위해 업계 전반에서 사용되는 개방형 표준인 콘텐츠 출처 및 진위 확인 연합(C2PA, Coalition for Content Provenance and Authenticity) 표준을 따릅니다. 서명된 메타데이터 라벨이 있으면 해당 파일이 처리되었음을 나타냅니다.

원문 보기
원문 보기 (영어)
Anthropic has signed the EU AI Act's Article 50(2) Code of Practice on Transparency of AI-Generated Content, as a provider of both generative AI models and generative AI systems. This article describes how we’re planning to put those commitments into practice, how marking works, and what its limitations are. We’ll update this article and publish more detailed technical guidance as it becomes available. Anthropic’s commitments under the EU AI Act’s Code of Practice on Transparency of AI-Generated Content What our marking commitments mean for Claude: New models will mark AI-generated content from day one. Claude models launched in the EU on or after August 2, 2026 will support machine-readable marking at launch. Generated text will carry embedded watermarks, and generated files will include digitally signed provenance metadata where supported. Marking works everywhere you use Claude. Marks will apply to output from supported Claude models across Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag, and wherever Claude is offered, worldwide. Some platforms or features may not support certain marking types. We'll help you detect Claude's marks. We'll support users and other third parties to detect Claude’s marks, as the Code requires, and we’ll share details in forthcoming documentation. Existing models are in progress. The law includes a transition period for Anthropic models launched before August 2, 2026, and we’re working to add marking support for those models as well. More details about our marking plans are below. Machine-readable marks in Claude-generated content As AI-generated content becomes commonplace, greater transparency and signals about where content comes from can give people useful context about the information they consume. To support transparency and comply with our legal obligations, Anthropic is working to include machine-readable marks in content that Claude generates. What’s covered Models. Claude models launched on or after August 2, 2026 support marking at launch. We’re also working to add marking support to Claude models released before that date, and we’ll update this article as that becomes available. Products. Claude markings cover output from supported models everywhere you use Claude, including Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag. Embedded watermarks will apply to all generated text. Provenance metadata will apply where Claude supports processing files. Cloud partners. Embedded watermarks will apply when supported Claude models are accessed through AWS, Google Cloud, or Microsoft Foundry. Signed provenance metadata may not be supported on every platform, depending on the features each platform offers. Regions. Marking will apply to output from supported models wherever Claude is offered, worldwide. How Claude marks content Claude uses two complementary techniques to mark content generated and processed by Claude: (1) watermarks embedded in text, and (2) signed provenance metadata attached to files. 1. Embedded watermarks in text When a supported Claude model generates text, it weaves an imperceptible watermark directly into the text itself. You won’t see it, and it doesn’t change the meaning, quality, or readability of Claude’s response. Because the watermark is part of the text, it will travel with the text when it’s copied and pasted elsewhere, and may persist through some editing. Watermarking will be applied at the model level, which means it will be present no matter which Claude product or surface the text comes from. 2. Signed provenance metadata When Claude generates a supported file type, such as a .svg, .png, or .jpg, it will attach signed provenance metadata. This metadata follows the Coalition for Content Provenance and Authenticity (C2PA) open standard, which is used across the industry to record information about content provenance. If a signed metadata label is present, it signals that a file was processed by Claude and lets you detect whether the file has been tampered with. Detecting Claude’s marks We’re also working to enable users and other third parties to detect Claude’s embedded watermarks and provenance metadata. Detection checks whether a piece of text or a file carries a supported Claude mark. If a supported mark is found, it indicates that the content may have been processed by Claude. We’ll share details on detection mechanisms in forthcoming technical documentation. Limitations Machine-readable marks provide important signals about content, but it’s worth understanding their limitations across all content types. A detected mark provides a signal that content was processed by Claude, but is not fully conclusive. Detecting a Claude mark tells you that the content may have been processed by Claude. It does not, on its own, confirm the full provenance of the content. For example: Claude may not be the original author. People often use Claude to proofread, translate, summarize, or convert files. The output can carry a Claude mark even if the underlying ideas, text, or data originated from another source; The content may have changed after Claude processed it. Marked content may be modified, excerpted, or combined with other material after Claude processed it. Lack of a detected mark doesn’t mean the content wasn’t AI-generated or processed. Content generated by Claude may not carry a detectable mark if, for example: It was generated by a model released before marking was supported; The text has been heavily edited, paraphrased, translated, or mixed into other writing; The passage is very short, leaving too little text for a reliable signal; A file’s metadata was stripped through format conversion, re-saving, screenshots, or other means; It was produced through a platform, feature, or file type where a particular marking type wasn’t supported. If you build with Claude If you deploy Claude in your own product, you should independently assess what Article 50 requires of your products and services. Consistent with our commitments under the EU Code, our goal is to support you in meeting your own transparency obligations, and we'll share technical guidance on our marking and detection approach as it becomes available.