메뉴
BL
The Decoder 1일 전

인도 법원, OpenAI 저작권 가처분 신청 기각… AI 학습 공정 이용 인정

IMP
8/10
핵심 요약

인도 델리 고등법원이 대형 통신사 ANI가 OpenAI를 상대로 제기한 저작권 침해 가처분 신청을 기각했습니다. 법원은 챗GPT가 기사를 그대로 복제했다는 ANI의 주장에 증거가 부족하며, AI 모델 학습 과정이 인도 저작권법상 '연구'를 포함한 사적 이용 예외에 해당될 수 있다고 판단했습니다. 이는 전 세계 법원이 AI 학습 데이터 활용의 적법성을 공정 이용으로 인정한 첫 사례로, 향후 AI 저작권 소송에 큰 선례를 남길 중요한 판결입니다.

번역된 본문

원문 제목: 인도 대형 통신사의 저작권 가처분 신청을 기각하며 OpenAI에게 승리를 안겨준 델리 고등법원 소스: 블로그 본문: 인도 대형 통신사의 저작권 가처분 신청을 기각하며 OpenAI에게 승리를 안겨준 델리 고등법원 Matthias Bastian이 작성 Matthias Bastian의 LinkedIn 프로필 보기 2026년 7월 27일

핵심 요약

  • 델리 고등법원은 저작권 침해 혐의와 관련하여 OpenAI에 대한 인도 통신사 ANI의 가처분 신청을 기각했습니다.
  • ANI는 ChatGPT가 자사 기사를 그대로 복제했다는 것을 입증하지 못했습니다.
  • 증거로 제출된 기사들은 AI 모델이 이미 학습을 마친 이후에 작성된 것들이었습니다.
  • 판사는 OpenAI와 통신사가 서로 다른 산업 분야에서 운영되기 때문에 ANI에 대한 경제적 피해가 없다고 판단했으며, 교육, 연구 및 접근성을 위한 언어 모델의 공익적 이점을 확언했습니다.

이 기사에 대해 묻기… 검색

임시 판결에서 델리 고등법원은 OpenAI에 대한 인도 통신사 아시안 뉴스 인터내셔널(ANI)의 가처분 신청을 기각했습니다. 인도 최대 규모의 통신사 중 하나인 ANI는 AI 학습 및 ChatGPT 출력물에 저작물을 무단 사용했다며 OpenAI를 고소했습니다. 아미트 반살(Amit Bansal) 판사는 두 가지 주장 모두에 대해 ANI의 구제 조치 요청을 기각했습니다.

이번 판결은 데이터 암기(Memorization), 검색 증강 생성(RAG), 그리고 AI 학습의 법적 지위를 다룹니다. AI 저작권법 전문가 안드레스 구아다무즈(Andres Guadamuz)는 이번 판결을 OpenAI의 중요한 초기 승리라고 평가했습니다.

광고

ANI의 증거는 자체 저작권 주장을 무너뜨렸다 ANI는 자사 기사를 실질적으로 복사한 것이라고 주장하며 여러 ChatGPT 출력물을 법원에 제출했습니다. 그러나 이는 오히려 화를 자초했는데, OpenAI가 사용된 모델인 GPT-4 및 GPT-4o는 2022년 4월과 2024년 4월까지의 데이터를 바탕으로 학습되었다는 것을 보여주었기 때문입니다. ANI가 인용한 기사는 대부분 2024년 8월과 9월에 작성된 것이었으므로, 학습 데이터에 포함될 수 없었습니다.

광고 DEC_D_Incontent-1

판사의 예비적 견해에 따르면, 이러한 유사성은 검색 엔진처럼 언어 모델이 온라인 정보를 실시간으로 검색할 수 있게 하는 RAG(Retrieval Augmented Generation)에서 비롯된 것으로 보았습니다. ANI는 소장에서 RAG 문제를 다루지 않았기 때문에 법원은 이 문제에 대해 최종적인 판결을 내릴 수 없었습니다. 판사는 RAG 기반의 출력물은 '대중에 대한 전달'로 인정될 수 있으며, 이 문제는 본안 심리에서 다룰 것이라고 밝혔습니다.

ANI의 소송 근거는 여기서부터 더욱 약해졌습니다. 이 통신사는 모델에게 기사를 '정확하게' 재현하라고 명시적으로 지시하는 적대적 프롬프트(Adversarial prompts)를 사용했습니다. 그럼에도 불구하고 ANI는 단 한 건의 완전한 동일 복제본도 제출하지 못했습니다. 판사는 일반적인 뉴스 기사의 사실 관계는 저작권으로 보호되지 않으며, 주제와 헤드라인을 재생하는 것이 이 사건에서 ANI와의 직접적인 경쟁에 해당하지 않는다고 판단했습니다.

광고

또한 이 증거들은 OpenAI가 학습 데이터를 모델에 영구적으로 저장하여 요청 시 통신사의 저작물을 그대로 복제할 수 있다는 ANI의 주장을 뒷받침하지 못했습니다. 그러나 법원은 본안 심리에서 해당 문제를 다시 다룰 예정입니다.

법원, AI 학습을 공정 이용으로 잠정적 판단 ANI는 AI 학습을 위해 자신들의 저작물을 복사하는 것이 저작권 침해에 해당한다는 점을 입증하는 데에도 실패했습니다. 양측은 OpenAI가 학습 중에 ANI 콘텐츠를 사용했다는 점에 동의했지만, OpenAI는 해당 자료가 전체 데이터셋에서 극히 일부를 차지하며, 모델은 문법, 통사론 및 언어 패턴과 같은 비표현적 요소만 추출했다고 주장했습니다.

광고 DEC_D_Incontent-2

판사는 인도 저작권법에 따른 예외 조항을 검토하고 '연구를 포함한 사적 또는 개인적 사용'을 다루는 조항에 의존하여, '연구'를 AI 학습을 포괄할 수 있을 만큼 넓게 해석했습니다.

광고

이 예외 조항이 성립하기 위해 판사는 조건을 설정했습니다. 학습에 사용되는 사본은 합법적인 출처에서 가져와야 하며, 불법 공유 사이트(Shadow libraries)나 허가 없이 접근한 유료 사이트여서는 안 됩니다. 또한 OpenAI는 학습 사본을 공개적으로 배포하지 않았고 내부적으로만 처리했습니다.

구아다무즈에 따르면, 이는 법원이 AI 학습이 사적 사용 예외에 해당한다고 명시적으로 판결한 첫 사례입니다. 법원은 세 부분으로 구성된 공정성 테스트를 거쳤으며, 세 가지 항목 모두에서 OpenAI의 손을 들어주었습니다. 데이터 암기나 복제가 입증되지 않았기 때문에 OpenAI의 ANI 저작물 사용은 학습 목적으로만 제한되었습니다. ANI는 두 회사가 서로 다른 산업 분야에서 운영되기 때문에 경제적 피해를 입증할 수도 없었습니다. 사용이 불법일지라도 (문맥상 생략된 내용으로 보임.)

원문 보기
원문 보기 (영어)
Delhi High Court hands OpenAI a win by rejecting major Indian news agency's copyright injunction Matthias Bastian View the LinkedIn Profile of Matthias Bastian Jul 27, 2026 Key Points The Delhi High Court has rejected Indian news agency ANI's request for a preliminary injunction against OpenAI over copyright infringement. ANI couldn't prove that ChatGPT reproduced its articles verbatim. The articles it submitted as evidence were published after the models had already been trained. The judge found no economic harm to ANI because OpenAI and the news agency operate in different sectors, and affirmed the public benefit of language models for education, research, and accessibility. Ask about this article… Search In an interim ruling, the Delhi High Court rejected a request by Indian news agency Asian News International (ANI) for a preliminary injunction against OpenAI. ANI, one of India's largest news agencies, sued OpenAI over its use of copyrighted material for AI training and in ChatGPT's outputs. Judge Amit Bansal denied the requested relief on both claims. The decision addresses memorization, Retrieval Augmented Generation (RAG) , and the legal status of AI training. AI copyright law expert Andres Guadamuz calls the ruling an important early win for OpenAI. Ad ANI's evidence undermined its copyright claims ANI submitted several ChatGPT outputs to the court that it claimed were substantial copies of its articles. The move backfired because OpenAI showed that the models used, GPT-4 and GPT-4o, were trained on data from April 2022 and April 2024. The articles ANI cited were mostly from August and September 2024, so they couldn't have been part of the training data. Ad DEC_D_Incontent-1 The judge's preliminary view was that the similarities came from RAG, which lets a language model retrieve online information in real time, much like a search engine. ANI hadn't addressed RAG in its filing, so the court couldn't make a final ruling on the issue. The judge said RAG-based outputs could qualify as "communication to the public," a question the court will address in the main proceedings. ANI's case got weaker from there. The agency had used adversarial prompts, explicitly telling the model to reproduce articles "exactly." Even so, ANI couldn't produce a single verbatim copy. The judge found that facts in news articles in general aren't copyrightable and that reproducing topics and headlines didn't amount to direct competition with ANI in this case. Ad The evidence also didn't support ANI's claim that OpenAI permanently stores training data in its models and can reproduce the agency's work verbatim on demand. But the court will revisit that question in the main proceedings. Court tentatively treats AI training as fair use ANI also failed to show that copying its work for AI training amounted to copyright infringement. Both parties agreed that OpenAI had used ANI content during training, but OpenAI argued that the material made up a tiny share of the overall dataset and that the model extracted only non-expressive elements such as grammar, syntax, and language patterns. Ad DEC_D_Incontent-2 The judge looked at exceptions under Indian copyright law and relied on a clause covering "private or personal use, including research," reading "research" broadly enough to cover AI training. Ad For that exception to hold, the judge set conditions. Training copies must come from lawful sources, not shadow libraries or paywalled sites accessed without permission. OpenAI also never made the training copies public and processed them only internally. Guadamuz says this is the first time a court has explicitly found that AI training falls under a private use exception. The court ran a three-part fairness test and sided with OpenAI on all three counts. OpenAI's use of ANI's works was limited to training, since no memorization or reproduction was proven. ANI also couldn't show economic harm because the two companies operate in different sectors. Even when users ask ChatGPT about ANI headlines, the model only returns topics and, at most, a few article titles. The judge cited U.S. cases including Bartz v. Anthropic and Kadrey v. Meta , where language model outputs were deemed transformative. He also pointed to the earlier Google Books ruling . The judge also found that trained language models improve access to information, support education, advance scientific research, help with software development, enable translation, and create tools for people with disabilities. International AI copyright cases paint a mixed picture The Delhi ruling joins a growing list of court decisions worldwide that have reached conflicting conclusions. In the U.S., a judge threw out the lawsuit filed by Raw Story and AlterNet against OpenAI because the plaintiffs couldn't show enough harm and the odds of exact copies were low. That court also held that facts aren't copyrightable. The GitHub Copilot case failed too, with plaintiffs unable to present a single example of identical code. The Intercept, on the other hand, won a partial victory through a DMCA complaint over copyrighted material that had been stripped out before training. In Ross Intelligence v. Thomson Reuters , a court denied fair use because the AI research tool directly competed with Thomson Reuters' legal database Westlaw, making the use non-transformative. The court stressed that this ruling applied only to this non-generative use case and couldn't be extended to large language models. In the Anthropic case , a federal court in San Francisco called AI training with copyrighted works "spectacularly" transformative, a strong signal favoring fair use. But the court drew a "Napster comparison" because Anthropic had used pirated books from shadow libraries as training data. Fair use doesn't cover unlawfully obtained material. Anthropic later paid $1.5 billion to book authors for using those pirated copies. The U.S. Copyright Office rejected the AI industry's argument that training on "vast troves of copyrighted works" broadly qualifies as fair use. The official who wrote the report was fired by the Trump administration shortly after it was published . In Europe, two courts reached opposite conclusions almost simultaneously. The Munich Regional Court ruled in the GEMA case that song lyrics were reproducible in the model weights, making it a copyright-relevant reproduction. The High Court in London dismissed the Getty Images v. Stability AI lawsuit , ruling that an AI model isn't an "infringing copy." New research showing that language models can memorize copyrighted books could ramp up the memorization debate further. Across all these cases, the same core questions remain unresolved. Do AI models permanently store training data? Can training qualify as fair use? Where's the line between lawfully and unlawfully obtained data? And do copies generated through adversarial prompts reflect normal use? AI and media face a problem beyond copyright Beyond copyright, another question looms for the media industry. Even if courts rule that AI training is lawful, AI-powered search products could still gut the news market. A recent Pew Research Center study shows that click-through rates to external websites drop to just 8 percent with Google's AI Overviews, compared to 15 percent without an AI summary. Users tend to stop searching right after getting the AI response. They don't check other sources. For news agencies like ANI, this means AI systems that summarize news and make clicking the original source unnecessary could erode the industry's business model over time, even without direct copyright infringement. The Munich I Regional Court also recently ruled that Google is directly liable for false claims in its AI summaries , since these count as independent content rather than search results. The limited liability that traditionally shielded search engine operators doesn't extend to AI-generated summaries. That ruling could become relevant for ChatGPT's RA
관련 소식
TD
The Decoder 1일 전
IMP 8

OpenAI: 전 직종의 절반, ChatGPT로 타인 업무 대체

OpenAI의 분석에 따르면 직장인들이 자신의 전문 분야가 아닌 다른 직군의 업무를 처리하는 '태스크 크로스오버(Task Crossover)' 현상이 43.5%에 달하는 것으로 나타났습니다. 특히 전문 인력이 부족한 중소기업일수록 마케팅, 엔지니어링, 데이터 분석 등의 타 직무 영역을 AI로 대체하는 경향이 뚜렷했습니다. 이는 기존 직무의 경계가 허물어지고 실제 업무 형태가 빠르게 변화하고 있음을 보여주는 중요한 지표입니다.

오픈에이아이 챗지피티 업무 생산성
TD
The Decoder 1일 전
IMP 8

마이크로소프트, 자체 보안 AI 모델 발표... 여전히 복잡한 작업은 OpenAI에 의존

마이크로소프트가 비용 절감과 성능 향상을 위해 자체적인 컴팩트 보안 AI 모델인 MAI-Cyber-1-Flash를 출시했습니다. 이 모델은 대부분의 보안 작업을 처리하지만, 여전히 가장 복잡한 추론 작업에 대해서는 OpenAI 모델에 의존하는 하이브리드 방식을 취하고 있습니다. 또한, 실시간 위협 모니터링 시스템인 Perception을 도입하며 방대한 보안 데이터를 기반으로 한 AI 오케스트레이터로서의 입지를 강화하고 있습니다.

마이크로소프트 사이버보안 OpenAI
MR
MIT Tech Review 1일 전
IMP 3

OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.

[요약 오류] OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.