메뉴
BL
TechCrunch AI • 23일 전

미국 정부, 저작물 학습 논란서 OpenAI 편 들어

IMP
8/10
핵심 요약

뉴욕타임스가 OpenAI를 상대로 제기한 소송에서 트럼프 행정부가 20페이지 분량의 의견서를 제출하며 OpenAI의 무단 저작물 활용 LLM 학습을 옹호했습니다. 의견서는 공정 이용(fair use) 원칙을 근거로 저작물 학습 제한이 미국 AI 산업 경쟁력을 저해한다고 주장했습니다. 이는 판결은 아니지만 AI 업계에 유리한 흐름을 강화하는 신호로 평가됩니다.

번역된 본문

뉴욕타임스가 OpenAI를 상대로 제기한 소송에서, 트럼프 행정부는 ChatGPT 개발사가 허가 없이 저작물을 활용해 LLM을 학습시킨 행위를 옹호하는 20페이지 분량의 의견서를 제출했다. 의견서는 "미국은 AI 활용의 실천과 절차에 대한 글로벌 표준을 정하는 강건하고 경쟁력 있는 인공지능 산업을 지속적으로 발전시키는 데 큰 이해관계를 갖고 있다… 그러므로 미국이 '인공지능 분야의 글로벌 리더십을 유지'하는 것이 중요하다"라고 밝히며, 트럼프 대통령이 작년에 서명한 행정명령을 인용했다.

ChatGPT, Claude, Gemini 같은 챗봇을 구동하는 LLM은 저작권이 있는 책, 기사, 기타 미디어를 포함해 이해하기 어려울 만큼 방대한 데이터베이스로 학습되며, AI 기업들은 이러한 데이터를 허가 없이 수집해왔다. 이번 사건의 뉴욕타임스를 포함한 많은 출판사들은 OpenAI 같은 AI 기업이 자신들의 저작물로 AI 모델을 학습시키는 것이 불법이라고 주장해왔다.

'저작물로 AI를 학습시킬 수 있는가'라는 이 질문은 흑백논리로 답할 수 있는 문제가 아니며, 그래서 이 주제를 둘러싼 광범위한 법적 논쟁이 이어지고 있다. 이러한 논의는 대체로 공정 이용(fair use)에 초점이 맞춰져 있는데, 이는 타인의 저작물을 허가 없이 사용하는 특정 상황에서 합법으로 인정할 수 있도록 저작권법에서 예외를 두는 조항이다. 이번 사건에서 공정 이용 논쟁의 쟁점은 AI 기업의 저작물 활용이 판사가 합법이라고 판결할 만큼 '변혁적(transformative)'인지 여부다. 의견서는 "공정 이용 원칙에 대한 오해 아래 LLM 개발을 제약하는 것은 창의적·과학적 진보를 가로막고 미국의 번영과 경제적 유동성을 저해할 것"이라고 주장한다.

지금까지 AI 학습과 저작권 침해 관련 사건들은 대체로 AI 기업에게 유리하게 흘러갔다. 작년 윌리엄 알숩 판사는 Anthropic에게 자사 AI 모델 학습에 사용된 저작물의 원작자들에게 15억 달러의 저작권 합의금을 지급하라고 명령했지만, Anthropic은 AI 학습 자체로는 처벌받지 않았다. 오히려 이 회사는 학습에 사용할 책을 불법 섀도우 라이브러리에서 해적판으로 구한 행위로 벌금을 물었다. 알숩 판사는 LLM의 학습을 인간이 책을 읽는 것에 비유하며 "작가를 꿈꾸는 독자처럼, Anthropic의 LLM은 작품을 앞질러 복제하거나 대체하기 위해 학습한 것이 아니라 새로운 방향으로 전환해 다른 것을 창조하기 위해 학습했다"라고 판시했다.

이번 트럼프 행정부의 의견서는 판결이 아니다. 해당 사건은 미국 뉴욕 남부연방지방법원에서 심리 중이며, 의견서 작성자들에게는 관할권이 없다. 그러나 트럼프 행정부의 이러한 개입은 여전히 상당한 무게를 실어줄 수 있다.

원문 보기
원문 보기 (영어)
In a lawsuit that the New York Times filed against OpenAI, the Trump administration has contributed a 20-page brief in defense of the ChatGPT maker's unlicensed use of copyrighted material to train its LLMs. "The United States has a strong interest in continuing to develop a robust and competitive artificial intelligence industry that sets the standard for the practice and procedure of AI use globally… As such, it is critical for the United States to ‘retain global leadership in artificial intelligence,'" the brief reads, referencing an executive order that President Donald Trump signed last year. The LLMs powering chatbots like ChatGPT, Claude, and Gemini are trained on incomprehensibly massive databases of published works, including copyrighted books, articles, and other media that AI companies feed into these databases without permission. Many publishers, including the New York Times in this case, have sought to argue that it is illegal for AI companies like OpenAI to train AI models on their copyrighted material. This question — can you use copyrighted material to train an AI? — isn't black and white, hence the extensive legal debate around the subject. These conversations often center on fair use, a carve out of copyright law that makes exceptions for certain scenarios when it can be ruled legal to use someone else's copyrighted work without permission. In this case, the fair use debate addresses whether AI companies' use of copyrighted work is "transformative" enough for a judge to rule it legal. "Constraining LLM development under a misunderstanding of fair use doctrine would thwart such creative and scientific progress while hindering American prosperity and economic mobility," the brief says. So far, cases about AI training and copyright infringement have largely been favorable to AI companies. Last year, Judge William Alsup ordered Anthropic to pay a $1.5 billion copyright settlement to a group of writers whose works were used to train the company’s AI models; but Anthropic wasn't dinged for its AI training. Rather, the company was fined for using illegal shadow libraries to pirate the books it used for training. “Like any reader aspiring to be a writer, Anthropic’s LLMs trained upon works not to race ahead and replicate or supplant them — but to turn a hard corner and create something different,” Judge Alsup wrote, comparing the LLM's training to a human reading a book. This new Trump administration brief is not a ruling, as the case is being tried in the U.S. District Court for the Southern District of New York, and the authors of the brief do not have jurisdiction over. However, this intervention by the Trump administration could still carry weight. Topics AI , OpenAI When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Amanda Silberling Senior Writer Amanda Silberling is a senior writer at TechCrunch covering the intersection of technology and culture. She has also written for publications like Polygon, MTV, the Kenyon Review, NPR, and Business Insider. She is the co-host of Wow If True, a podcast about internet culture, with science fiction author Isabel J. Kim. Prior to joining TechCrunch, she worked as a grassroots organizer, museum educator, and film festival coordinator. She holds a B.A. in English from the University of Pennsylvania and served as a Princeton in Asia Fellow in Laos. You can contact or verify outreach from Amanda by emailing amanda@techcrunch.com or via encrypted message at @amanda.100 on Signal. View Bio October 13 - 15 San Francisco Don't miss out . The startup community will gather to answer a pivotal question: How do you build sustainably in the AI era? REGISTER NOW Most Popular Apple shares ‘shocking evidence' against former employee accused of stealing company data for OpenAI Amanda Silberling Microsoft tests fix for latest hours-long Outlook outage Sarah Perez MapQuest's app surges to No. 1 in Navigation after refusing to rename Lake Ontario Sarah Perez Musk's faster path to more gas turbines comes with pollution problem Connie Loizos Nvidia’s AI advantage is moving beyond the GPU Russell Brandom Hugging Face is selling a cute $399 open source duck robot, Microduck Rebecca Bellan Nvidia closes in on Hugging Face acquisition Connie Loizos
관련 소식