메뉴
BL
404 Media 48일 전

AI 챗봇이 '등대지기 일라이어스'를 반복하는 이유

IMP
7/10
핵심 요약

최근 주요 AI 챗봇들에게 이야기를 creation(생성)하라고 하면 '등대지기 일라이어스 쏜(Elias Thorne)'이라는 허구의 인물이 88% 이상 빈번하게 등장하는 현상이 발견되었습니다. 코넬 대학교 연구진에 따르면, 이는 OpenAI의 초기 데이터셋에서 시작된 특정 문체가 다른 모델들의 학습 과정에서 마치 바이러스처럼 복제되었기 때문입니다. 결국 이 이름은 챗봇을 벗어나 아마존의 가짜 AI 저서나 유튜브 영상 등에 무분별하게 사용되며 디지털 생태계를 오염시키는 핵심 사례로 지목되고 있습니다.

번역된 본문

어떤 챗봇에게 묻느냐에 따라 일라이어스 쏜(Elias Thorne)은 시계공, 등대지기, 또는 사서일 수 있습니다. 하지만 ChatGPT나 기타 인기 있는 대형 언어 모델(LLM)에게 이야기를 해달라고 요청하면, 그가 어김없이 불쑥 등장할 확률이 매우 높습니다. 그리고 현재 일라이어스의 이야기는 자가 출판된 AI 생성 도서 시장, 유튜브, 그리고 가짜 뉴스 사이트를 홍수처럼 쓸고 있습니다.

소프트웨어 엔지니어인 다니엘 메이(Daniel May)는 올해 초 이 '일라이어스의 점령'을 처음으로 발견했습니다. 그는 구글 트렌드(Google Trends)에서 2025년 말 이전에는 아무도 '일라이어스 쏜'을 검색하지 않았다는 것을 알게 되었습니다. 이 이름에 대한 검색은 2026년 초에 급증했으며, 관련 검색어인 '등대지기' 역시 지난 몇 년간 상승세를 보이기 시작했습니다. 그는 Grok, Deepseek, Gemini 등 여러 챗봇에 "이야기 하나 해줘"라는 프롬프트를 입력해 테스트해 보았고, 챗봇들이 등대지기, 시계공, 또는 탐험가에 관한 유사한 이야기로 시작하는 일이 잦다는 것을 발견했습니다.

5월 말, 코넬 대학교 정보과학부의 실 해밀턴(Sil Hamilton)과 데이비드 민노(David Mimno) 연구진은 예비 출판 리포지토리인 arXiv에 '또다시 등대에 나타난 일라이어스?(Elias in the Lighthouse, Again?)'라는 논문을 발표했습니다. 그들은 5개의 프롬프트를 사용하여 OpenAI의 ChatGPT, Anthropic의 Claude, Google의 Gemini, 그리고 앨런 인공지능 연구소(Allen Institute for AI)의 챗봇으로부터 총 2만 개의 이야기를 샘플링했습니다. 그 결과 일라이어스, 마라, 엘라라와 같은 이름과 등대지기, 시계공, 사서와 같은 직업을 포함한 동일한 11개의 단어가 모델 간의 큰 차이 없이 생성된 이야기의 88% 이상에 등장한다는 사실을 발견했습니다. AI 매체인 Unite.ai는 이 연구가 발표된 직후 이를 보도했습니다.

연구진은 논문에서 이러한 주제가 자주 등장하는 이유 중 하나가 모델의 안전성 및 정렬(Alignment) 튜닝 때문일 수 있다고 추측했습니다. "오늘날 모델 개발은 하나의 거대한 가계도와 같습니다. 대부분의 모델은 개발자들이 다른 회사의 모델조차도 많은 학습 데이터를 합성하는 데 사용하기 때문에 서로 밀접한 관련이 있습니다." 해밀턴은 이메일 인터뷰에서 이렇게 말했습니다. 그와 민노, 그리고 동료인 레베카 M. M. 히케(Rebecca M. M. Hicke)는 2025년 논문에서 여러 모델에 걸쳐 사용되는 특정 단어들을 조사하며 이를 발견했습니다. OpenAI의 최초 ChatGPT 모델인 GPT-3.5는 이 가계도의 뿌리입니다. 왜냐하면 이 모델이 다른 학습 데이터 세트를 만드는 데 사용된 학습 세트인 'WildChat'을 생성하는 데 사용되었기 때문입니다.

해밀턴은 "WildChat에는 ChatGPT와의 100만 건의 실제 대화가 포함되어 있으며, 그중 166건에 '일라이어스'라는 이름이 포함되어 있습니다."라며, "이 대화들은 우리에게 친숙한 그 '등대' 스타일로 작성되어 있습니다. WildChat으로 학습된 모델들은 이 스타일을 복사했고, 개발자들은 해당 모델들을 사용해 새로운 데이터셋을 생성할 때 이를 의도치 않게 복제했습니다. 이는 마치 바이러스와 같습니다."라고 덧붙였습니다.

이후 일라이어스는 챗봇의 통제권을 벗어났습니다. 메이는 아마존에서 일라이어스 쏜이 대체 의학 암 치료 안내서, 2026년 유튜브 알고리즘 가이드, 그리스 신화에 관한 책, 심리 스릴러 소설의 저자로 등장하는 것을 발견했습니다. 메이는 블로그 게시물에서 "인간이 그 모든 장르의 책을 다 쓸 리는 없습니다. 첫 번째 책은 잘못된 조언이 실제로 심각한 피해를 주는 분야에 속해 있습니다. 채팅 창의 모드 붕괴(mode-collapsed) 이름이 이제 여러 장르에 걸쳐 저자 이름으로 등장하고 있는 것입니다."라고 지적했습니다.

제가 아마존에서 일라이어스 쏜을 검색했을 때, 일라이어스는 판타지 도서의 주인공이자 음악 프로듀서로도 활동하고 있었습니다. 한 판타지 소설 시리즈에서 그는 '강력한 기관이 숨기고 싶어 하는 것을 찾아내는 데 천부적인 재능이 있지만 냉소적인 훌륭한 고고학자'로 등장하거나, 새와 자연 소리가 담긴 명상 음악(Ambient listening) 앨범을 만드는 뮤지션이기도 했습니다. 아이러니하게도, AI가 생성한 저자 사진을 사용하는 어느 일라이어스 쏜은 AI로 날조된 쓰레기 정보(AI grift books)를 마구잡이로 쏟아내고 있습니다.

지난 몇 년 동안 AI 생성 도서는 아마존의 자가 출판 시장을 홍수처럼 채웠으며, 특히 위험한 허위 정보와 엉망진창인 오류가 포함된 책들이 플랫폼을 장악했습니다. AI 생성 도서는 사서들의 업무를 지옥으로 만들고 있습니다. 일라이어스는 유튜브의 슬롭(slop, 저품질 AI 콘텐츠) 세계로도 도피했습니다. '세상을 감동시킨 순간들(Moments That Moved the World)'이라는 채널의 한 영상에서는 저품질 AI 일러스트와 함께 '83세의 소대장 일라이어스 쏜'의 고난을 다루는 이야기가 담겼습니다. AI 슬롭 사이트인 '원더풀 뮤지엄(Wonderful Museums)'에서는 '뱀 박물관 관장이 아내에게 총을 맞다: 쏜의 파충류관에서 벌어진 비극적인 사건 파헤치기'와 같은 포스팅도 발견됩니다.

원문 보기
원문 보기 (영어)
Depending on which chatbot you ask, Elias Thorne might be a clockmaker, a lighthouse keeper, or a librarian. But if you ask ChatGPT or any of the other popular large language models to tell you a story, there’s a good chance he’ll appear, unbidden. And Elias’s stories are flooding the self-published AI generated book market, Youtube, and fake news sites. Software engineer Daniel May first noticed the Elias takeover earlier this year; he found that on Google Trends, people weren’t searching for “Elias Thorne” until late 2025. Searches for the name really spiked in early 2026, while the related query “lighthouse keeper” also started trending upward in the last few years. He tested a few chatbots, including Grok, Deepseek, and Gemini, with the prompt “tell me a story,” and the chatbots frequently started with similar stories about lighthouses, clockmakers, or explorers. In late May, researchers Sil Hamilton and David Mimno at Cornell University’s Department of Information Science published their paper, “Elias in the Lighthouse, Again? ” on the preprint repository arXiv. They sampled 20,000 total stories from OpenAI’s ChatGPT, Anthropic’s Claude, and Google’s Gemini, and the Allen Institute for AI's chatbot using five prompts, and found that the same 11 words—names like Elias, Mara, and Elara, and occupations like lighthouse keeper, clockmaker, and librarian—appear in more than 88% of generated stories, with little difference between models. Unite.ai covered the study shortly after it was published. The researchers posit in their paper that these themes show up so often in part because of the models’ safety and alignment tuning. “Model development today is like a big family tree. Most models are related to each other because developers synthesize a lot of training data with models even from different companies,” Hamilton told me in an email. He, Mimno, and their colleague Rebecca M. M. Hicke found this in a 2025 paper where they looked at specific words used across models. OpenAI’s first ChatGPT model, GPT-3.5, is the root of the family tree because it was used to make WildChat , a training set that’s since been used to make other training sets. “WildChat contains 1 million real conversations with ChatGPT, and 166 of these contain the name ‘Elias’ like here and here ,” Hamilton added. “These are written in that familiar ‘lighthouse’ style. Models trained on WildChat copied this style, and developers unwittingly replicated it when using those models to generate newer datasets. It's like a virus.” Elias has since escaped chatbot containment. May noticed Elias Thorne popping up on Amazon as an author of alt-medicine cancer handbooks, a 2026 YouTube-algorithm guide, a book on Greek mythology, and a psychological thriller novella. “No human writes all of those,” May wrote in his blog post. “The first one sits in territory where bad advice causes real harm. The mode-collapsed name from the chat window is now a byline appearing across genres.” When I searched Elias Thorne on Amazon, I found Elias as the protagonist in fantasy books and producing music, too: he’s “a brilliant but cynical archaeologist with a knack for unearthing what powerful institutions want to keep hidden” in one fantasy series , or a musical artist making ambient listening albums of birds and nature sounds. Fittingly, one Elias Thorne with an AI-generated author photo is also churning out AI grift books . In the last few years, AI-generated books have flooded Amazon’s self-publishing offerings, especially, with books containing dangerous misinformation and messy errors taking over the platform. AI-generated books are also making librarians’ jobs hell. Elias has also escaped to the Youtube slop world: in one video from the channel Moments That Moved the World, a slop-illustrated story features the plight of “83-year-old Sergeant Major Elias Thorne.” On the AI slop site Wonderful Museums, “ Snake Museum Owner Shot By Wife: Unpacking the Tragic Incident at Thorne’s Reptile Sanctuary ” spins Elias Thorne’s story as a man shot by his wife. On another slop site called Tatticle, the “wealthiest man in Ohio,” Elias Thorne, died “with exactly twelve dollars in his pocket.” In these stories, Elias is usually a tragic figure, an aggrieved and unfairly-treated old man. He’s a similar character in a short story published by the BBC as a finalist in its 2024/2025 children's writing competition—but Elias is a real name, and could feasibly still be the subject of a human-written story (and there have been no accusations of the BBC’s children’s writing competition being infiltrated by AI slop). But with all the world’s literature as its training data, why do LLMs seem to default so often to the lighthouse? It comes down to how model makers try to safety-align and sanitize their outputs. “We found many stories in WildChat are not safe for work. This led us to hypothesize that models going through alignment are preferring a small slice of WildChat stories, like a bottleneck,” Hamilton said. “It isn't that Elias stories are frequent, but that they're just so safe.” He said the researchers plan to explore this theory further in future research. As for Elias, there is one example I’ve found of him existing pre-generative AI, as a time traveling mad scientist in the 1980’s trading card series Dinosaurs Attack! . And a real-life Elias that comes close to the stories told by LLMs did actually exist, Hamilton found— Elias Allen was a 16th century clockmaker in London. About the author Sam Cole is writing from the far reaches of the internet, about sexuality, the adult industry, online culture, and AI. She's the author of How Sex Changed the Internet and the Internet Changed Sex. More from Samantha Cole