메뉴
BL
404 Media 18시간 전

서브스택 AI 탐지 도구, '마녀사냥' 논란

IMP
7/10
핵심 요약

뉴스레터 플랫폼 서브스택(Substack)이 AI가 작성한 글을 판별하는 새로운 탐지 도구를 도입하자, 작가들은 부정확한 탐지 결과로 인해 명예가 훼손될 수 있다며 강하게 반발하고 있습니다. 서브스택은 AI 사용 자체를 금지하는 것은 아니며 작가가 자체적으로 제작 과정을 밝힐 수 있는 기능을 지원하지만, 탐지 도구의 낮은 정확도와 오남용에 대한 우려는 여전히 남아있는 상황입니다.

번역된 본문

서브스택(Substack) 기고자들은 뉴스레터 플랫폼 내에서 AI가 생성한 글에 표시를 남기는 회사의 결정을 거부하며, 일부 크리에이터들은 새로운 AI 탐지 기능이 일종의 '마녀사냥'이라고 비판했습니다. 글쓰기나 기사 편집에 AI를 사용하는 일부 서브스택 사용자들은 이 새로운 정책이 자신들을 부당하게 차별한다고 말하는 반면, 글쓰기에 AI를 전혀 사용하지 않는 다른 사용자들은 서브스택이 자신들의 글을 AI가 생성한 것으로 잘못 표시하여 명예를 훼손할까 봐 우려하고 있습니다.

'Backstage Pass'라는 소규모 서브스택을 운영하는 맥 콜리어(Mack Collier)는 다음과 같이 적었습니다. "창작 과정에 AI를 사용한 것에 대해 사과할 생각은 없습니다. 저는 AI 없이 20년 동안 글을 써왔고, 원한다면 다시 그렇게 할 수도 있습니다. 산출량은 줄어들고, 게시물의 구조는 덜 체계적이 될 것이며, 훌륭한 편집자가 필요하다는 사실이 더 명백해질 것입니다. AI를 사용하면 제 글의 전반적인 품질이 향상됩니다. 그래서 사용하는 것입니다."

고스트라이터이자 디지털 작문 코치인 앨리스 르메(Alice Lemee)는 서브스택의 새로운 AI 탐지 기능에 대해 만든 영상에서 다음과 같이 말했습니다. "이러한 탐지기들은 정말 악명 높을 정도로 엄청나게 부정확합니다. 단 한 번의 허위 고발만으로도 작가의 명예가 돌이킬 수 없을 정도로 훼손될 수 있습니다."

서브스택은 화요일 크리스 베스트(Chris Best) CEO의 게시물을 통해 AI 탐지 기능을 발표했습니다. 그는 "이제 인터넷에서 무엇이 진짜인지 구별하기가 점점 더 어려워지고 있으며", "누구에 의해서도 만들어지지 않은 콘텐츠가 인간이 있어야 할 인터넷의 영역을 장악하게 되면, 공동체의 공간을 오염시키고 인간의 목소리를 발견하고 듣기 어렵게 만든다"고 말했습니다. 베스트는 서브스택이 AI 탐지 도구인 '팬그램(Pangram)'과 파트너십을 맺었다고 설명했습니다. 이제 이 도구는 서브스택에 기본 탑재되어 모든 사용자가 글을 스캔할 수 있게 해주며, 팬그램은 글의 얼마나 많은 부분이 AI 또는 인간에 의해 생성되었는지 백분율 점수로 산출합니다. (참고: 팬그램은 이전에 404 Media에 광고를 게재한 적이 있습니다.)

우리는 이전에 팬그램의 연구나 해당 AI 탐지기에 의존하는 연구를 보도하며, AI가 생성한 글이 인터넷의 모든 구석을 어떻게 잠식하고 있는지 보여준 바 있습니다. 하지만 우리가 이전에 보도했듯이, AI 탐지기는 완벽하지 않으며 팬그램 자체도 인간이 작성한 콘텐츠를 AI가 만든 것으로 표시하거나, 반대로 AI가 만든 것을 인간이 쓴 것으로 표시하는 오류를 벗어날 수 없습니다. 팬그램의 CEO인 맥스 스페로(Max Spero)는 최근 우리와의 인터뷰에서 회사가 오류를 최소화하기 위해 끊임없이 노력하고 있으며, 오탐률(거짓 양성 비율)을 대략 10,000분의 1로 추정하고 있다고 말했습니다.

교수이자 'Slow AI'의 저자인 샘 일링워스(Sam Illingworth)는 자신의 서브스택에 '서브스택의 AI 탐지기와 마녀사냥의 귀환'이라는 제목의 게시물에서 다음과 같이 적었습니다. "우리는 이제 거울과 같은 형태의 시스템을 만들었습니다. 팬그램은 확률론적 추측을하기 위해 방대한 양의 텍스트로 학습된 진짜 기계이며, 우리는 사람이 그 안에 숨어 있는지 알아내기 위해 이 기계를 여러분의 글에 겨누었습니다. 서브스택은 반대편에 인간이 있는지 결정하기 위해 기계에게 묻고 있습니다."

서브스택은 또한 사용자가 자신의 창작 과정을 설명하고 AI 사용 여부를 공개할 수 있는 'How I make this(이 글을 어떻게 만들었는가)' 명시 기능을 추가했습니다. 베스트는 "우리는 사람들이 업무를 보조하기 위해 AI를 사용하는 것에 반대하지 않으며, 자신을 표현하기 위해 어떤 도구를 사용할지 자유롭게 선택해야 한다고 생각합니다. 플랫폼 내 일부는 인간 직접 작성을 최우선으로 하는 사람들이고, 다른 일부는 자신이 자부할 수 있는 작업을 만들기 위해 세부 사항을 고민하며 AI 도구 사용의 최전선을 공개적으로 탐구하고 있습니다. 우리 역시 서브스택에서 소프트웨어를 작성하고, 연구를 수행하고, 클리핑, 번역 등 제품 기능을 구축할 때 항상 AI를 사용합니다. 하지만 사람들은 자신이 무엇을 얻고 있는지 알아야 합니다."라고 말했습니다.

일링워스는 나에게 이렇게 말했습니다. "서브스택이 이 문제를 해결할 수 있는 가장 좋은 방법은 추가적인 대화를 이끌어내는 것이라고 생각합니다. 사실 저는 작가들이 AI를 활용해 뉴스레터를 어떻게 구성하는지 독자들에게 알릴 수 있는 새로운 기능이 정말 마음에 듭니다. 이것이야말로 우리가 잘못되었다는 것을 알고 있는 점수를 부여하는 대신, 서브스택이 장려해야 할 종류의 대화입니다. 제가 우려하는 것은 이러한 AI 탐지기가 모국어가 영어가 아닌 작가나 신경다양성(neurodiverse)을 가진 작가에게 허위 경고를 발생시킬 뿐만 아니라, 대화의 기회를 박탈한다는 것입니다."

원문 보기
원문 보기 (영어)
Substack contributors are rejecting the company’s decision to flag AI generated writing on the newsletter publishing platform, with several creators saying the new AI detection features are a “witch hunt.” Some Substack users who use AI to help them write or edit articles say the new policy unfairly discriminates against them, while other users who don’t use AI in their writing at all are worried that Substack will mistakenly flag their writing as being AI generated and tarnish their reputation. “I’m not going to apologize for using AI in the creation process. I wrote for 20 years without AI, I could do it again if I wanted to,” Mack Collier, who has a small Substack called Backstage Pass, wrote . “My output would fall, my posts would be less structured, and it would be more obvious that they needed a good editor. Using AI improves the overall quality of my writing. That’s why I use it.” “These detectors are notoriously, wildly inaccurate,” Alice Lemee, a ghostwriter and digital writing coach, said in a video she made about Substack’s new AI detection features. “All it takes is one false accusation for a writer to have their reputation almost irreversibly tarnished.” Substack announced the AI detection features on Tuesday in a post from CEO Christ Best , who said that “It’s getting harder to tell what’s real on the internet” and that “when content made by no one takes over parts of the internet that are supposed to be human, it pollutes the commons and makes it hard to discover and hear human voices.” Best explained that Substack has partnered with the AI detection tool Pangram, which is now built in to Substack and allows any user to scan a piece of writing. Pangram then produces a percentage-based score determining how much of the writing was AI or human generated. (Disclosure: Pangram previously bought an ad on 404 Media). We’ve covered research from Pangram, or research that relies on its AI detector before, showing how AI generated writing is flooding every corner of the internet. As we’ve previously reported, AI detectors are not perfect, and Pangram itself is not immune to labeling human content as being AI, and labeling AI as human. Max Spero, the CEO of Pangram, recently told us that the company is constantly working on minimizing errors, and that it estimates its false positive rate at roughly one in 10,000. “We have now built the mirror image. Pangram is a real machine, trained on vast amounts of text to make a probabilistic guess, and we have pointed it at your writing to work out whether a person is hidden inside,” Sam Illingworth, a professor and the author of Slow AI , said on his Substack in a post titled “Substack’s AI Detector and the Return of the Witch Hunt.” “To decide if there is a human on the other end, Substack asks a machine.” Substack also added a way for users to add a “How I make this” statement, where they can explain their writing process and disclose if they use AI. “We’re not against people using AI to assist their work, and we think people should be free to choose which tools they use to express themselves,” Best said. “Some on the platform are human-writing maximalists , others are publicly exploring the frontier of using AI tools while sweating the details to make work they stand behind. We use AI all the time at Substack to write software, do research, and build product features like clipping , translations , and more. But people should know what they’re getting.” “I think the best way for Substack to address this is to invite further dialogue,” Illingworth told me. “I actually really like the new feature that lets authors tell their readers how they construct their newsletters using AI, and this is the kind of dialogue they should encourage rather than ascribing a score that we know to be incorrect. My worry is that not only do these AI detectors generate false flags for non-native English speakers and neurodiverse writers, but they also remove any opportunity for dialogue because they start from a position of suspicion. Whereas what we need to be doing is developing opportunities for trust between readers and authors.” "The tools we've introduced today are intended to increase transparency and give readers more context, not to prohibit or penalize AI-assisted writing. They do not impact discovery on the platform,” a Substack spokesperson told me. “We're also encouraging Substack publishers to add a ‘How I make this’ statement, where they can explain their process directly, including how they use (or don't use) AI, and set expectations for readers. Creators can disable detection on their posts pre- and post-publication, as well as report and remove scans on their own work that they believe are mistaken. Find more information on how the features work here .” Substack’s attempt to detect and label AI generated content was also celebrated by many readers who are exhausted by AI slop and other internet platforms that can’t or don’t want to label it. I just published story about how there is so much unlabeled AI slop on Spotify that people are now creating their own sites to detect and label it without the company’s help. But the backlash from some Substack users shows that companies can’t just flip a switch if they want to ban or give users the ability to filter out AI content from their feeds. As the flood of AI writing and images we see online and the real world every day makes clear, generating slop is easy. What we’re going to do about it is still an unsolved problem . About the author Emanuel Maiberg is interested in little known communities and processes that shape technology, troublemakers, and petty beefs. Email him at emanuel@404media.co More from Emanuel Maiberg