메뉴
BL
404 Media • 3일 전

OpenAI, AI로 AI를 학습시킨 계약직 검수자들 해고

IMP
6/10
핵심 요약

OpenAI의 모델 개선을 위해 고용된 수천 명의 계약직 데이터 검수자들이 업무에 AI를 사용하다 적발되어 해고된 사실이 404 Media의 내부 문서 확보로 드러났습니다. 이는 AI가 생성한 텍스트로 학습할수록 모델 성능이 저하되는 '모델 붕괴(model collapse)'를 막기 위해 인간의 검수가 필수적인 상황에서 벌어진 일이라 아이러니합니다. 검수자들은 반복적 단어 사용, em dash 과다 사용 등 AI 사용의 단서를 파악하도록 지시받고 있으며, 내부 문서에는 AI 탐지 도구나 Grammarly 같은 AI 도구 사용 자체도 금지되어 있습니다.

번역된 본문

OpenAI에는 실제 ChatGPT 사용자들의 프롬프트와 다른 데이터를 읽고 챗봇의 응답을 개선하는 데 도움을 주는 대군의 계약직 인력이 있다. 이들의 역할은 OpenAI의 모델에 당연히 '인간의 손길'을 더하는 것이다. 하지만 모든 계약직이 그렇게 하고 있는 것은 아니다. 404 Media는 OpenAI의 모델 개선을 위해 고용된 여러 계약직이 AI를 사용해 AI를 학습시키다 해고된 사실을 확인했다. 이는 모델 자체에도 좋지 않지만, AI 학습 회사들이 OpenAI를 위해 일하면서 직원들이 AI를 사용했다는 이유로 해고하는 것은 — OpenAI의 존재 이유가 직장에서 사람들이 AI를 사용하게 만드는 것임을 고려하면 — 큰 아이러니가 아닐 수 없다.

일부 AI 모델은 이미 '모델 붕괴(model collapse)'의 징후를 보이고 있는데, 이는 AI가 생성한 텍스트로 추가 학습된 AI 모델이 점점 더 나빠질 수 있는 현상이다. 이번 사례에서는 이러한 현상을 부분적으로 막기 위해 고용된 사람들 중 일부가 스스로 AI 생성 응답을 사용해 OpenAI의 모델을 학습시키고 있었던 것이다.

한 계약직은 "AI를 사용하는 사람을 정말 자주 보고, 그 때문에 해고되는 사람도 정말 자주 있다. 사실상 그게 즉시 퇴출당하는 유일한 이유"라고 말했다. 이 사람은 "수천 명 규모의 그룹에서 수많은 사람이 적발됐다"고 덧붙였다.

💡 OpenAI, Anthropic 또는 다른 AI 회사에서 프롬프트 검수자로 일하고 계신가요? 제보를 기다립니다. 업무용이 아닌 기기에서 Signal(joseph.404)로 안전하게 연락하거나 joseph@404media.co로 이메일을 보내주세요.

지난주 404 Media는 '프로젝트 릴리(Project Lily)'를 보도했다. 이 프로젝트에서 OpenAI는 수백 명의 계약직이 개인정보를 포함할 수 있는 실제 ChatGPT 사용자들의 프롬프트와 대화를 읽게 한다. 이 계약직들은 ChatGPT가 생성한 응답에 대해 평가하고 비평하며, 응답이 지나치게 아부하지 않는지, ChatGPT를 의인화하지 않는지 확인하는 역할도 한다. 이 보도는 내부 문서와 해당 프롬프트 업무에 종사하는 관계자와의 대화를 바탕으로 했다.

404 Media는 이후 다른 내부 문서를 확보하고 다양한 프로젝트에서 OpenAI를 위해 일하는 세 명의 계약직과 대화를 나눴다. 내부 문서 중 하나에 따르면 이러한 프로젝트에는 1만 명이 넘는 계약직이 포함될 수 있다. 문서 중 하나는 계약직이 업무에 스스로 AI를 사용해서는 안 된다고 명시하고 있다.

다른 계약직들의 업무를 검토하는 계약직(여기에는 AI 사용 적발도 포함)의 업무를 설명하는 한 문서에는 이렇게 적혀 있다. "AI 탐지 도구를 사용하지 마시고, 직접 AI도 사용하지 마세요. GPTZero나 다른 AI 탐지 도구를 사용하지 마세요. 신뢰할 수 없습니다. 검토자는 Grammarly나 AI 번역을 포함하여 검토, 피드백 작성, 코멘트 작성에 AI를 사용할 수 없습니다."

문서는 이어서 다음과 같이 말한다. "평가자에게 AI 사용이 의심되는 이유를 알려주지 마세요. 무엇을 찾는지 알면 숨기기가 더 쉬워집니다. 하나의 단서가 아니라 전체적인 패턴으로 판단하세요."

세 명의 계약직 모두 검토자들에게 업무에 AI를 사용하지 말라는 지시를 받았다고 말했다. 두 명의 소식통은 AI 사용으로 인해 사람들이 해고되거나 퇴출(offboard)되었다고 밝혔다. 404 Media는 이들이 언론과 대화할 수 없는 처지였기에 익명성을 보장했다.

다른 계약직의 업무를 검토하는 계약직들은 AI 사용의 징후를 경계하도록 지시받는다. 여기에는 반복적인 단어 사용, AI 스타일의 문장 부호 — 예를 들어 em dash(—)의 과도한 사용 — 그리고 계약직이 업무를 매우 빠르게 끝내는 것 등이 포함될 수 있다.

한 계약직은 사람들이 서로 조언을 구하는 관련 Slack 채널에서 많은 사람이 예시를 올리며 '이거 AI인가요?'라고 묻는다고 말했다. 이 사람은 "대개 대답은 '그렇다'"라고 했다.

한 계약직은 OpenAI 모델 학습을 도우면서 AI를 사용했으며, 해고 통보서로 보이는 것을 공유했다. 거기에는 고용주가 그들의 업무 '진정성(authenticity)'에 문제를 확인했다고 적혀 있었다.

이 계약직은 404 Media에 "저는 나쁜 사람이나 나쁜 직원이 아닙니다. 그저 조금의 도움이 필요했고 AI에 의지했고, 결국 그것이 몰락으로 이어졌죠"라고 말했다. "저는 이 일에서 아무런 기쁨도 느끼지 못했고, 사회에 기여하고 있다는 느낌도 없었습니다."

404 Media와 대화한 계약직 중 두 명은 계약직을 고용하는 AI 학습 회사인 Mercor에서 일하고 있었다.

원문 보기
원문 보기 (영어)
OpenAI has an army of contractors who read real ChatGPT users’ prompts and other data to help improve the chatbot’s responses. The idea is that the contractors provide an, obviously, human touch to OpenAI’s models. Well, not all of the contractors are doing that. 404 Media has found multiple contractors hired to improve OpenAI’s models have been fired for using AI to train the AI. That’s not great for the models themselves, but there is also obviously a great irony in AI training companies working for OpenAI firing people for using AI when OpenAI’s whole thing is to make people use AI at work. Some AI models already exhibit signs of “model collapse,” which is where AI models further trained on AI-generated text can become worse and worse. In this case, some of the people hired to partially stop that happening are themselves using AI-generated responses to train OpenAI’s models. One contractor said they see people using AI “all the time and people are let go for it all the time, it’s pretty much the one thing that will get you kicked off ASAP.” The person said, “in a group of thousands there are tons that have been caught.” 💡 Do you work as a prompt reviewer for OpenAI, Anthropic, or another AI company? I would love to hear from you. Using a non-work device, you can message me securely on Signal at joseph.404 or send me an email at joseph@404media.co. Last week, 404 Media revealed Project Lily, in which OpenAI has hundreds of contractors reading real ChatGPT users’ prompts and conversations which can include personal information. Those contractors then rate and critique the responses ChatGPT generated, including making sure that the responses are not too sycophantic or anthropomorphize ChatGPT. That reporting was based on internal documents and conversations with a person who works on the prompts. 404 Media has now obtained other internal documents and spoken to three contractors doing work for OpenAI across various projects. Those projects can include more than ten thousand contractors, according to one of the internal documents. One of those documents says contractors must not use AI themselves for their work. “Do not use AI detection tools, or AI yourself,” one document describing the work of contractors who are hired to review the work of other contractors, including catching them for using AI, says. “Do not use GPTZero or any other AI detection tool. They are not reliable. Reviewers may not use AI either, including Grammarly and AI translation, to review, write feedback, or write comments.” The document continues, “Do not tell evaluators why you suspect AI. It is easier for them to hide if they know what you look for. Judge the overall pattern, not one clue.” All three of the contractors said reviewers are told not to use AI in their work. Two of the sources said people have been fired or offboarded for using AI. 404 Media granted the contractors anonymity as they weren’t permitted to speak to the press. The contractors who review other contractors’ work are told to be on the look out for tell-tale signs of AI use. That can include repetitive words, AI-style punctuation — which might include over zealous use of the em dash — and contractors finishing their work very quickly. In related Slack channels where people ask each other for advice, a lot of people will post an example with the question, ‘Is this AI?,’ one contractor said. “Usually the answer is yes,” the person said. One contractor said they used AI while helping to train OpenAI’s models and shared what they presented as their termination letter. It said their employer had identified issues with the “authenticity” of their work. “I’m not a bad person or worker. I just needed a little boost and turned to AI to help me which eventually led to my downfall,” the contractor told 404 Media. “I felt no joy in the work or that I was contributing to society in any way.” Two of the contractors 404 Media spoke to worked for Mercor, an AI-training company that hires the contractors who in turn review ChatGPT-related material. A Mercor spokesperson told 404 Media in a statement: “Our experts are hired for their expertise and judgement, which is essential to the ongoing advancement of AI. Our contracts strictly prohibit the use of LLMs to complete projects and we enforce that. We invest heavily in our tools and systems to detect misuse and ensure our experts comply with project rules and contract terms. When we confirm an expert has used AI to complete a task, we immediately remove them from the project.” 404 Media spoke to a fourth contractor who has worked on training models for various AI companies. They said they sometimes purposefully chose the worst responses because they wanted to actively sabotage the models’ training. “I did feel guilty about doing this kind of work at the start,” they said. “I either pay zero attention to the results and choose randomly or purposely choose the [worst] output. I’m not sure how much of a difference it actually makes since there are hundreds of other people also rating prompt results, but it does feel like I’m getting paid to make AI worse.” OpenAI declined to comment on its contractors being fired for using AI. About the author Joseph is an award-winning investigative journalist focused on generating impact. His work has triggered hundreds of millions of dollars worth of fines, shut down tech companies, and much more. More from Joseph Cox
관련 소식