메타가 자사 AI 에이전트 '뮤즈(Muse)'의 신규 기능으로 홍보한 'AI가 대신 전화 걸기'가 실제로는 콜센터의 인간 상담원이 전화를 처리하는 방식으로 테스트되고 있는 것으로 드러났다. 404 미디어가 확보한 내부 문서에 따르면 메타는 내부적으로 '인간 에이전트 레이어'를 추가했다고 직원들에게 안내했으며, 직원들은 프라이버시 문제와 'AI가 부족해서 인간이 필요하다'는 부정적 인식을 우려하고 있다.
번역된 본문
지난주 메타 경영진은 크게 홍보해온 AI 에이전트 '뮤즈(Muse)'의 새로운 기능을 발표했다. 레스토랑 예약이나 미용실 예약 같은 일을 대신 해주기 위해 기업에 전화를 걸어줄 수 있다는 것이었다. 그러나 실제로는 메타가 이 전화들을 콜센터의 인간 상담원이 걸도록 테스트하고 있다는 사실을 404 미디어가 확인했다. 뮤즈 수석 엔지니어인 라이언 폭스(Ryan Fox)는 9월 16일 X(구 트위터)에 "미국 기업 대상 아웃바운드 콜에 대한 @muse 베타를 확장했다"고 게시했다. 메타의 최고 AI 책임자(CAIO) 알렉산드르 왕(Alexandr Wang)도 "뮤즈의 전화 베타를 확장한다!"고 게시했다. 그러나 내부적으로는 회사가 직원들에게 이 기능을 발표하며 "전화 완료를 위해 인간 에이전트 레이어를 추가했다"고 안내했고, "뮤즈 인간 에이전트 콜은 사내 도그푸딩(내부 테스트) 준비가 완료됐다"고 밝혔다. 404 미디어가 입수한 내부 게시판 글에는 "뮤즈는 단순히 번호를 다이얼하지 않는다. 사용자를 대신해 기업에 전화를 걸고, 대화를 처리하고, 요청을 완수한 뒤, 녹취록과 요약과 함께 결과를 보고한다"고 적혀 있다. "또한 더 이상 혼자 일하지 않는다. 뮤즈는 이제 요청을 훈련된 인간 에이전트에게 넘길 수 있으며, 인간 에이전트가 전화를 걸어 일을 처리한다." 이 게시물은 "이것은 여전히 기밀 유지 중인 출시 전 제품"이라며 "회사 외부의 누구와도 이 제품이나 그 결과물을 공유하지 말라"고 덧붙였다. 또다시, 주요 AI 기술의 핵심 기능이 사실상 콜센터에서 일하는 사람들이 회사가 AI가 한다고 말한 작업을 수행하는 것에 불과한 셈이다. 언제, 얼마나 자주 전화가 이른바 '인간 에이전트'로 연결되는지, 그리고 언제 전화가 순수하게 AI에 의해서만 처리되는지는 명확하지 않다. 전화에 인간이 개입한다는 사실은 직원들의 의문을 불러일으켰다. 사용자의 전화 요청을 다른 인간에게 공유하는 것의 프라이버시 영향과, 자사 AI 어시스턴트가 스스로 전화를 처리할 만큼 advanced하지 않다는 인식이 그것이다. 한 직원은 메타 내부 게시판에 "내가 뭘 잘못 이해하고 있는지 모르겠지만, 왜 이게 '기능'인가?"라고 썼다. "부정적 PR이 될 잠재력이 너무 크다. '그들의 AI가 충분히 좋지 않아서 여전히 인간이 필요하다'는 식의 보도가 나올 수 있다." 메타가 적어도 때때로 인간이 사용자의 채팅을 읽고 사용자를 대신해 전화를 걸도록 한다는 것은, 기업이 AI 제품을 출시하며 AI가 한다고 주장하지만 나중에 인간이 그 작업의 일부 또는 전부를 하고 있었다는 사실이 밝혀지는, 유서 깊은 전통의 연속이다. 그리고 이번 사례는 기술 거대기업이 잠재적으로 매우 민감한 정보를 인간에게 넘기는 최신 사례이기도 하다. 누군가 뮤즈에게 민감한 병원 예약을 부탁했는데, 사용자는 그 전화를 인간이 걸 수도 있다는 사실을 모르는 상황을 상상하기 어렵지 않다. 404 미디어가 확인한 내부 커뮤니케이션에서 메타는 사용자가 AI와 공유한 잠재적으로 민감한 정보가 인간 계약자에게 전달되더라도, 계약자들이 '모든 데이터를 안전하고 보호하기 위한 충분한 훈련'을 받았기 때문에 여전히 안전하다고 직원들에게 주장했다. 한 직원은 그런 훈련은 '보안 메커니즘이 아니다'라고 지적했다. "기본 활성화 상태로 공개 출시하면 엄청난 반발이 일어날 것이다. 얼리 어답터들로부터 얻고 있는 호감과 자연스러운 언론 보도를 모두 죽일 것"이라며 "'메타가 무대 뒤에서 인간을 사용해 전화를 건다' 또는 '뮤즈 AI는 사실 외주 계약자'라는 헤드라인이 어떻게 외칠지 생각해보라. 데이터 처리의 프라이버시와 보안에 대한 엄청난 악성 보도로 이어질 것이다. […] 부디 부디 이 상태 그대로 출시하지 말고, 굳이 출시해야 한다면 최소한 사용자 설정 옵션으로 만들어 달라. 제발." 시스템을 테스트한 다른 직원들은 이것이 "정말정말 나쁜 아이디어"라고 말했다. 그들은 전화가 걸린 이후에야 인간이 전화를 대신 걸었다는 사실을 통보받았다. 한 직원은 "테스터는 발신자가 인간이라는 사실을 미리 알리지 않았고, 전화가 끝난 후에야 통보받았다"고 썼다. "그래서 그들은 자신의 정보가 어떻게 활용되는지에 대해 우려하고 있"
Last week, Meta executives announced that its much-hyped AI agent, Muse, had a new feature: It could call businesses for you to do things like make a restaurant reservation or a haircut appointment. But in reality, Meta is testing having these calls being made by human beings in call centers, 404 Media has learned. “We just expanded the @muse beta for outbound calls to US businesses,” Ryan Fox, the principal engineer on Muse posted on X on Sept. 16. Alexandr Wang , Meta’s chief AI officer, posted “we’re expanding our phone beta for muse!” Internally, however, the company announced the features to employees and said that it had “added a human agent layer for calls to get completed,” and that “Muse human agent calls [sic] is ready for company dogfooding,” or testing. “Muse doesn’t just dial a number. It calls a business on your behalf, handles the conversation, completes your request, and reports back with a transcript and a summary,” the internal message board post to employees, seen by 404 Media, reads. “It also no longer works alone. Muse is now able to hand requests to a trained human agent, who places the call and works it through.” “This is still a confidential, pre-launch product,” the post adds. “Please don’t share it, or anything it producers, with anyone outside the company.” Once again, a core feature of a major piece of AI technology will actually just be people working in a call center, doing work that a company says AI is doing. It is not clear how often or when a call is routed to a so-called “human agent,” and when a call is done exclusively by AI. The fact that human beings are involved in the call raised questions among employees, who wondered both about the privacy implications of sharing people’s call requests with other human beings as well as the perception that its AI assistant wasn’t advanced enough to do calls on its own. “Not sure what I am missing here but why is this a ‘feature?,’ one employee wrote on Meta’s internal message board for workers. “This has potential for so much negative PR. It could portray us as ‘their AI is not good enough so they still need humans’ kind of coverage for this launch.” That Meta is at least sometimes letting human beings read users’ chats and make phone calls on their behalf is part of a time honored tradition in which companies launch an AI product, claim that it is AI, only for it to be later learned that human beings are doing some or all of the work. And it is the latest example of a tech giant passing potentially highly sensitive information to human beings; it is easy to imagine someone asking Muse to make a sensitive doctor’s appointment, for example, and the user not knowing that a human being might make that call. In internal communications seen by 404 Media, Meta suggested to employees that potentially sensitive information shared by the user with AI — and then passed to human contractors — was still safe because the contractors had undergone “a lot of training to make sure all your data is safe and secure.” An employee commented that the training “is not a [security] mechanism.” “This is absolutely going to create a ton of outcry if we publicly launch this as default-on. It will kill all the goodwill and organic press we’re getting from early adopters,” they wrote. “Please think about how the headlines will scream ‘Meta uses humans behind the scenes to make calls’ or ‘Muse AI is actually independent contractors’ and result in a ton of bad press about the privacy and security of our data handling […] please PLEASE do not release this as-is and at least make this a user preference if we absolutely must launch this. Please.” Other employees who had tested the system said that it was a “bad bad idea,” and that they were not made aware that a human being did the call for them until after the call had been made: The tester “was not made aware of the caller being human and was only told after,” one employee wrote. “So they were concerned with information that they shared with the assumption that it is the secure AI which is making the call. Even with clear indication of human calling, I can think of many scenarios where I wouldn’t be comfortable with this and I am not positive about how our users will react to this.” Various Meta employees and regular users have posted about Muse’s ability to make phone calls, and their experience with it. Ravid Shwartz Ziv, who does AI research at Meta, posted on X that “this week, Muse called customer service on my behalf. It navigated the phone tree, waited on hold, spoke with a human rep, and resolved the issue (worked great!). The crazy thing for me (besides that it works) is that the rep didn’t blink. Talking to a bot was completely natural to them.” Others have posted that the call function didn’t work, or that the person on the other end hung up on the agent (some have said that Muse did what it was supposed to do). “Internal testing, aka dogfooding, is core to the product development process. While the response from employees has been overwhelmingly positive, the entire point is to get feedback so we can implement safety and privacy protections and improve features before we release them publicly,” a Meta spokesperson told 404 Media. “We’re working with merchants to continue improving this potential calling feature, and will only roll it out when it's ready and with the proper disclosures.” About the author Jason is a cofounder of 404 Media. He was previously the editor-in-chief of Motherboard. He loves the Freedom of Information Act and surfing. More from Jason Koebler