메뉴
BL
TechCrunch AI • 17일 전

NYU 수학자, 밀레니엄 문제를 둘러싼 OpenAI의 불공정 행위 주장

IMP
7/10
핵심 요약

NYU의 트리스탄 벅마스터(Tristan Buckmaster) 교수는 앤스로픽(Anthropic) 소속 수학자 레벤트 알퍼지(Levent Alpöge)와 협업해 나비에-스토크스 문제 관련 세 가지 증명을 발표했습니다. 그런데 연구 진행 정보가 OpenAI에 새어나간 뒤 OpenAI가 막대한 컴퓨팅 자원을 동원해 같은 접근법으로 경쟁 증명을 시도했다는 논란이 불거졌습니다. OpenAI의 세바스티앙 부벡(Sébastian Bubeck)은 이 주장을 부인했으나, 밀레니엄 상금 문제를 둘러싼 AI 활용 연구의 윤리와 공정성 문제가 핵심 쟁점이 되었습니다.

번역된 본문

뉴욕대학교(NYU) 수학과 교수 트리스탄 벅마스터(Tristan Buckmaster)는 화요일, 이론 수학의 주요 미해결 문제 중 하나에 대한 예비적 결과를 포함해 세 가지 증명을 발표했다. 이 성과는 앤스로픽(Anthropic) 소속 수학자 레벤트 알퍼지(Levent Alpöge)와의 협업으로 이루어졌으며, OpenAI의 코덱스(Codex)와 앤스로픽의 클로드(Claude) AI 모델을 모두 활용한 것이다. 이 결과 자체로도 의미가 크지만, 동시에 같은 문제를 풀려던 OpenAI의 시도를 둘러싼 이례적인 논란이 함께 나타났다. 벅마스터는 증명 발표 성명에서 "이 이야기에는 또 다른 부분이 있으며, 솔직히 말해 신경 쓰고 싶지 않았던 부분"이라고 밝혔다.

벅마스터와 알퍼지가 결과를 마무리하던 중, 두 사람은 "우리의 진행 상황에 대한 정보가 OpenAI에 전달되었다"는 사실을 알게 되었다. 이에 대해 OpenAI에 문의하자, OpenAI는 이미 핵심 문제에 대한 완전한 증명을 달성했다는 답변을 받았다. 하지만 OpenAI가 언제 이 문제 연구를 시작했는지, 사람의 개입이 얼마나 있었는지 등 후속 질문을 하자 답변은 점점 회피적으로 변했다. 벅마스터는 "문제 전체를 전담하는 팀이 있었고, 막대한 양의 컴퓨팅 자원이 사용된 것이 드러났다"며 "결국 첫 프롬프트는 우리 연구 정보가 OpenAI에 도달한 이후 지난 며칠 사이에 보내진 것으로 합의되었다"고 말했다. 사실이라면, OpenAI 팀이 벅마스터와 알퍼지의 접근법이 옳다고 확신하고, 컴퓨팅 자원에서의 물질적 우위를 활용해 형식적 증명을 먼저 완성하려 했다는 것을 시사한다.

OpenAI의 수학 연구를 이끄는 세바스티앙 부벡(Sébastian Bubeck)은 이러한 주장이 "거짓이고 선동적"이라고 반박했다. 부벡은 벅마스터의 성명 이후 올린 글에서 "분명히 하자면, 나는 학계 규범에 따라 이 논의에 임했고, 이런 상황까지 온 것이 유감"이라며 "나를 아는 사람이라면 학문적 기준이 나에게 얼마나 중요한지 알 것"이라고 썼다. 부벡은 곧 더 완전한 성명을 내겠다고 덧붙였다.

이 분쟁의 중심에는 '나비에-스토크스 존재性与 매끄러움(existence and smoothness)' 문제가 있다. 이는 7개의 '밀레니엄 상 문제(Millennium Prize problems)' 중 하나로, 클레이 수학연구소(Clay Mathematics Institute)가 각 문제의 해결자에게 100만 달러의 상금을 걸어둔 주요 미해결 수학 문제들이다. 나비에-스토크스 방정식은 유체역학에서 널리 사용되지만 이론적으로는 잘 이해되지 않는다. 이 문제가 해결된다면 수리물리학에 대한 인류의 집단적 이해가 크게 진전될 것이다.

이 문제는 수학자들 사이에서 널리 연구되지만, 벅마스터와 그의 공동 연구자가 택한 구체적 전략은 훨씬 드문 것이었다. 그렇기에 벅마스터는 OpenAI가 마침 같은 시기에 같은 접근법을 택한 것이 수상하다고 판단했다. 벅마스터는 "매끄러운 힘을 통한 클레이 문제 접근, 즉 페퍼만(Fefferman)의 문제 서술에서 선택지 c와 d는 루이스와 디에고가 열어놓은 길이자 레벤트와 내가 조용히 공략하기로 선택한 길이었다"고 썼다. 그는 "내가 아는 한 거의 다른 누구도 이 방향을 연구하고 있지 않았다"며 "모델에게 문제 서술을 주고 며칠 만에 도달할 수 있는 방향이 아니다"라고 덧붙였다.

알퍼지는 앤스로픽 소속이지만 이 연구를 회사를 대신해 수행한 것은 아니었다. 그래서 두 사람은 주로 OpenAI의 코덱스에 의존하며 여러 모델을 섞어 사용했다. 그럼에도 앤스로픽 소속이라는 점이 OpenAI의 심기를 건드린 것으로 보이며, 벅마스터는 부벡이 제안한 타협안의 일부로 알퍼지의 공로 표기를 빼달라고 요청했다고 주장했다. 벅마스터가 이 분쟁을 공개하겠다고 밀자 부벡은 "왜 당신 커리어를 망치려 하냐"고 답했다고 한다. 벅마스터가 반발하자 부벡은 "내가 친절할 필요가 없다면, 친절하지 않아도 된다"고 말했다고 한다.

벅마스터는 또한 프로젝트 수행에 코덱스를 광범위하게 사용했기 때문에 자신의 연구 정보가 OpenAI의 문제 해결 노력에 영향을 줬을 수 있다는 우려도 제기했다. OpenAI는 코덱스 상호작용을 모델 학습에 사용할 권리를 보유하고 있다.

원문 보기
원문 보기 (영어)
NYU mathematics professor Tristan Buckmaster announced three proofs on Tuesday with a preliminary finding on one of the major unsolved problems in theoretical mathematics. The findings, made in collaboration with Anthropic mathematician Levent Alpöge and using both Codex and Claude AI models, are significant in themselves — but they’re also accompanied by an unusual controversy surrounding OpenAI’s attempts to solve the same problem. “There is another part of this story,” Buckmaster wrote in his statement announcing the proofs, “and one that, honestly, I very much wish I did not have to be concerned with.” While Buckmaster and Alpöge were finalizing their results, they learned that “information about our progress had been passed to OpenAI.” When they contacted OpenAI about this, they were told that OpenAI had already achieved a full proof of the central problem. But when they asked follow-up questions about when OpenAI had begun its research into the problem and how much human input was involved, the answers became more evasive. “It emerged that an entire team had been working on the problem,” Buckmaster said, "and that an insane amount of compute had been used…. Eventually, it was agreed that [the first prompt] had been sent in the past few days, after information about our work had reached OpenAI." If true, that would suggest the OpenAI team had become convinced that Buckmaster and Alpöge’s approach was the right one, and decided to use its material advantage in computing resources to reach a formal proof first. Sébastian Bubeck, who leads OpenAI’s mathematical research, says those claims are “false and inflammatory.” “To clarify, I came into the discussion following academic norms, and I'm disappointed that it has come to this,” he wrote in a post after Buckmaster’s statement. “Anyone who knows me knows that academic standards are of the highest importance to me.” Bubeck said he would make a more complete statement soon. The dispute centers on the “Navier-Stokes existence and smoothness” problem, one of the seven " Millennium Prize problems" — a set of major unsolved math problems, each carrying a $1 million bounty from Clay Mathematics Institute for the first person or group to provide a solution. The Navier-Stokes equations are widely used in fluid mechanics but poorly understood in theoretical terms. A solution would represent a significant advance in the collective understanding of mathematical physics. Although the problem is widely pursued among mathematicians, the specific tactic taken by Buckmaster and his collaborator is far less common. As a result, Buckmaster found it suspicious that OpenAI ended up taking the same approach at the same time. “The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack," Buckmaster wrote. "Almost nobody else I know of was working on it,” he continued. “It is not the direction one arrives at in a few days by giving a model the problem statement.” While Alpöge is employed by Anthropic, he was not conducting this research on the company's behalf. As a result, the duo used a mix of models, relying primarily on OpenAI's Codex in their work. Even so, Alpöge's affiliation with a rival lab seems to have been a sore point for OpenAI, and Buckmaster alleges that Bubeck asked him to remove Alpöge's credit as part of a proposed compromise. When Buckmaster pushed to make the dispute public, he says that Bubeck replied: "Why would you ruin your career?" Buckmaster says that when he pushed back, Bubeck followed up with: "If you don't want me to be nice, then I don't have to be nice." Buckmaster also raised concerns that, because he used Codex extensively in assembling the project, information from his work could have informed OpenAI’s own efforts to solve the problem. OpenAI reserves the right to train models on Codex interactions, although users are able to opt-out. If the OpenAI team used a model trained on Buckmaster’s own Codex interactions, it’s plausible that it could have regurgitated his work when faced with a similar problem. OpenAI did not respond to a request for comment on this possibility. Regardless, the issue is likely to reignite the ongoing debate about AI’s role in mathematical research, and OpenAI’s specific incentives. For his part, Buckmaster seems to believe the best answer is to get as much information about the research out into the public eye. “I have not seen OpenAI’s proof," Buckmaster wrote. "I do not know what their model did, or how. I do not know whether our data was used. I am not accusing anyone of anything. I am stating what I was told, when, and what was proposed to me. I am stating it because the alternative is to let a sequence of announcements say something I know to be false.” Topics AI , mathematics When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Russell Brandom AI Editor Russell Brandom has been covering the tech industry since 2012, with a focus on platform policy and emerging technologies. He previously worked at The Verge and Rest of World, and has written for Wired, The Awl and MIT's Technology Review. He can be reached at russell.brandom@techcrunch.com or on Signal at 412-401-5489. View Bio October 13 - 15 San Francisco Don't miss out . The startup community will gather to answer a pivotal question: How do you build sustainably in the AI era? REGISTER NOW Most Popular A secret new Elizabeth Holmes documentary stuns Telluride Connie Loizos TechCrunch Mobility: Tesla Cybercab hits the road — and a snag Kirsten Korosec Hikers rescued after using Google Gemini for planning Anthony Ha Feds launch investigation into Tesla's Cybercab deployment Sean O'Kane Kirsten Korosec Tesla is asking people if they want to buy and run Cybercab fleets Kirsten Korosec Norway considers ban on camera-enabled wearable ‘pervert glasses' Zack Whittaker Uber is laying off 10% of staff, or 3,300 people Ram Iyer
관련 소식