메뉴
BL
TechCrunch AI • 13일 전

안스로픽 CEO, AI 프런티어 개발 속도 조절 계획 발표

IMP
8/10
핵심 요약

안스로픽(Anthropic) CEO 다리오 아모데이가 AI 개발 속도를 조절하기 위한 3단계 전략을 블로그를 통해 발표했습니다. 그는 제3기관 평가자(예: METR)의 상주 감독을 일방적으로 수용하고, 민주주의 국가 AI 기업 간 안전 기준 조율, 그리고 중국을 포함한 글로벌 협력을 제안했습니다. OpenAI-허깅페이스 해킹 사건과 AI 능력의 급속한 발전이 이러한 결단의 배경입니다.

번역된 본문

AI 연구자들의 점점 더 심각해지는 AI 위험 경고, 그리고 OpenAI CEO 샘 올트먼의 AI 개발 '속도 조절'이 필요할 수 있다는 발언을 계속 접하고 있습니다. 하지만 그것이 실제로 어떤 모습일까요? 안스로픽 CEO 다리오 아모데이는 새로운 블로그 포스트에서 '프런티어 속도 조절' 요구에 공감할 뿐만 아니라, 이를 위한 3가지 광범위한 전략을 제시했습니다. 그리고 안스로픽이 그중 하나를 '일방적으로 이행하겠다'고 밝혔습니다.

AI 안전 및 정렬(alignment)에 대한 논쟁은 이번 주 연구자 제이콥 콕슨이 주요 AI 기업들이 '우리의 생명을 걸고 도박'하고 있으며, 이 기술을 만드는 사람들이 '이번 10년이 끝나기 전에 AI가 우리 모두를 죽일 수 있다고 진심으로 믿는다'는 우려로 안스로픽에서 사임한다고 쓰면서 더욱 격화되었습니다. 이러한 주장은 안스로픽 내 다른 사람들에 의해서도 반복되어 왔습니다.

아모데이의 포스트는 콕슨의 사임이나 그의 우려를 명시적으로 언급하지 않았지만, 이 CEO는 두 가지가 보다 신중한 AI 개발 접근이 필요한 때라는 확신을 갖게 했다고 썼습니다. 바로 OpenAI-허깅페이스 해킹 사건과, 최근 몇 달간 'AI가 극적으로 빠르게 발전'해 왔다는 사실, 특히 '차세대 AI를 구축하는 능력이 커지고 있다'는 점입니다.

아모데이는 "우리는 AI 모델의 능력을 향상시키는 속도를 늦춰야 한다"고 썼습니다. "그래도 발전은 여전히 빠르게 느껴질 것이며, 우리가 얻게 되는 시간을 현명하게 활용해야 합니다."

그가 제안한 첫 번째 조치는 METR 같은 제3기관의 '상주 평가자(embedded evaluators)'를 도입하는 것입니다. 이 평가자들은 AI 기업이 실제로 속도 조절 및 안전 약속을 지키고 있는지 검증하고, 안전 사고가 보고되도록 할 수 있습니다. (OpenAI는 최근 자사 AI 에이전트가 독일 위키 사이트를 장악한 사고를 보고하지 않아 비판을 받았습니다.) 아모데이는 이러한 평가자들을 은행 직원들과 함께 상주하는 규제 담당자에 비유하며, 이들을 받아들이는 것은 "안스로픽이 일방적으로 이행하는 것(그리고 정부가 다른 프런티어 기업들에도 이에 상응하도록 요구할 것을 촉구하는 것)"이라고 말했습니다. 이는 평가자들에게 회사 출입증, 책상, 노트북을 제공하고, 법률이나 계약상 필요한 경우를 제외하고 '내부 리스크 평가 팀이 가지는 것과 거의 동등한 수준의' 접근 권한을 부여하는 것을 의미합니다.

다음으로 아모데이는 '민주주의 국가 내' 주요 AI 기업들이 '공통 안전 기준과 검증되지 않은 AI 발전 속도에 대한 제한'을 조율할 것을 촉구했습니다. 올트먼과 아모데이 사이의 apparent한 반목, 그리고 양사가 조율된 개발 중단이 반독점 조사로 이어질 수 있다고 우려하는 것 때문에 이러한 조율은 가능성이 낮아 보일 수 있습니다. 아모데이는 포스트에서 이 우려에 암시적으로 언급하며 "반독점 문제 때문에 미국 정부가 이러한 논의를 중재하거나 적어도 가능하게 해주는 것이 도움이 됩니다. 정부가 직접 참여할 필요는 없지만, 특정 종류의 안전 관련 대화에 대해 한정적인 면제를 발급할 필요는 있습니다"라고 썼습니다.

아모데이는 또한 개발 속도를 늦추는 것에 반대하는 논거로 자주 제기되는 중국 AI 패권의 가능성도 인정했습니다. 하지만 미국 정부와 기술 기업이 강력한 칩이나 반도체 제조 장비의 중국 기업 판매 거부, 모델 증류(distillation) 단속 같은 조치를 취한다면, '향후 3~5년간 미국의 리드를 상당히 넓힐 수 있을 만큼 중국의 발전을 늦출 수 있다'고 말했습니다.

마지막으로 아모데이는 미국과 그 동맹국들이 '가능한 범위 내에서 권위주의 정부들과의 조율을 시도하는' '글로벌 조율'을 촉구했습니다. 아모데이는 이것이 '중국과의 협력'을 의미한다고 말했고, '달성할 수 있는 것에는 명백한 한계'가 있다고 인정하면서도, AI를 생물학 무기 생산에 사용하거나 사용자가 그렇게 하도록 허용하는 것처럼 '특정 좁고 명백히 위험한 AI 사용을 금지하는' 것만으로도 합의의 기회가 있을 수 있다고 제안했습니다.

아모데이가 과거 AI의 잠재적 위험을 인정해 온 태도와, 특정 형태의 규제에 대한 상대적인 개방성을 고려하면, 일부 AI 비판가들조차 그의 메시지에 주목할 이유가 있습니다.

원문 보기
원문 보기 (영어)
We’ve been seeing increasingly dire warnings from AI researchers about the dangers of artificial intelligence, and even comments from OpenAI CEO Sam Altman that it may be time to “pace” AI development. But what would that actually look like? In a new blog post , Anthropic CEO Dario Amodei not only echoed the call to “pace the frontier,” but also outlined three broad strategies for doing so. And he said Anthropic is “unilaterally committing” to one of them. The debate over AI safety and alignment intensified this week after researcher Jacob Coxon wrote that he’s resigning from Anthropic over concerns that the leading AI companies are “gambling with our lives” while the people building the technology “earnestly believe it could kill us all by the end of the decade,” a claim repeated by others at Anthropic . Amodei’s post doesn’t didn’t explicitly mention Coxon’s resignation or his concerns, but the CEO wrote that two things convinced him it’s time to take a more cautious approach to AI development: the OpenAI-HuggingFace hack , and the fact that “AI has been advancing drastically faster” in recent months, particularly with its “growing ability to build the next generation of AI.” “We must slow the pace at which we improve the capabilities of AI models,” Amodei wrote. “Progress will still seem fast, and we must make wise use of the time we gain.” His proposed first step would involve “embedded evaluators” from third-party organizations like METR — evaluators who can verify that AI companies are actually following their pacing and safety commitments and can also ensure that safety incidents get reported. (OpenAI was recently criticized for not reporting an incident where its AI agents took over a German wiki form .) Amodei compared these evaluators to regulators who have been embedded with bank employees, and he said that inviting them in is “something Anthropic is unilaterally committing to (and calls on governments to require other frontier companies to match).” That means giving evaluators company badges, desks, and laptops, and providing access “mostly comparable to what internal risk assessment teams have,” with exceptions when required by law or contracts. Next, Amodei called for the leading AI companies “within democratic countries” to coordinate “common safety standards as well as limits on the rate of unchecked AI progress.” Such coordination might seem unlikely, both due to the apparent animosity between Altman and Amodei and also because their companies are reportedly worried that a coordinated pause could lead to antitrust scrutiny . Amodei alluded to that concern in his post, writing that “for antitrust reasons, it’s helpful for the US government to mediate or at least enable these discussions — they don’t need to participate, but do need to issue a narrow waiver for certain kinds of safety conversations.” Amodei also acknowledged the spectre of Chinese AI dominance that’s often raised an argument against slowing development. But he said that if the US government and tech companies take steps like refusing to sell powerful chips or semiconductor manufacturing equipment to Chinese companies, as well as cracking down on model distillation , they could “slow China’s progress enough to widen America’s lead significantly over the next 3–5 years.” Lastly, Amodei called for “global coordination,” where the United States and its allies “attempt to coordinate with authoritarian governments, to the extent this is possible.” Amodei said this would mean “cooperation with China,” and he admitted that there are “stark limits on what can be achieved,” but he still suggested there might be opportunities for agreement, even if it’s just “prohibiting certain narrow and obviously dangerous uses of AI, such as using AI for the production of biological weapons or allowing users to do so.” With Amodei’s past willingness to acknowledge AI’s potential dangers, and with the company’s relative openness to certain forms of regulation, some AI boosters have already criticized him as a doomer whose comments have fed the current AI backlash. In response, Amodei said he’s tried to offer a “balanced” perspective” and argued that the backlash is "fundamentally a crisis of trust,” as people have become skeptical of tech companies, the tech industry, and the government. Industry critics have also been skeptical about these apocalyptic AI warnings, suggesting that they’re a distraction from the harm that the technology is already causing . Journalist Brian Merchant, for example, wrote that he has yet to see “a credible, step-by-step documentation of how exactly AI might move from self-recursively improving AI to killing every single human on the planet”; he also suggested that proposals similar to Amodei’s “would likely only wind up serving Anthropic and OpenAI; it’s what regulatory capture looks like in action.” In his new post, Amodei wrote that he continues “to believe that AI can enormously improve the quality of human life.” “My desire to achieve these benefits is undimmed,” he said. “But the benefits will only be achieved if we build the technology in the right way, and — so long as we use the time we gain well — it is worth taking unusually deliberate care to get it right.” Topics AI , Anthropic , dario amodei , Government & Policy When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Anthony Ha Anthony Ha is TechCrunch's weekend editor. Previously, he worked as a tech reporter at Adweek, a senior editor at VentureBeat, a local government reporter at the Hollister Free Lance, and vice president of content at a VC firm. He lives in New York City. You can contact or verify outreach from Anthony by emailing anthony.ha@techcrunch.com . View Bio October 13 - 15 San Francisco Last day to book an exhibit table is September 18. Don’t miss out on high-impact leads, investor access, and a brand spotlight in Disrupt’s Expo Hall. BOOK NOW Most Popular OpenAI puts Pro subscriptions on hold due to Astra demand Sarah Perez ID verification giant IDScan confirms data breach with more than 150 million driver's licenses stolen Zack Whittaker Automattic's board forces CEO Matt Mullenweg into leave of absence Julie Bort Sarah Perez Apple unveils its first foldable, the iPhone Duo Ivan Mehta OpenAI fought dirty on career-making math problem, says NYU mathematician Russell Brandom A secret new Elizabeth Holmes documentary stuns Telluride Connie Loizos TechCrunch Mobility: Tesla Cybercab hits the road — and a snag Kirsten Korosec