메뉴
HN
Hacker News 5일 전

오픈소스 AI를 반대하는 주장은 빈약하다

IMP
8/10
핵심 요약

최근 일부 정치권과 빅테크 기업들은 오픈소스 AI가 위험하다며 통제를 주장하지만, 이는 시장 독점을 위한 변명에 불과합니다. 소프트웨어 산업의 역사를 볼 때, 오픈소스는 상용 소프트웨어의 근간이며 억압하려는 시도는 오히려 역효과를 낳습니다. 특히 AI 모델의 개방은 특정 국가의 현상이 아니라 엔비디아 같은 글로벌 기업들의 이해관계가 맞물린 필연적인 흐름입니다.

번역된 본문

최근 'Kimi K3' 모델이 출시되면서新一轮新一輪新一轮新一轮新一轮 새로운 불안과 혼란스러운 논쟁이 촉발되었습니다. 언론인, 비즈니스 리더, 정치인들을 중심으로 오픈소스 AI가 위험한 위협이라고 주장하는 목소리가 매우 거셉니다.

OpenAI 관계자인 딘 볼(Dean Ball)은 다음과 같이 말했습니다: "오픈 가중치 모델(Open-weight-model)이 지배하는 세계의 한 가지 가능한 결과는 완전한 AI 공산주의입니다... 시장 상품이 아니라 AI가 '공공재'가 되는 것이죠." 누구나 무료로 사용할 수 있는 AI라니, 얼마나 끔찍한가요!

최첨단 AI 연구소(Frontier labs)들이 오픈소스 AI를 반대하는 논리는 본질적으로 다음과 같습니다. 오픈소스 모델은 위험(하고 비미국적!)합니다. 우리가 AI라는 판도라의 상자를 열어야 하지만, 반드시 '책임감 있는' 관리자(가급적이면 우리가 독점하는 통행료 징수원!)를 통해서만 열어야 합니다. 오직 신뢰할 수 있는 사용자(우리의 가장 수익성 높은 고객들)만이 이를 사용할 수 있어야 합니다.

저는 오픈소스 AI를 반대하는 몇 가지 잘못된 주장을 반박하고자 합니다. 하지만 그 전에, 이 논쟁이 프레이밍되는 방식에 대해 몇 가지 짚고 넘어갈 점들이 있습니다.

오픈소스 소프트웨어는 상용 소프트웨어의 기초입니다

볼의 주장은 오픈소스 소프트웨어가 모든 독점 소프트웨어의 기초라는 사실을 간과하고 있습니다. 여기에는 결국 소프트웨어 제품에 불과한 최첨단 AI 모델들도 포함됩니다. 소프트웨어 산업 밖의 사람들에게 오픈소스 소프트웨어는 직관적이지 않습니다. 왜 열심히 제품을 만들어 놓고 무료로 나눠주는 것일까요?

하나의 소프트웨어 프로그램은 여러 프로그램들이 쌓여 있는 형태이며, 각 층은 그 아래 층 위에 구축됩니다. 우버(Uber) 같은 서비스를 만들려면 프로그래밍 언어 프레임워크, 웹 트래픽을 주고받는 소프트웨어, 데이터 분석 도구 등 수많은 구성 요소가 필요합니다. 이러한 대부분의 요소들은 상업 기업에게 차별점이 되지 않으므로, 기업들은 하위 구성 요소에 대해서는 협력하고 진정으로 자신들의 제품을 차별화하는 상위 수준에서 경쟁하는 것이 이익입니다.

최첨단 AI 연구소들은 AI 모델이 '경쟁하는 것이 의미 없을 정도로 흔해지는' 범주에 속하지 않기를 매우 바랍니다. 그렇게 될지 여부는 두고 봐야 알 일입니다.

오픈소스 소프트웨어를 억제하는 것은 매우 어렵습니다

현실에서 오픈소스 모델을 억제하자는 논쟁은 대부분 핵심을 벗어나 있습니다. 역사는 오픈소스 소프트웨어를 억제하는 것이 매우 어려우며, 그렇게 하려는 시도는 기업들을 국제 경쟁자들에게 취약하게 만들 뿐이라고 말해줍니다.

암호화 기술의 짧은 역사가 이를 잘 보여줍니다. 오늘날 PGP는 누구나 사용할 수 있는 흔한 도구이며, 대부분의 개발자가 익숙하게 알고 있습니다. 하지만 1991년 필 짐머만(Phil Zimmermann)이 이를 발명했을 때, 미국 정부는 암호화 기술을 군사 기술로 간주했습니다. 짐머만에 대한 형사 조사가 시작되었습니다. 넷스케이프(Netscape)가 SSL을 만들었을 때, 미국 정부는 국제적으로 약화된 버전만 출시하도록 허용했습니다.

이러한 통제는 역효과를 낳았습니다. 약화된 '국제' 버전을 구하는 것이 훨씬 쉬웠기 때문에 많은 미국인들조차 그것을 사용했습니다. 수출 통제는 정부가 원했던 대로 암호화를 제한하는 데 실패했습니다. SSL, PGP 및 유사한 도구들은 전 세계 어디서나 쉽게 구할 수 있었고, 통제는 오히려 미국인들에게 불이익을 주었습니다. 결국 법원은 암호화 소스 코드를 배포하는 것은 표현의 자유로 보호된다고 판결했고, 미국 정부는 암호화 수출 통제를 완화했습니다.

억제 대상을 '중국산' 모델로 좁힌다고 해서 상황이 쉬워지는 것은 아닙니다. 도대체 무엇이 AI 모델을 중국산으로 만드는가요? 최첨단 모델 주장처럼 미국 모델에서 증류(distill)되었다면 중국산인가요? 미국인이 중국 모델을 미세조정(fine-tune)하면 어떻게 되나요? 기껏해야 이런 식으로 AI를 규제하는 것은 (일시적으로) 미국인들에게 붉은 테이프(까다로운 규제)를 두르고 세계 다른 지역에 비해 AI 접근성을 떨어뜨리게 될 것입니다.

오픈소스 AI는 단지 중국의 현상이 아닙니다

오픈소스 AI 논쟁에는 오픈소스 모델이 오직 중국 정부만이 개발할 인센티브가 있는 것이라는 가정이 깔려 있습니다. 하지만 현실에는 오픈소스 AI를 개발할 충분한 동기를 가진 수많은 상업적 주체들이 있습니다.

칩메이커: 엔비디아(Nvidia)의 젠슨 황(Jensen Huang) CEO는 엔비디아가 구축하고 있는 것을 '토큰 공장(token factories)'이라고 묘사했습니다. 엔비디아는 자사의 칩이 최첨단 모델을 실행하는 데 사용되는지 여부에는 신경 쓰지 않습니다.

원문 보기
원문 보기 (영어)
The release of Kimi K3 has opened a fresh round of angst and confused discourse. There's a loud cohort of journalists, business leaders, and politicians arguing that open source AI is a dangerous threat. OpenAI's Dean Ball : One probable outcome of an open-weight-model-dominant world is full AI communism... rather than a market product, AI is a "public good" Freely available AI for anyone? The horror! Frontier labs' case against open source AI is essentially: Open source models 1 are dangerous ( and un-American! ). We should open the AI Pandora's Box, but only with responsible gatekeepers ( toll collectors, preferably us! ). Only trusted users ( our most profitable customers ) should be able to use it. I want to address some bad arguments against open source AI, but some corrections on how the argument is being framed are in order: Open source software is the foundation for commercial software ​ Ball's framing strolls past the fact that open source software is the foundation of all proprietary software. This includes frontier models, which at the end of the day are software products. Open source software is counterintuitive to people outside of the software industry. Why work hard on a product, and give it away for free? A software program is a stack of programs, with each layer built on top of another. To build Uber, you need programming language frameworks, software to send and receive web traffic, data analysis tools, and countless other components. Most of these are not differentiators for a commercial enterprise, so it serves commercial actors to cooperate on lower components in the stack and compete on the higher level pieces that actually differentiate their products. Frontier labs would very much like AI models to not fall into the category of "so commonplace that it doesn't make sense to compete on". Whether that happens remains to be seen. Open source software is very difficult to suppress ​ In reality, the argument about suppressing open source models is mostly beside the point. History tells us that suppression of open source software is extremely difficult, and attempting to do so only serves to weaken companies against international competitors. A brief history of encryption is illustrative: Today, PGP is a commonplace tool anyone can use, and most devs are at least familiar with. But when Phil Zimmermann invented it in 1991, the U.S. government considered encryption to be military technology . A criminal investigation was opened against Zimmermann. When Netscape created SSL , the U.S. government allowed it to only release a weakened version of it internationally. These controls backfired: it was much easier to acquire the weakened, "international" version, so even many Americans used it . Export controls did not succeed in limiting encryption as the government wished. SSL, PGP, and similar tools were readily available throughout the world, and the controls disadvantaged Americans. Eventually, courts ruled that releasing encryption source code is protected speech, and the U.S. government relaxed encryption export controls . Narrowing suppression to "Chinese" models won't make things easier. What, exactly, makes an AI model Chinese? Is it Chinese if, as frontier models allege, it was distilled from American models? What about if an American fine-tunes a Chinese model? At best, regulating AI in this way will (temporarily) encumber Americans with red tape and diminished AI access relative to the rest of the world. Open source AI is not just a Chinese phenomenon ​ There's an assumption baked into the open source AI debate that open source models are something that only the Chinese government has an incentive to develop. In reality there are many commercial actors with ample incentive to develop open source AI: Chip makers : Nvidia CEO Jensen Huang has described what Nvidia is building as “token factories” 2 . Nvidia doesn't care if its chips are used to run frontier models or cheap open source models 3 - it just wants to produce and generate demand for as many tokens as possible. And indeed Nvidia has itself released a suite of open source models . American Startups : Thinking Machines Lab recently released a powerful open source model . They and others are betting that models will be commoditized, and a defensible moat can be built around auxiliary services that complement or customize models. Enterprise AI users : Frontier model customers aren't currently all that active in open source AI development, but they will be. They will want lower-cost models for low-complexity tasks, and more fine-grained control over customer-facing features. BigCos : You can be sure that Google and Meta are watching OpenAI's new ad product closely. Should frontier model ad products gain traction, it would be well worth it for these behemoths to commoditize ad-free, open source models to squash ad competition. The "AI race" is... what, exactly? ​ Much of the angst around China's models centers on "losing the AI race". But what's the goal of this race? Is it to develop the best model? To sell the most tokens? To destroy humanity first? Talking about an "AI Race" doesn't make more sense than talking about an "Internet Race". We're not competing to be the first to send a rocket to the moon, we're reacting to a new, transformational technology. To the extent there's a race between nations, it's to absorb this transition and grow economies. In this framing, free AI models are a boon, not a threat. Bad arguments to fear Chinese AI models ​ China is "AI dumping!" ​ Scott Galloway has argued that free Chinese AI is an attempt to eliminate competitors in the long run: This is what China did to solar panels, steel, EVs, and batteries. First, they match Western quality, or they don't even match it. 89%. Close. Actually, match it with cars, they've matched it, but go ahead. Then they cut the price by two thirds, then they own the market. But apart from chips, AI isn't a physical good. Solar panels and steel require physical supply chains, each link of which cannot easily exist on its own. If no one is manufacturing solar panels in your country, it's difficult to build a business selling solar-grade silicon wafers. Software isn't like that. An open source model coming from China doesn't prevent a fine-tuning business from succeeding in the US - quite the opposite! They will spread propaganda! ​ It's not unreasonable to assume that Chinese models will be shipped with a pro-China point of view. But this is not a reason to suppress them. The models are open source! If any American has an issue with the political slant of Chinese AI models, they are free to change and release an "Americanized" one. At least within the U.S., it's difficult to foresee a model seen as having a distorted pro-China bias outcompeting a substantially similar model with a distorted pro-U.S. bias. They will add backdoors! ​ AI does not change the basic market for vulnerabilities: responsible actors patch them, attackers exploit them. Limiting tools for responsible actors only serves attackers. It's theoretically possible for a bad actor to embed hidden adversarial behavior in a model. But if this happens, it serves the interests of responsible actors to find these exploits as soon as possible, and the best way to do this is to let anyone who wants to inspect them. Open source AI is coming ​ It doesn't matter much what policy makers or business leaders want: open source AI is too powerful, and too difficult to control. It's coming, and attempts to squash it will not amount to anything more than noise along the way. Footnotes ​ I'll use the terms "open source model" as in, "open weights model". ↩ Quote taken from Derek Thompson's recent article on Chinese AI. Which, while we're here, gets a few things wrong: whoever is on the frontier is the best