메뉴
BL
TechCrunch AI 9일 전

누구나 무료로 쓰는 'AI의 월드 와이드 웹' 구축에 나선 비영리 단체

IMP
8/10
핵심 요약

주요 AI 시스템이 빅테크 기업의 독점으로 남아있는 가운데, 비영리 단체 'Current AI'는 프랑스 정부 및 주요 기업들로부터 4억 달러의 자금을 조달하여 공공 AI 인프라를 구축하고 있습니다. 이들은 전 세계 소수 언어와 지역 사회의 데이터를 보호하고, 언어 장벽 없이 누구나 무료로 접근할 수 있는 개방형 AI 생태계를 만드는 데 초점을 맞추고 있습니다.

번역된 본문

인도 시골의 한 농부가 시들어가는 식물의 사진을 찍었다. 그녀는 인터넷으로 이에 대해 알아보고 싶지만 영어를 하지 못한다. 하지만 그녀가 굳이 영어를 할 필요는 없다. 이것이 바로 'Current AI'라는 비영리 단체가 개방적이고 공공적인 AI 인프라를 구축함으로써 해결하고자 하는 문제다. 이 단체는 2월 인도 AI 정상회의에서 인도 정부의 AI 언어 부서인 '바시니(Bhashini)'와 협력했다. 그 결과 탄생한 것이 '수노 수트라(Suno Sutra)'로, 힌디어로 '듣는 연대기'라는 뜻이다. 이는 인터넷 연결 없이 22개의 인도 언어로 AI를 구동하는 주머니 크기의 오프라인 기기다.

Current AI의 아야 비데르(Ayah Bdeir) 최고경영자(CEO)는 테크크런치와의 인터뷰에서 "인도에는 수백 개의 다양한 언어와 방언이 있지만, 현재 AI는 이를 제대로 반영하지 못하고 있다"고 말했다. 이 기기는 오픈소스로 제공되어 개발자 커뮤니티가 이를 기반으로 다양하게 발전시킬 수 있다. 마르티 티스네(Martin Tisne)가 2025년 2월에 설립한 이 비영리 단체는 빠른 속도로 움직이고 있다. 지난달에는 4개 조직의 프로젝트에 320만 달러의 지원금을 배정했으며, 최근인 지난주에는 스위스 제네바에서 열린 'AI for Good Summit'에서 오픈소스 AI 챗봇을 공개했다. 모질라(Mozilla)의 AI 전략을 이끌었던 비데르 CEO는 올해 1월에 합류했다. 그녀는 2019년 스페로(Sphero)에 매각되어 수백만 명의 어린이에게 도달했던 STEM 교육 기업인 '리틀비츠(littleBits)'를 설립한 바 있다. 비데르는 테크크런치에 Current AI가 정부, 기업, 자선단체를 한데 모아 공익 기술에 자금을 지원하는 '공공-민간 파트너십' 형태로 운영된다고 밝혔다. 프랑스 정부가 1억 달러의 시드 머니를 지원했으며, 포드 재단, 맥아더 재단, 딥마인드(DeepMind), 세일즈포스(Salesforce)가 동참해 현재까지 약정된 총 자금 규모는 4억 달러에 달한다. 비데르는 "그들은 투자자가 아니라 자금 지원자(funders)"라고 강조했다.

이 단체가 해결하고자 하는 문제는 명확하다. OpenAI부터 구글, 앤스로픽(Anthropic)까지 오늘날의 모든 주요 AI 시스템은 사기업의 소유다. 비데르는 "AI가 진정으로 혁신적인 기술이며 삶의 모든 측면을 바꿀 것이라면, '월드 와이드 웹(World Wide Web)'처럼 누구나 무료로 사용할 수 있는 공공의 대안이 반드시 있어야 한다"고 말했다. 전 세계 절반의 구어가 멸종 위기에 처해 있다. 비데르는 "영어가 가장 큰 언어 모델과 AI 시스템을 주도함에 따라, 세계의 수많은 언어, 그리고 그에 따른 문화와 공동체가 뒤처지고 있다"고 덧붙였다. 빅테크 기업들의 다국어 지원 움직임에 대한 질문에 그녀는 확고한 선을 그었다. "빅테크는 동의나 문맥과 상관없이 단지 시장을 확장하기 위해 다국어 모델을 구축한다"고 지적했다. 그 결과는 매우 구체적인 문제로 나타난다. "원주민 언어의 경우, 해당 공동체가 어떠한 규칙도 정하기 전에 선교사들의 성경 번역본이 AI의 학습 데이터로 먼저 사용되곤 한다."

단순히 언어에 관한 문제가 아니다. AI가 언어를 구사하는 능력은 학습해야 할 요소의 일부일 뿐이다. "언어는 지식, 전통, 기억, 정체성이 한 세대에서 다음 세대로 전해지는 방식입니다. 따라서 기술이 당신의 언어를 말할 수 없다면, 그것은 당신의 문화 또한 담아낼 수 없습니다." Current AI에 대한 그녀의 비전은 초기 웹을 모델로 한 개방형 시스템이다. 이 시스템에서는 개선의 이점이 모두에게 돌아가고, 누구도 소외되지 않으며, 커뮤니티가 자체 데이터에 대한 통제권을 유지하게 된다. 지난달 발표된 Current AI의 첫 지원 그룹에는 케냐, 레바논, 브라질 아마존에 걸친 4개 조직에 320만 달러를 배포하는 것이 포함되었다. 케냐의 마사카네(Masakhane) 프로젝트는 건강, 농업, 교육을 위해 50개 이상의 아프리카 언어로 된 AI 데이터셋을 구축하는 것이다. 레바논의 세계 창조 연구소(Institute for Worldmaking)는 (기술 회사가 아닌) 지역 사회가 통제할 수 있는 기계 판독 가능 데이터베이스에 아랍 문화 역사와 현대적 실천을 디지털화하고 있다. 브라질의 'Portal sem Porteiras'는 원주민 아마존 공동체와 함께 데이터를 영토 내에 유지하면서 오프라인 AI 도구를 구축하고 있다. 그리고 케냐의 아프리카 인터넷 권리 연합(African Internet Rights Alliance)은 아프리카 전역의 AI 시스템에 책임을 묵기 위한 감사 도구를 개발 중이다.

데이터의 소유권은 누구에게 있는가? 데이터 소유권 문제에 대해 비데르는 단호하게 말했다. "다양한 커뮤니티 내에서 데이터 소유권이 누구에게 있는지에 대해 여러 가지 모델과 제안이 있지만, 한 가지는 확실합니다. 그것은 (데이터가) 기업의 것이어서는 안 된다는 것입니다."

원문 보기
원문 보기 (영어)
A farmer in rural India takes a photo of a dying plant. She wants to research it on the internet but she doesn't speak English. She shouldn't have to. That's the type of problem a nonprofit called Current AI is trying to solve by building open, public AI infrastructure. In February at the India AI Summit, it teamed up with Bhashini , the Indian government's AI language division. The result became Suno Sutra, Hindi for "listening chronicles," a pocket-sized, offline device that runs AI in 22 Indian languages, no internet required. "In India, there are hundreds of different languages and dialects, and right now AI is not representing them," Current AI CEO Ayah Bdeir said in an interview with TechCrunch. The device is open-sourced, available for developer communities to build on. The nonprofit, founded in February 2025 by Martin Tisne, is moving fast. Last month, it allocated $3.2 million in grants to projects across four organizations; most recently (last week) it launched an open-source AI chatbot at the AI for Good Summit in Geneva . Bdeir, joined in January after leading Mozilla's AI strategy. She previously founded littleBits, the STEM education company that reached millions of kids before selling to Sphero in 2019 . Current AI operates as a "public-private partnership" bringing together governments, companies, and philanthropies to fund public interest tech, she told TechCrunch. The French government seeded Current AI with $100 million, joined by the Ford Foundation, MacArthur Foundation, DeepMind, and Salesforce — bringing total committed funding to $400 million. "They're not investors; they're funders," Bdeir said . The problem it aims to solve is straightforward: every major AI system today, from OpenAI to Google to Anthropic, belongs to a private company. "If AI is truly a transformative technology, if it's going to change every aspect of everyone's life, there has to be a public alternative," Bdeir said. "Like the World Wide Web, available to anyone, for free." Half the world’s spoken languages face extinction . “And with English driving the largest language models and AI systems, a bulk of the world’s languages and, consequently, cultures and communities are left behind,” Bdeir said. When asked about Big Tech's multilingual push, Bdeir drew a sharp distinction. "Big tech builds multilingual models to expand their market," she said, "regardless of consent or context." The consequences are concrete. "For Indigenous languages, missionary Bible translations become training data before communities have set any rules," she said. Not just about language An AI's ability to speak a language is only part of what it needs to learn. “Language is how knowledge, tradition, memory and identity get carried from one generation to the next. So when a technology can't speak your language, it can't hold your culture either,” she said. Her vision for Current AI is an open system modeled on the early web, where improvements benefit everyone, no one gets locked out, and communities keep control of their own data. Current’s first cohort grant round, announced last month, involved deploying $3.2 million to four organizations across Kenya, Lebanon, and the Brazilian Amazon. The project in Masakhane, Kenya, involves building AI datasets across more than 50 African languages for health, farming, and education; Lebanon's Institute for Worldmaking is digitizing Arab cultural history and contemporary practice into machine-readable databases that communities (not tech companies) control. Brazil's Portal sem Porteiras is building offline AI tools with Indigenous Amazon communities, keeping data within the territory. And Kenya's African Internet Rights Alliance is developing audit tools to hold AI systems accountable across the continent. Who owns the data? On the question of data ownership, Bdeir didn't mince words. "There are different models and proposals for who owns data in various communities, but one thing is sure: it shouldn't be a company in Silicon Valley trying to make a select few thousand people wealthier," she told TechCrunch. The nonprofit's approach is to store models and data locally, bringing in community experts before anything is built, or writing consent protocols into the pipeline so communities can halt the process at any point. None of Current AI's grantees have fully solved it yet. But Bdeir sees that as the point. "Every one of them has built the question into their work," she said, "rather than accepting the usual default, where complexity becomes the excuse to let a government or a tech company decide for everyone." As for how much progress can be made with a $3.2 million budget split across four organizations, Bdeir says, “Scale is not always the measure. That is the Big Tech paradigm,” she said. "This could look like an Indigenous elder in the Brazilian Amazon using a tool built in Kenya to be able to pass down ecological knowledge in their own language.” Building the stack Earlier this month in Geneva, Current launched Alpha Chat , an open-source chatbot assembled in seven weeks by a coalition of ten organizations, including Hugging Face, Mozilla, and MIT Media Lab. Each contributor brought a piece of the stack, including a language model, safety tooling, and computing power. Current AI also struck a deal with Sakana AI , a Tokyo-based startup known for its work on what it calls Sovereign AI. The two organizations plan to build a shared open-source AI stack, one designed to support the Japanese language and culture, but also communities across the Global South that dominant AI systems have largely ignored. Topics AI , current ai , large language models , nonprofit When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Kate Park Reporter, Asia Kate Park is a reporter at TechCrunch, with a focus on technology, startups and venture capital in Asia. She previously was a financial journalist at Mergermarket covering M&A, private equity and venture capital. View Bio November 4 Boston Last chance to save up to $190 on TechCrunch Founder Summit. Join 1,000+ founders and VCs at all stages for real-world scaling insights and connections that move the needle. Savings end June 26, 11:59 p.m. PT . REGISTER NOW Most Popular Tesla driver in fatal Texas crash pressed accelerator 100%, NTSB confirms Sean O'Kane Amid hardware legal battle, OpenAI releases a $230 keyboard for Codex Lucas Ropek OpenAI's first hardware device is reportedly a screenless speaker that can move Lucas Ropek Anthropic's newest ad is creeping people out Lucas Ropek Satya Nadella has issued a shocking warning to companies using AI Julie Bort The wildest allegations in Apple’s trade secrets lawsuit against OpenAI Sarah Perez Anthropic starts localizing Claude pricing for India, its biggest market after the US Jagmeet Singh