메뉴
BL
TechCrunch AI 21일 전

프랑스 AI 스타트업 ZML, 다양한 칩에서 추론 속도를 높이는 무료 제품 출시

IMP
8/10
핵심 요약

튜링상 수상자 얀 르쿤의 지지를 받는 프랑스 AI 스타트업 ZML이 엔비디아, AMD, 구글 TPU 등 다양한 칩에서 최적의 LLM 추론 성능을 낼 수 있는 무료 소프트웨어 'ZML/LLMD'를 출시했습니다. 이는 기존 엔비디아 중심의 시장 독점과 벤더 종속 문제를 타파하고, 기업들이 비용과 에너지를 절감하며 AI 인프라를 구축할 수 있게 돕는 시장 판도 변경자로 평가받고 있습니다.

번역된 본문

엔비디아의 압도적인 시장 지배력이 끝난 것은 아니지만, 이제 곳곳에서 도전자와 대안들이 등장하고 있습니다. 튜링상 수상자인 얀 르쿤(Yann LeCun)의 지지를 받으며 떠오르는 프랑스 AI 스타트업 ZML은 엔비디아, AMD, 구글 TPU, 애플 메탈(Apple Metal), 인텔 아크(Intel Arc) 등 다양한 칩에서 여러 오픈소스 대규모 언어 모델(LLM)을 실행할 수 있게 해주는 추론(Inference) 성능 최적화 소프트웨어를 출시했습니다.

ZML의 창립자 스티브 모린(Steeve Morin)은 테크크런치(TechCrunch)와의 인터뷰에서 새롭게 출시된 LLM 추론 서버인 ZML/LLMD를 통해 기존의 사일로(독립적인 폐쇄망)를 깨부수고, 다양한 칩들이 AI 사용 사례에서 최대 성능으로, 어떨 때는 그 이상으로 작동할 수 있게 만드는 것이 회사의 목표라고 밝혔습니다. 모린은 AI가 우리의 업무와 일상에 통합됨에 따라 프롬프트 처리 과정인 '추론'을 최적화하는 것이 모델 학습보다 더 중요해졌지만, 소프트웨어와 아키텍처 장벽으로 인해 특정 업체에 종속되는 벤더 종속(Vendor lock-in) 문제가 발생하는 등 이면적인 부분은 여전히 파편화되어 있다고 설명했습니다.

다양한 칩에서 최고 성능을 끌어내겠다는 약속은 기술적인 성과일 뿐만 아니라, AI 관련 비용에 대한 우려가 커지고 있는 상황에서 시장의 판도를 바꿀 수 있는 혁신적인 요소입니다. ZML은 기업과 클라우드 고객들에게 비용이 저렴하거나 에너지 소비가 적은 여러 칩을 혼합하여 사용할 수 있는 옵션을 제공하고자 합니다. 모린은 "사람들에게 자신만의 시스템을 구축할 수 있는 힘을 돌려주고, AI가 널리 보급될 수 있도록 하는 실질적인 효율성을 달성하는 것이 우리의 목표"라고 말했습니다.

모린은 악셀레라(Axelera), 프랙타일(Fractile), 칼레이(Kalray), 올릭스(OLIX), 큐안트(Q.ANT), 스피펄(SiPearl), 스팅클라우드(SpiNNcloud), 브이소라(VSORA) 등 신생 AI 칩 제조사들(그중 다수가 유럽 기업임)에게도 이러한 소프트웨어적 지원이 큰 도움이 될 것이라고 덧붙였습니다. 하지만 그에게 중요한 것은 기업의 출신 지역이 아니라, ZML이 이들과 협력하여 "전 세계 어디에서도 시도되지 않았던 새로운 일"을 해낼 수 있다는 점입니다.

물론 모린이 엔비디아의 전망을 어둡게 보는 것은 아닙니다. 그는 엔비디아의 막강한 기존 공급망을 언급하며 엔비디아에 대해 비관적이지 않다고 밝혔습니다. 그는 ZML이 추론 시장의 부상에 대비하고 있는 거대 AI 칩 기업 엔비디아와 좋은 관계를 유지하고 있다고 전했습니다. 요즘 추론 분야는 막대한 투자가 쏟아지면서 이른바 '추론 골드러시(Inference gold rush)'로 불리고 있습니다. 이에 따라 최근 130억 달러(약 17조 원)의 기업가치를 받은 베이스텐(Baseten), 오픈소스 프로젝트 vLLM 개발자들이 만든 인페락트(Inferact), SGLang을 개발한 상업 기업 라딕스아크(RadixArk) 등이 경쟁자로 존재합니다. vLLM과 SGLang 모두 LLMD와 부분적으로 경쟁하지만, ZML의 모린이 품고 있는 야심은 더 큰 스펙트럼을 아우릅니다.

그는 "우리는 이제 실리콘을 공동 설계(co-designing)하는 단계에 도달했다"고 밝혔습니다. 또한 파리에 기반을 둔 이 20명 규모의 소규모 정예 팀 덕분에 민첩하게 움직일 수 있었으며, 앞으로도 더 많은 제품을 출시할 계획이라고 덧붙였습니다. 이 작은 팀이 빠르게 성장할 수 있었던 또 다른 이유는 이들의 규모에 비해 충분한 자금 지원을 받았기 때문입니다. 2017년 스냅챗(Snapchat)에 9자릿수 금액으로 인수된 젠리(Zenly)의 엔지니어링 부사장 출신인 모린은 20VC, >commit, AALVC, 드라이스데일 벤처스(Drysdale Ventures), 자비에르 닐(Xavier Niel)의 키마 벤처스(Kima Ventures), 킨드레드 캐피탈(Kindred Capital), 로컬글로브(LocalGlobe), 퍼즐 벤처스(Puzzle Ventures) 등의 벤처 캐피털로부터 2천만 달러(약 270억 원)를 투자받았습니다.

2024년에 출시되어 3월에 업데이트된 첫 공개 프로젝트(추론 중심 ML 프레임워크)와 달리, 이번 ZML/LLMD는 오픈소스가 아닙니다. 하지만 고객들의 실사용 데이터를 파악하기 위해 무료 제품으로 출시되었습니다. 모린은 "처음부터 너무 욕심을 부려 성장을 저해하는 것보다, 가장 효과적인 부분에서 성과를 측정하고 수익을 창출하는 것이 낫다"고 밝혔습니다. ZML/LLMD가 언제 유료 제품으로 전환되고 사용자들의 채택률은 어떨지 예측하기엔 아직 이르지만, 스타트업의 주주 명단에는 솔로몬 하이크스(Solomon Hykes, Docker 및 Dagger 창립자), 클레망 들랑게(Clément Delangue)와 줄리앙 쇼몽(Julien Chaumond, Hugging Face 창립자), 그리고 현재 AMI Labs에 소속된 얀 르쿤 등 다른 창립자들도 포함되어 있어 업계의 높은 관심을 보여줍니다. 이는 유럽의 AI 스타트업들 역시 자신들의 땅에서 훌륭한 성과를 낼 수 있다는 가능성을 증명합니다.

모린은 "나는 파리가 아니면 ZML을 할 수 없었을 것"이라고 덧붙였습니다.

원문 보기
원문 보기 (영어)
The days of Nvidia’s unparalleled market dominance aren’t over, but challengers and choices are arising from all directions. ZML , a hot French AI startup endorsed by Turing Award winner Yann LeCun, has released inference -performance software that allows a variety of open-source large language models to run on a variety of chips — including Nvidia’s, AMD’s, Google’s TPU, Apple Metal and Intel Arc. With ZML/LLMD , the newly launched LLM inference server, the company's ambition is to break existing silos and make different chips available for AI use cases at their maximum available speed, and sometimes faster, ZML founder Steeve Morin told TechCrunch. As AI becomes integrated into our work and everyday lives, optimizing inference — aka, the processing of prompts — has been outpacing model raining in importance, but often feels patchy behind the scenes, with software and architecture barriers that lead to vendor lock-in, Morin said. The promise of achieving peak performance across a variety of chips is a technological feat, but it could also be a market disruptor, amid mounting fears over AI-related costs. ZML hopes to provide enterprises and clouds with the option to use a mix of chips, some of which might be less costly or consume less energy. “The idea is to give people back the power to create their own system and achieve real efficiency gains that allow [AI] to be disseminated,” Morin said. Such a software assist may help novel AI chipmakers, many of which happen to be from Europe, Morin observed, citing Axelera , Fractile , Kalray , OLIX , Q.ANT , SiPearl , SpiNNcloud , and VSORA . But more than their region of origin, what matters to him is that ZML can work with them on “things that haven’t been done before anywhere in the world.” That doesn't mean Morin is bearish on Nvidia. He's not , in part because of its existing supply. He told TechCrunch that ZML has a good relationship with the AI chip giant, which has been gearing up for the rise of inference. Inference has been an area of such intense investment, that the trend has been hailed the “ inference gold rush ." So ZML has competition such as Baseten , recently valued at $13 billion; Inferact , from the creators of open source project vLLM ; as well as RadixArk , the commercial company behind SGLang . Both vLLM and SGLang partially compete with LLMD, but Morin's ambitions for ZML cover a broader spectrum. “We have reached the point where we are co-designing silicon,” he said. He further credited ZML’s lean team of 20 people as the reason why the Paris-based startup has been able to move fast, with more releases in the plans. It also helped that this small team is well funded for its size. Thanks to his track record as VP of engineering of Zenly , which Snapchat acquired for nine figures in 2017 , Morin raised $20 million from venture firms including Harry Stebbings’ 20VC, >commit, AALVC, Drysdale Ventures, Xavier Niel’s Kima Ventures, Kindred Capital, LocalGlobe, and Puzzle Ventures. Unlike ZML’s first public project, the inference-focused ML framework released in 2024 and updated in March , ZML/LLMD is not open source. But it is launching as a free product with the goal of learning about usage. “I’d rather measure and [then generate revenue] where it is most effective without hindering my growth stupidly because I have been too greedy from the get-go,” Morin said. It is too early to tell when ZML/LLMD might become a paid product, and what its adoption will look like. But the startup’s cap table confirms that other founders are paying attention, including Dagger and Docker founder Solomon Hykes, Clément Delangue and Julien Chaumond from Hugging Face, as well LeCun, now with AMI Labs . This also builds the case that Europe’s AI startups can now build from home . “I couldn’t do ZML anywhere but in Paris,” Morin said. Topics AI , AI inference , Europe , Exclusive , France , Fundraising , Startups , ZML When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Anna Heim Freelance Reporter Anna Heim is a writer and editorial consultant. You can contact or verify outreach from Anna by emailing annatechcrunch [at] gmail.com. As a freelance reporter at TechCrunch since 2021, she has covered a large range of startup-related topics including AI, fintech & insurtech, SaaS & pricing, and global venture capital trends. As of May 2025, her reporting for TechCrunch focuses on Europe’s most interesting startup stories. Anna has moderated panels and conducted onstage interviews at industry events of all sizes, including major tech conferences such as TechCrunch Disrupt, 4YFN, South Summit, TNW Conference, VivaTech, and many more. A former LATAM & Media Editor at The Next Web, startup founder and Sciences Po Paris alum, she’s fluent in multiple languages, including French, English, Spanish and Brazilian Portuguese. View Bio November 4 Boston Last chance to save up to $190 on TechCrunch Founder Summit. Join 1,000+ founders and VCs at all stages for real-world scaling insights and connections that move the needle. Savings end June 26, 11:59 p.m. PT . REGISTER NOW Most Popular Reddit is using LLMs to solve a problem LLMs largely created Amanda Silberling Amazon will stop accepting new customers for Mechanical Turk Anthony Ha 5 desk gadgets that can make your workday better Aisha Malik Chevy built an all-American EV truck — why is nobody buying it? Tim De Chant Mark Zuckerberg tells staff that AI agents haven't progressed as quickly as he'd hoped Lucas Ropek Jersey Mike's IPO illustrates how bad the AI hype has become Julie Bort After $18B IPO, Bending Spoons founder says success comes from minimizing luck Anna Heim