AI 스타트업 Decart가 실시간으로 실사 수준의 주행 환경을 생성하는 인터랙티브 월드 모델 'Oasis 3'를 API 형태로 전격 공개했습니다. 자율주행 및 로봇 공학을 주요 타겟으로 삼고 있으며, 처음부터 API를 개방하여 개발자 생태계를 구축하는 데 사활을 걸고 있습니다. 뛰어난 하드웨어 최적화 기술을 바탕으로 경쟁사 대비 압도적으로 저렴한 비용으로 무한한 환경을 생성할 수 있다는 것이 이 모델의 가장 큰 강점입니다.
번역된 본문
AI 스타트업 Decart가 실시간으로 실사 수준의 사진 같은 주행 환경을 생성할 수 있는 최신 인터랙티브 월드 모델(World Model)인 Oasis 3를 수요일에 공개했다고 TechCrunch가 단독으로 보도했습니다. 이 모델은 현재 API를 통해 사용할 수 있습니다. 이 스타트업은 초기 대규모로 희귀한 주행 시나리오를 시뮬레이션해야 하는 자율주행 차량 기업을 타겟으로 삼고 있으며, 향후 로봇 공학 및 기타 물리적 AI(Physical AI) 애플리케이션으로 확장할 계획입니다.
하지만 더 큰 목표는 개발자들입니다. 첫날부터 API 액세스를 제공함으로써, Decart는 OpenAI가 대형 언어 모델로 달성한 것과 유사한 방식으로 월드 모델을 중심으로 한 개발자 생태계를 구축하려 하고 있습니다. Decart의 공동 창립자이자 CEO인 딘 라이터스도르프(Dean Leitersdorf)는 TechCrunch와의 인터뷰에서 "이것은 사람들이 실제로 프로그래밍할 수 있는 최초의 유용한 월드 모델이 될 것"이라며 "이를 기반으로 하는 전체 개발자 커뮤니티가 탄생할 것이라고 생각한다"고 말했습니다. 이 스타트업은 이미 10만 명 이상의 개발자 커뮤니티를 보유하고 있으며, 이들 중 다수는 주로 전자상거래와 라이브 스트리밍 분야에서 실시간 비디오 모델인 '루시(Lucy)'를 기반으로 제품을 구축하고 있습니다. Oasis 3는 해당 파운데이션 모델을 기반으로 하며, 회사의 물리적 AI 진출을 상징하는 모델입니다. 이용료는 초당 0.02달러로 책정되었으며, 기업용 가격은 사용 사례에 따라 다르다고 Decart는 밝혔습니다.
Decart는 점점 더 치열해지는 월드 모델 경쟁 시장에 뛰어들고 있습니다. 작년에 구글은 연구 프리뷰 버전으로 Genie 3를 공개했고, 페이페이 리(Fei-Fei Li)의 World Labs는 상업적 사용 사례를 위해 Marble을 출시했습니다. 또한 Luma 및 Runway와 같은 비디오 생성 스타트업들도 물리 법칙을 이해하는 비디오 모델을 월드 모델로 변환하고 있습니다.
Oasis 3의 출시는 창립 2년 차인 Decart가 3억 달러(약 4,000억 원)를 유치한 지 몇 주 만에 이루어졌습니다. 라이터스도르프는 이 자금 조달이 전자상거래, 라이브 스트리밍 및 물리적 AI 분야에서 "우리가 구축한 모델에 대한 수요가 급증한 데" 따른 것이라고 말했습니다. 이번 펀딩 라운드를 통해 Decart의 기업가치는 약 40억 달러(약 5조 3천억 원)로 급등했으며, 도요타, 어도비, 이베이와 같은 전략적 투자자들이 참여했습니다. 라이터스도르프는 이 기업들이 모두 잠재 고객이라고 밝혔습니다. 기존 투자자인 엔비디아(Nvidia) 역시 이번 라운드에 참여했습니다.
Oasis 3의 강점은 모델의 실사 수준의 사진 같은 구현과 무한 생성 기능에 있습니다. 이는 회사의 또 다른 주요 제품인 DOS(Decart Optimization Stack) 소프트웨어를 통해 구동되는 효율성 기술 덕분입니다. 이 소프트웨어는 엔비디아, 아마존, 구글 하드웨어에서 모델이 효율적으로 실행되도록 하여, 타사 모델에 비해 운영 비용을 훨씬 저렴하게 만듭니다. 라이터스도르프는 "이것은 하드웨어 수준까지 완벽하게 최적화한 전체 실시간 스택 위에 구축되었습니다"라며 "이렇게 수직적으로 통합함으로써 이러한 모델을 실행하는 데 있어 업계 어느 곳보다 10배 이상 저렴할 수 있습니다"라고 설명했습니다. 그에 따르면 이 스타트업의 모델은 매우 효율적이어서 창립 이후 지금까지 1억 달러보다 '현저히 적은' 비용만 소진했습니다.
Oasis 3는 훈련 및 테스트 시스템을 위해 물리적으로 정확한 다중 카메라 환경(전면 1대, 측면 2대)을 생성합니다. 또한 제한된 데모나 연구 프리뷰만 제공하는 대신, 개발자가 시나리오를 무한히 생성할 수 있도록 허용합니다. 제가 테스트해 본 구글의 Genie 3나 World Labs의 Marble 같은 다른 모델들과 비교했을 때, Oasis 3는 단일 텍스트 프롬프트에서 가장 실사 수준의 사진 같은 환경을 제공했습니다. 또한 사용자가 몇 시간 동안 환경과 상호작용할 수 있다는 점은 Decart의 경쟁사들에게는 없을 수도 있는 수준의 효율성을 시사합니다.
하지만 오랜 시간 세계를 생성하도록 내버려 두면 모델의 품질도 현저히 떨어집니다. 테스트 결과, 시스템은 프롬프트와 일치하는 강력한 초기 장면을 일관되게 구축할 수 있었지만, 공간을 이동할수록 주제의 일관성이 급격히 저하되는 현상을 발견했습니다. 오전의 뉴욕 거리를 생성하라고 프롬프트를 주었더니 아름답게 생성해 냈습니다. 하지만 운전을 계속하자 환경은 뉴욕이라기보다는 서구의 일반적인 도시 어디든 볼 수 있는 풍경처럼 변했습니다. 방향을 바꾸어 처음 교차로로 돌아가려고 하자 그곳은 사라지고 완전히 새로운 환경으로 대체되어 있었습니다. 게다가,
AI startup Decart on Wednesday unveiled Oasis 3, its latest interactive world model that can generate photorealistic driving environments in real time, TechCrunch has exclusively learned. The model is currently available via API. The startup is initially targeting autonomous vehicle companies that need to simulate rare driving scenarios at scale, and plans to expand into robotics and other physical AI applications. But the bigger bet is on developers: By offering API access from day one, Decart is trying to build a developer ecosystem around world models much like how OpenAI did with language models. “It's going to be the first usable world model that people can actually program on top of,” Dean Leitersdorf, co-founder and CEO of Decart, told TechCrunch. “I think there's going to be an entire developer community that emerges on top of this.” The startup already has a community of more than 100,000 developers, many of whom are building products on top of its real-time video model Lucy, largely in e-commerce and live streaming. Oasis 3 is based on that foundation model, and it represents the company’s push into physical AI. Access is priced at $0.02 per second, and enterprise pricing depends on use cases, Decart said. Decart is playing in an increasingly packed world model arena. Last year, Google released Genie 3 in research preview, Fei-Fei Li’s World Labs launched Marble for commercial use cases, and video generation startups like Luma and Runway are also translating their physics-aware video models into world models. Oasis 3's release comes a few weeks after two-year-old Decart raised $300 million, which Leitersdorf says followed “huge demand increases for the models we built” in e-commerce, live streaming and physical AI. The round boosted Decart’s valuation to nearly $4 billion, and brought a series of strategic investors such as Toyota, Adobe and eBay. All of these companies are potential customers, says Leitersdorf. Nvidia, an existing investor, also participated in the round. Oasis 3’s edge lies in the photo-realism of its models and infinite generation capability. That’s due to some efficiency wizardry on Decart’s part, powered by the company’s other main product: the DOS (Decart Optimization Stack) software that allows models to run efficiently on Nvidia, Amazon and Google hardware, making its models far less expensive to run than competitors. “This is built on top of our entire real-time stack, which we optimize all the way down to the hardware,” Leitersdorf said. “By being so vertically integrated, we’re able to be more than an order of magnitude cheaper than anyone else in the industry in order to run these models.” The startup's models are so efficient, per Leitersdorf, that it has burned through “drastically less” than $100 million in its lifetime. Oasis 3 generates physically accurate, multi-camera environments — one front-facing and two-side facing — for training and testing systems. And instead of offering limited demos and research previews, Decart allows developers to generate scenarios infinitely. Compared to other models I’ve tried, like Google’s Genie 3 or World Labs’s Marble, Oasis 3 delivers the most photorealistic environments from a single text prompt I’ve seen. And the fact that you can interact with them for hours suggests a level of efficiency that Decart's rivals might lack. But by letting you generate a world for so long, the model also degrades significantly. In my testing, I found the system could consistently set up a strong initial scene that matches the prompt, but the thematic integrity degraded rapidly as I moved through the world. I prompted it to generate a New York City street in the morning, it did so, beautifully. But as I drove along, the environment looked less like New York and more like a standard version of any urban, Western city. When I tried to turn around and make my way back to the initial intersection, it was gone, replaced by an entirely new environment. On top of that, the controls aren't very responsive, and I often lost control over where the car was moving (again, a drawback shared by other world models I've tested). The experience felt less like a coherent simulation and more of a dream-like, disjointed stream of consciousness that quickly grows nonsensical. Another issue, which I've also seen in other world models, is that the car will just drive through other cars, meaning the model doesn’t simulate physics properly in the environment. Leitersdorf calls this a “major research problem that we’re cracking now,” attributing it to the fact that “there’s drastically more data on good driving compared to accidents.” Part of what makes this physics consistency difficult is fundamental to how this world model works. Oasis 3 is auto-regressive, meaning it generates one frame at a time, and looks back at what it previously generated to decide what comes next. This is a key architectural feature of many world models, and it is a compute-intensive one, too. In order to maintain consistency, Leitersdorf says the Decart team is working to improve the length of the model's memory. “Every frame we generate is roughly 8,000 tokens,” he said. “Generating this at tens of frames per second — that’s hundreds of thousands of tokens per second. The context window fills up very quickly. We’re researching how to do longer context to store millions more tokens, and how to compress the memory into fewer tokens.” Leitersdorf thinks the consistency issue might be partially solved in the model's next version, which will allow users to start generating worlds based on a video of an environment rather than an image. He acknowledged that world models as a field are still early. Still, the founder is less focused on the current limitations of his tech than what will happen when developers get their hands on it. “It takes me back to the early days of LLMs, when OpenAI invented the API for models,” he said, pointing to the emergence of a developer community that advanced the field by finding and building new use cases. "When we talk again in three months, we’ll be like, ‘Here’s 100 developers that all built 100 different applications with Oasis that surprised all of us,'" he said. Topics AI , decart , Exclusive , oasis 3 , world models When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Rebecca Bellan Senior Reporter Rebecca Bellan is a senior reporter at TechCrunch where she covers the business, policy, and emerging trends shaping artificial intelligence. Her work has also appeared in Forbes, Bloomberg, The Atlantic, The Daily Beast, and other publications. You can contact or verify outreach from Rebecca by emailing rebecca.bellan@techcrunch.com or via encrypted message at rebeccabellan.491 on Signal. View Bio June 18 Los Angeles Get an inside look at what it takes to scale and succeed from leaders at Mach Industries, Founders Fund, and Shinkei Systems. Through candid fireside chats and high-impact networking, you'll walk away with valuable insights and new connections. REGISTER NOW Most Popular Google just fired a warning shot in the AI subscription price wars Lucas Ropek Connie Loizos WWDC 2026: Everything announced on Siri AI, iOS 27, Apple Intelligence, and more Morgan Little Aisha Malik Anthropic's Claude Fable 5 is a version of Mythos the public can access today Rebecca Bellan It's not FAANG anymore. It's MANGOS. Julie Bort Microsoft's open source tools were hacked to steal passwords of AI developers Zack Whittaker Google will pay SpaceX $920M per month for compute Sean O'Kane Mira Murati steps back into the spotlight, carefully Connie Loizos