메뉴
BL
TechCrunch AI 47일 전

인도 AI 스타트업 아바타르, 저렴하고 빠른 맞춤형 비디오 AI 공개

IMP
7/10
핵심 요약

인도의 AI 스타트업 아바타르(Avataar)는 알리바바의 오픈소스 모델을 경량화해 인도의 문화적 맥락을 이해하는 초저가 비디오 생성 AI '바리야(Varya)'를 선보였습니다. 이 모델은 기존 모델 대비 10배 빠르고 20분의 1 수준의 저렴한 가격으로 제공되며, 인도 정부의 AI 플랫폼을 통해 오픈웨이트로 공개됩니다.

번역된 본문

미국, 유럽, 중국과 비교할 때 인도의 AI 모델 생산 속도는 다소 더뎠습니다. 모델을 출시하는 스타트업은 소수에 불과하며, 대부분 대형 언어 모델(LLM)이나 음성 모델에 국한되었습니다. 개발을 촉진하기 위해 인도 정부는 약 12억 달러 규모의 '인도 AI 미션(India AI Mission)' 이니셔티브를 출범했습니다. 이 프로그램의 일환으로, 선정된 스타트업들에게는 보조금이 적용된 GPU 컴퓨팅 자원을 제공하고 그 대가로 모델을 공개하도록 하고 있습니다.

이 프로그램에 선정된 12개 스타트업 중 하나인 아바타르 AI(Avataar AI)는 다양한 지역 축제, 음식, 의상 등 로컬 문맥을 이해하도록 설계된 새로운 비디오 모델인 '바리야(Varya)'를 출시했습니다. 이커머스용 비디오 제작 도구에 주력하는 이 피크 XV(Peak XV) 투자 유치 스타트업은 바리야를 처음부터 새로 개발하지는 않았습니다. 알리바바가 공개한 비디오 생성 모델인 Wan 2.2를 기반으로 삼아, 모델의 기능을 압축해 아바타르의 특정 사용 사례에 최적화된 보다 가볍고 빠른 버전으로 만드는 '증류(Distillation)' 기술을 활용했습니다.

그 결과, 이 모델은 Wan 2.2의 50단계가 아닌 단 4단계의 과정을 거쳐 작동하며, 기존보다 10배 빠르고 비용은 극히 일부만 소모됩니다. 구체적인 수치로 설명하자면, 엔비디아 H200 GPU를 사용할 때 바리야는 5초 분량의 720p 클립을 45초 만에 생성합니다. 이는 Wan 2.2가 소요하는 1,230초와 비교되는 엄청난 속도입니다.

바리야의 가장 돋보이는 특징은 역시 가격입니다. 이 회사는 자사 호스팅 서비스에서 영상 1초당 0.48 루피(약 0.005달러)를 청구할 계획입니다. 이는 일반적으로 영상 1초당 0.10달러 이상을 부과하는 Veo, Kling, Luma, Runway 등 기존 모델들에 비해 훨씬 저렴한 가격으로, 약 20배의 가격 차이가 납니다.

피크 XV의 라잔 아난단(Rajan Anandan) 이사는 TechCrunch와의 인터뷰에서 "인도는 비디오가 우선시되는 시장입니다. 인도의 모든 대형 소비자 인터넷 제품에서 비디오가 텍스트를 이깁니다. 현재의 AI 비디오 모델은 인도의 인구 규모에 맞춰 사용하기에는 너무 비쌉니다. 비디오 AI가 학생, 교사, 중소기업(MSME), 크리에이터, 기업 및 공공 서비스에 도달하려면 비용이 획기적으로 낮아져야 합니다. 비용이야말로 인도에서 AI 도입을 촉진하는 가장 큰 열쇠입니다"라고 말했습니다.

이미지 및 비디오 생성 모델은 종종 문화적 뉘앙스를 놓치고 고정관념에 갇히거나 평범한 결과물을 생성하는 문제가 있습니다. 아바타르 AI는 큐레이션된 데이터를 사용해 바리야가 음식, 의상, 건축물 및 축제를 포함한 인도의 문화적 특징을 인식하도록 학습시켰다고 설명했습니다.

바리야는 자체 학습 데이터와 함께 인도 정부의 공개 AI 모델 및 데이터셋 중앙 저장소인 'AI Kosh' 포털에 오픈 웨이트(Open-weight) 모델로 공개될 예정입니다. 이로써 개발자들은 자체 서버에 호스팅하거나 필요에 맞게 모델을 수정할 수 있게 됩니다. 아바타르는 기업 고객에게도 이 모델을 제공할 계획이며, Higgsfield 및 Adobe Firefly 등 타 비디오 도구와의 파트너십에도 열려 있다고 밝혔습니다. 현재 누구나 텍스트 프롬프트나 참조 이미지를 통해 웹사이트에서 이 모델을 시험해 볼 수 있습니다.

바리야의 출시는 인도의 AI 야망에 대한 근본적인 트레이드오프를 보여줍니다. 업계 베테랑들은 인도가 파운데이션 모델(Foundation Model) 경쟁을 하는 대신, 애플리케이션 구축과 탄탄한 개발자 생태계 조성을 통해 AI 분야에서 입지를 다져야 한다고 지적해 왔습니다. 이러한 실용주의에는 분명한 이유가 있습니다. 컴퓨팅 자원 부족과 양질의 데이터 확보의 한계로 인해 인도의 모델 개발은 글로벌 경쟁국들보다 더뎠기 때문입니다.

'인도 AI 미션'은 이러한 격차를 줄이기 위한 정부의 광범위한 노력의 일환입니다. 작년에 정부는 12개 스타트업(아바타르 AI 포함)을 선정하여 AI 모델을 개발하도록 지원하고 비용 효율적인 컴퓨팅 자원을 제공했습니다. 올해 초, 인도의 전자정보기술부 장관 아쉬비니 바이슈나우(Ashwini Vaishnaw)는 인도가 2028년까지 2,000억 달러의 AI 투자를 유치하고, 6개월 이내에 GPU 수용 능력을 두 배 이상 늘리는 것을 목표로 하고 있다고 밝혔습니다.

원문 보기
원문 보기 (영어)
India's AI model output has been slow compared to the U.S., Europe, and China. Only a few startups are releasing models, and most of them are large language models or voice models. To encourage more development, the government launched the India AI Mission , a roughly $1.2 billion initiative that — among other things — gives selected startups access to subsidized GPU compute in exchange for releasing their models publicly. One of the 12 startups selected for the program, Avataar AI , has launched a new video model called Varya that is built to understand local context — such as identifying different festivals, food, and clothing. The Peak XV-backed startup, which focuses on creating video tools for e-commerce , didn't build Varya from scratch. It started with Wan 2.2, a publicly available video generation model released by Alibaba, and used a technique called distillation — essentially compressing the model's capabilities into a leaner, faster version optimized for Avataar's specific use cases. The result is a model that runs in four steps rather than Wan 2.2's 50, producing video 10 times faster and at a fraction of the cost. To put that in concrete terms: using an NVIDIA H200 GPU, Varya can generate a 5-second 720p clip in 45 seconds, compared to 1,230 seconds for Wan 2.2. The most striking aspect of Varya may be its price. The company plans to charge ₹0.48 ($0.005) per second of video on its hosted service — far cheaper than models like Veo, Kling, Luma, and Runway, which typically charge $0.10 or more per second. That's a roughly 20x price difference. "India is a video-first market. We see this across every large consumer internet product in India: video wins over text. Current AI video models are too expensive for population-scale use in India. If video AI is going to reach students, teachers, MSMEs, creators, enterprises, and public services, costs have to come down dramatically. Cost is the biggest unlock for AI adoption in India," Peak XV's managing director Rajan Anandan told TechCrunch. Image and video generation models often miss cultural nuances and produce stereotyped or generic outputs — a problem TechCrunch has reported on before . Avataar AI says it has used curated data to train Varya to recognize cultural nuances including food, clothing, architecture, and festivals. Varya will be released as an open-weight model on India's AI Kosh portal — the Indian government's centralized repository for publicly available AI models and datasets — along with its training data, meaning developers can self-host or modify it for their own needs. Avataar also plans to make the model available to its enterprise customers and says it is open to partnerships with video tools including Higgsfield and Adobe Firefly. Anyone can try it now on its website using text prompts or reference images. Varya's launch reflects a fundamental tradeoff in India's AI ambitions. Industry veterans have noted that India can make its mark in AI by creating applications and a robust developer ecosystem rather than competing on foundation models. And there's a reason for that pragmatism: model development has been slower in India than in global rivals due to a lack of compute and limited quality data availability. The India AI Mission is also part of a broader government push to close that gap. Last year, it selected 12 startups — Avataar AI among them — to develop AI models and provided them with cost-efficient compute. Earlier this year, IT minister Ashwini Vaishnaw said India aims to attract $200 billion in AI investment by 2028 and more than double its GPU capacity within six months. Topics AI , Avataar , India , india ai , Peak XV When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence. Ivan Mehta Ivan covers global consumer tech developments at TechCrunch. He is based out of India and has previously worked at publications including Huffington Post and The Next Web. You can contact or verify outreach from Ivan by emailing im@ivanmehta.com or via encrypted message at ivan.42 on Signal. View Bio June 18 Los Angeles Get an inside look at what it takes to scale and succeed from leaders at Mach Industries, Founders Fund, and Shinkei Systems. Through candid fireside chats and high-impact networking, you'll walk away with valuable insights and new connections. REGISTER NOW Most Popular Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable Lorenzo Franceschi-Bicchierai Google just fired a warning shot in the AI subscription price wars Lucas Ropek Connie Loizos WWDC 2026: Everything announced on Siri AI, iOS 27, Apple Intelligence, and more Morgan Little Aisha Malik Anthropic's Claude Fable 5 is a version of Mythos the public can access today Rebecca Bellan It's not FAANG anymore. It's MANGOS. Julie Bort Microsoft's open source tools were hacked to steal passwords of AI developers Zack Whittaker Google will pay SpaceX $920M per month for compute Sean O'Kane