메뉴
BL
MarkTechPost • 52일 전

엔비디아, 로보택시 자율주행 34B 오픈소스 모델 공개

IMP
8/10
핵심 요약

엔비디아가 자율주행 및 로보택시를 위한 34B 크기의 비전-언어-액션(VLA) 모델 '알파마요 2 슈퍼(Alpamayo 2 Super)'를 상업적 활용 및 파인튜닝이 가능한 OpenMDW-1.1 라이선스로 전격 공개했습니다. 이 모델은 강력한 추론 기반 모델과 확산 액션 디코더(diffusion action decoder)를 결합하여 단 한 번의 처리만으로 주행 궤적과 인과 관계 추적 등 핵심 주행 데이터를 정확하게 생성해 내는 것이 특징입니다.

번역된 본문

엔비디아는 자율주행을 위한 34B(340억 매개변수) 규모의 비전-언어-액션 모델(Vision-Language-Action Model)인 '알파마요 2 슈퍼(Alpamayo 2 Super)'를 OpenMDW-1.1 라이선스 하에 출시했습니다. 이 라이선스는 파인튜닱, 2차적 파생물 제작, 상업적 재배포를 허용하는 관대한 규정입니다. 이 모델은 32B 크기의 '코스모스 3 슈퍼 리즈너(Cosmos 3 Super Reasoner)' 백본(backbone)을 2.3B 규모의 확산 액션 디코더(diffusion action decoder)와 결합하여 LingoQA 벤치마크에서 79.2점을 기록했습니다. 또한 단 한 번의 처리(pass)만으로 주행 궤적(trajectories), 인과 관계 추적(Chain-of-Causation traces), 메타 액션(meta-actions), 자동 라벨링(auto-labels), 그리고 그라운디드 VQA(grounded VQA, 시각 기반 질의응답)를 생성해 냅니다.

엔비디아 출시 소식: 알파마요 2 슈퍼(Alpamayo 2 Super)는 로보택시와 자율주행을 위한 OpenMDW-1.1 기반의 34B 오픈소스 비전-언어-액션 모델입니다 (이 글은 MarkTechPost에 가장 먼저 게재되었습니다).

원문 보기
원문 보기 (영어)
NVIDIA released Alpamayo 2 Super, a 34B vision-language-action model for autonomous driving, under OpenMDW-1.1 — a permissive license covering fine-tuning, derivatives and commercial redistribution. It pairs a 32B Cosmos 3 Super Reasoner backbone with a 2.3B diffusion action decoder, scores 79.2 on LingoQA, and emits trajectories, Chain-of-Causation traces, meta-actions, auto-labels and grounded VQA from a single pass. The post NVIDIA Releases Alpamayo 2 Super: A 34B Open Vision-Language-Action Model for Robotaxis and Autonomous Driving Under OpenMDW-1.1 appeared first on MarkTechPost.