BL
MarkTechPost • 11일 전
리워드 AI, 원격조작 없이 인간 시연 데이터만으로 학습한 로봇 정책 'OM-1' 공개
IMP 7/10
핵심 요약
Reward AI가 7축(7-DoF) 웨어러블 글러브로 수집한 인간 시연 데이터만으로 학습된 범용 로봇 조작 정책 OM-1(Omnibody Model 1)을 공개했습니다. 이 정책은 산업용 로봇팔과 휴머노이드에서 인간 속도로 작동하며, 30분 미만의 데이터로 새 작업을 학습할 수 있습니다. 아직 가중치, 코드, API는 공개되지 않았습니다.
번역된 본문
Reward AI가 OM-1(Omnibody Model 1)을 공개했습니다. 이는 7자유도(7-DoF) 웨어러블 글러브로 캡처한 인간 시연 데이터만으로 완전히 학습된 범용 조작 정책으로, 원격조작(teleoperation) 데이터나 실제 로봇에서 수집한 온-로봇 데이터는 전혀 사용되지 않았습니다. 이 정책은 산업용 로봇팔과 휴머노이드에서 인간과 같은 속도로 작동하며, 30분 미만의 데이터만으로 새로운 작업을 학습할 수 있습니다. 또한 전자기식 손 추적 기술(초당 67cm 속도에서 시각-관성 추적 방식보다 오버슈트가 60% 낮음)과 자체 클록으로 작동하는 강화학습(RL) 기반 제어 레이어를 결합했습니다. 현재까지 가중치, 코드, API는 공개되지 않았습니다.
원문 보기 (영어)
Reward AI has released OM-1 (Omnibody Model 1), a general-purpose manipulation policy trained entirely on human demonstrations captured with a 7-DoF wearable glove, with no teleoperation or on-robot data. The policy runs on industrial arms and humanoids at human speed, learns a new task from under 30 minutes of data, and pairs electromagnetic hand tracking (60% lower overshoot than visual-inertial at 67 cm/s) with an RL-trained control layer that runs on its own clock. No weights, code, or API are public yet.
The post Reward AI Releases OM-1: A Robot Policy Trained on Human Demonstrations Only, With No Teleoperation or On-Robot Data appeared first on MarkTechPost.