본문으로 건너뛰기

Cosmos 3, physical AI용 open world model을 로봇·차량·엣지로 확장

Original: Into the Omniverse: How Open World Models Push the Frontier of Physical AI View original →

Read in other languages: English日本語
Humanoid Robots Aug 7, 2026 By Insights AI 1 min read Source

physical AI 개발에서 모델을 따로 조립하던 부담을 줄이려는 NVIDIA의 답은 Cosmos 3다. NVIDIA는 2026년 8월 6일 글에서 Cosmos 3를 open physical AI foundation omni-model family로 설명했다. 하나의 계열 안에서 vision reasoning, world generation, action prediction을 다루고, 로봇·자율주행·vision AI가 같은 기반 모델 위에서 장면 이해와 synthetic data 생성, 미래 상태 예측을 수행하도록 하는 접근이다.

구성도 구체적이다. Cosmos 3 Super는 64B 규모의 high-fidelity world modeling용 모델이고, Cosmos 3 Nano는 16B 규모로 reasoning과 post-training에 초점을 둔다. Cosmos 3 Edge는 4B 모델로, on-device vision reasoning과 robot policy deployment를 겨냥한다. NVIDIA는 Edge가 RTX GPU, DGX systems, Jetson, Jetson Thor 플랫폼에서 실행될 수 있다고 설명했다.

성능 주장은 benchmark 중심이다. NVIDIA에 따르면 Cosmos 3는 open weights text-to-image와 image-to-video generation에서 Artificial Analysis 1위, PAI-Bench world generation과 Physics-IQ image-to-video 부문 1위, robot policy용 RoboLab 1위를 기록했다. Cosmos 3 Super는 VANTAGE-Bench vision understanding에서 가장 높은 open model로 제시됐다. 단일 지표 하나가 아니라 생성, 물리 예측, robot policy, vision understanding을 함께 내세운 점이 이 릴리스의 무게다.

채택 사례도 제조와 이동체 쪽으로 넓다. NVIDIA는 Doosan Robotics, LG Electronics, Samsung Electronics, Skild AI가 robotics 분야에서, Li Auto, Xiaomi, Afari가 autonomous vehicles 분야에서 Cosmos를 쓰고 있다고 밝혔다. 모델과 dataset은 Hugging Face와 GitHub에서 접근할 수 있다. 다음 쟁점은 benchmark 1위보다 현장 배포다. simulation에서 만든 world model이 실제 공장, 물류, 차량, smart space에서 얼마나 안정적으로 정책 학습과 안전 검증을 줄여주는지가 핵심이다.

Share: Long

Related Articles