Artificial Intelligence
August 6, 2026
Aug. 6, 2026 — In July, NVIDIA joined more than 200 companies and organizations in signing “Open Weights and American AI Leadership,” an open letter arguing that AI leadership will be measured not by any single frontier model but by whether an open ecosystem reaches every sector.
Open models, which anyone can download, inspect, modify and run on their own infrastructure, are what make that possible. Nowhere is that more crucial than in physical AI, where every deployment is a specialization problem.
Physical AI has to understand and predict consequences, not just appearances.
To make this possible, world models learn how physical environments behave, what may happen next and which following actions make sense. They can generate physically grounded world and action data, simulate future states and provide a foundation that teams can specialize for a robot, autonomous vehicle or vision AI system.
Open world models are already being used to generate training data, test policies and specialize physical AI systems. NVIDIA Cosmos 3 brings these capabilities together in an open model family, with leading benchmark results and adoption across robotics, autonomous vehicles and vision AI.
And NVIDIA Omniverse libraries, part of NVIDIA Agent Toolkit, provides prebuilt capabilities for building simulation-ready worlds that physical AI teams can use to train, test and validate systems before real-world deployment.
World Models Are the Foundation of Physical AI
The data behind physical AI is difficult and expensive to collect at the scale required. Rare events and long-tail scenarios can be especially difficult to reproduce safely and repeatedly.
- More useful data by learning physical relationships from large-scale multimodal scenarios.
- More diverse environments that vary in weather, lighting, objects and trajectories.
- A better foundation to build on and adapt to a particular robot, vehicle, sensor configuration, task or operating environment.
A general model hasn’t seen a team’s particular robot, sensors or operating environment. Closing that gap requires access to model weights, a license that permits adaptation and the tools needed for post-training.
NVIDIA Cosmos world foundation models are available under the Linux Foundation’s OpenMDW 1.1 license, enabling teams to post-train models on their own data and hardware. Specialization is where openness becomes a practical technical requirement.
Specializing a model is only part of the workflow. Teams also need environments to generate data, run simulations and test behavior.
Omniverse libraries help developers build simulation-ready environments, while OpenUSD provides the open framework for composing, reusing and exchanging complex 3D data across digital twins, simulations and synthetic data generation workflows. Together, Omniverse and OpenUSD cut the duplicated work that can otherwise pile up every time assets, sensor configurations or environmental conditions change.
NVIDIA Cosmos 3 — a frontier open physical AI foundation omni-model built on a mixture-of-transformers architecture — combines vision reasoning, world generation and action prediction, letting developers use one model family to understand scenes, generate synthetic data, simulate future states and build specialized world action models.
Developers can use Cosmos 3 as a vision language model, as a physics-grounded world simulator that predicts future world states and generates large-scale synthetic data, or as the backbone for world action models, instead of assembling and maintaining a separate model for each capability.
The family includes Cosmos 3 Super (64B) for high-fidelity world modeling, Cosmos 3 Nano (16B) for efficient reasoning and post-training, and Cosmos 3 Edge (4B) for on-device vision reasoning and robot policy deployment. Lightweight enough to run on edge GPUs, Cosmos 3 Edge can be deployed across NVIDIA RTX GPUs, NVIDIA DGX systems and NVIDIA Jetson, including Jetson Thor platforms.
Across benchmark evaluations, Cosmos 3 ranks No. 1 on Artificial Analysis for open weights text-to-image and image-to-video generation, on PAI-Bench for world generation and in the image-to-video category of Physics-IQ. For robot policy, it ranks No. 1 on RoboLab. Cosmos 3 Super is also the highest-ranked open model on VANTAGE-Bench for vision understanding.
In addition to Cosmos, NVIDIA’s physical AI stack includes Isaac GR00T for robotics, Alpamayo for autonomous vehicles and Metropolis for vision AI.
How Developers Are Putting Cosmos 3 to Work
Across industries, developers are building on NVIDIA Cosmos for physical AI applications: Doosan Robotics, LG Electronics, Samsung Electronics and Skild AI in robotics; Li Auto, Xiaomi and Afari in autonomous vehicles; and Centific, Fogsphere, Linker Vision, Milestone Systems and Yuan for vision AI agents powering industrial AI and smart spaces applications.
The NVIDIA Cosmos Coalition extends this work by bringing together world model builders, AI developers and physical AI leaders to contribute models, research and evaluation methods. NVIDIA recently expanded the coalition to Japan, where robotics and manufacturing leaders intend to join and develop open world models for factories, logistics, agriculture, construction, healthcare and transportation.
Together, these implementations and collaborations are establishing open world models as an adaptable foundation for physical AI across robots, autonomous vehicles and vision AI systems.
The Big Data Inside Amazon’s New Fire Phone
The new Fire phone that Amazon launched this week looks like your ordinary black smartphone,…
Three Reasons to be Scared of the Internet of Things
We know the Internet of Things forecasts: 50 billion connected devices by 2020. Apparently, there’s…
Wanted: Intelligent Middleware That Simplifies Big Data Analytics
We’ve seen tremendous technological innovation in the data analytics space over the past 10 years….
GPUs Tackle Massive Data of the Hive Mind
LIVE from GTC12 — The flock of birds that weaves seamlessly through the sky, propelled…
Why Hadoop on IBM Power
In the quest to achieve data-driven insight, Hadoop running on Intel X86-based processors has emerged…
Datanami’s Leverage Big Data Summit Wraps Up
Dialog and networking were on the Datanami agenda this week as we kicked off our…
What Google DeepMind’s Departures Say About the AI Talent War
Google DeepMind has spent the past few years at the center of the AI boom….
Why AI Still Hasn’t Had Its Einstein Moment
AI is having a remarkable run in science. It’s helping researchers predict protein structures, improve…
AI Infrastructure Spending Has Surpassed $1 Trillion. Here’s Where the Money Is Going and What’s Next
AI used to be a race to build the best model. Now it’s a race…
Token Optimization in Enterprise AI: How Context Architecture Determines Whether Your AI Program Scales Economically
Enterprise AI budget forecasting is mostly guesswork right now. Most organizations set their initial allocations…
AI Is Creating a New Market for Infrastructure Data. S&P Global Wants In
The AI boom has kicked off a massive wave of data center construction around the…
Why the AI Semantic Layer Is Becoming the Foundation of Enterprise AI
BigDATAwire recently spoke with Cindi Howson, Chief Data and AI Strategy Officer (CAIO) at ThoughtSpot,…
