NVIDIA’s open push is aggressive, but I would not read it as model-lab openness. The package spans 10T language tokens, 500K robotics trajectories, 455K protein structures, and 100TB of vehicle sensor data across Nemotron, Cosmos, GR00T, Alpamayo, and Clara. Those assets fit NVIDIA’s training, simulation, and inference stack best.
I don’t buy the “advance AI across every industry” framing. OpenAI and Anthropic fight for API mindshare. Meta used Llama to win developer distribution. NVIDIA is grabbing the data formats, benchmarks, and toolchain entry points around physical AI and enterprise agents. The article does not give license detail or commercial-use boundaries, so the “open” label still needs a discount.