NVIDIA has made the NemotronLabs-VoiceChat-11B voice chat model available on Hugging Face. The repository is described as supporting full duplex communication capabilities.
NVIDIA releases NemotronLabs-VoiceChat-11B model on Hugging Face
Nvidia releases Nemotron-Labs-Audex-30B-A3B unified audio-text LLM
Nvidia has released Nemotron-Labs-Audex-30B-A3B, a unified audio-text large language model built on the Nemotron-Cascade-2-30B-A3B text-only MoE backbone. The model extends the original architecture with an audio encoder for inputs and discrete audio tokens for outputs, enabling capabilities in speech recognition, translation, text-to-speech, and audio generation.
NVIDIA releases Alpamayo 2 Super, a 34B open VLA model for autonomous driving
NVIDIA has released Alpamayo 2 Super, a 34-billion-parameter vision-language-action (VLA) model designed for robotaxis and autonomous driving under the OpenMDW-1.1 license. The model targets long-tail events by combining a 32B Cosmos 3 Super Reasoner backbone with a 2.3B diffusion-based action decoder to generate trajectories, causal explanations, and meta-actions from multi-camera video.
Anthropic releases Claude Opus 5; NVIDIA calls for open-weight policies
Anthropic introduced Claude Opus 5 as a more efficient model that approaches the capabilities of Claude Fable 5 at half the price, becoming the default for Claude Max. Concurrently, NVIDIA is advocating for US government policies to support open-weight AI models to foster innovation and enhance American leadership in the sector.
Jensen Huang cites open-weight model in Hugging Face incident to launch Open Secure AI Alliance
NVIDIA CEO Jensen Huang stated that an open-weight frontier model helped contain a security intrusion during the recent Hugging Face incident, whereas closed AI systems blocked essential forensic analysis. He cited this event as the primary motivation for establishing the Open Secure AI Alliance.
NVIDIA releases Cosmos 3 Edge, a 4B open world model for on-device robot reasoning
NVIDIA has released Cosmos 3 Edge, a 4-billion-parameter open world model designed to run on-device for robots and vision AI agents. Released July 20 on Hugging Face, it enables systems to understand surroundings, reason in real time, and generate robot actions locally without cloud dependency.