NVIDIA has released the Nemotron-3.5-Lightning-30B-A3B model on Hugging Face.
The checkpoint is available under the identifier nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.
NVIDIA has released the Nemotron-3.5-Lightning-30B-A3B model on Hugging Face.
The checkpoint is available under the identifier nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.
Alibaba has released the open weights for Qwen3.8-2.4T-A95B (also known as Qwen3.8-Max), its largest open-weight model, designed to bring near-frontier capabilities to the open ecosystem.
NVIDIA has released two open-source artifacts designed for building always-on AI agents: the Nemotron 3.5 Lightning model and the NeMo Switchyard routing library. These tools address the high cost and latency of sending every step of long-running agent workflows to frontier reasoning models by providing specialized execution and routing capabilities.
NVIDIA has released an update to its Magpie Multilingual TTS model, expanding support to twelve languages with the addition of Modern Standard Arabic, Korean, and Brazilian Portuguese. The release also includes quality improvements across existing languages through updated training data and architectural changes.
Meta has released Muse Glimmer, a 30B open-weight dense model with a 120K+ context window designed for local AI agentic work. Optimized to run across NVIDIA edge, desktop, and workstation platforms, the model delivers 20K tokens/sec on a single GPU.
NVIDIA has released NemotronLabs VoiceChat 11B, an open 11B parameter end-to-end speech-to-speech model designed for real-time, full-duplex conversation. Unlike cascaded stacks that chain ASR, LLM, and TTS, this unified network performs streaming speech understanding and generation simultaneously, achieving a measured smooth turn-taking latency of 448 ms on Full-Duplex-Bench 1.0.
We use cookies to measure traffic and improve the site. You can accept or decline analytics cookies. Privacy policy