NVIDIA has released the Nemotron-4 340B model family, comprising Base, Instruct, and Reward variants, under a permissive open license. The models are designed to fit on a single DGX H100 with 8 GPUs in FP8 precision and were trained on 9 trillion tokens.
- Nemotron-4-340B-Instruct surpasses comparable instruct models in instruction following and chat capabilities.
- Nemotron-4-340B-Reward achieves top accuracy on RewardBench, outperforming proprietary models like GPT-4o-0513 and Gemini 1.5 Pro-0514.
- Over 98% of the data used for model alignment was synthetically generated.
- NVIDIA is open-sourcing the synthetic data generation pipeline, including prompts and quality filtering tools.
The release aims to support community research and commercial applications by providing high-quality models and tools for generating synthetic training data.