NVIDIA has released the Nemotron-4 340B model family, comprising Base, Instruct, and Reward variants, along with the synthetic data generation pipeline used for alignment. The models are open access under the NVIDIA Open Model License Agreement and are sized to fit on a single DGX H100 with 8 GPUs in FP8 precision.

  • Nemotron-4-340B-Base, Instruct, and Reward variants are available under a permissive license allowing distribution and modification.
  • The models perform competitively against open-access benchmarks and were optimized to fit on a single DGX H100 with 8 GPUs in FP8 precision.
  • Over 98% of the data used in the model alignment process was synthetically generated, demonstrating the effectiveness of the approach.
  • NVIDIA is open-sourcing the synthetic data generation pipeline to support open research and facilitate training smaller language models.

The release aims to benefit the community in research studies and commercial applications, particularly for generating synthetic data to train smaller language models.