Claims and evidence
- Released Primary source[1]page date; first and fourth paragraphs
- NVIDIA-hosted API endpoint for Nemotron-4-340B-Instruct deprecated The notice gives 04/02/2025, read month first as NVIDIA writes dates elsewhere (its blogs use forms such as 10/21/2024). Applies to the hosted API in the NVIDIA API catalog only, not to the downloadable weights. Primary source[4]deprecation notice
- Status AvailableWeights remain downloadable from NVIDIA's Hugging Face organisation. The NVIDIA-hosted API endpoint for the Instruct model was deprecated (see milestones). Primary source[2] [5]
- Change · Size 340 billion parameters, against 15 billion for Nemotron-4 15B, with a similar architecture.Compared with Nemotron-4 15B Primary source[3]
- Change · Training data Uses the same data blend as Nemotron-4 15B but trains on 9 trillion tokens: the first 8 trillion plus 1 trillion tokens of continued pre-training.Compared with Nemotron-4 15B Primary source[3]
- Change · Availability Weights released for download under the NVIDIA Open Model License, whereas Nemotron-4 15B was published as a technical report only.Compared with Nemotron-4 15B Primary source[1] [2]
- Open weights YesReleased under the NVIDIA Open Model License, which allows commercial use and distribution of derivative models. Primary source[1]fourth paragraph [2]License [3]Abstract
- Context window 4,096 tokens tokens Primary source[2]Model Overview [5]Model Overview
- Parameters 340 billionThe technical report gives 9.4 billion embedding and 331.6 billion non-embedding parameters. Primary source[2]Model Overview
- Access Open-weights download, APIDownloads from NGC and Hugging Face at launch; the announcement said access at ai.nvidia.com as an NVIDIA NIM would follow, and the Instruct card points to build.nvidia.com. The hosted Instruct endpoint was later deprecated. Primary source[1]fourth paragraph [5]Model Overview [4]
Lineage
Predecessors
No known predecessor.
Successors
No known successor.
Based on
Not derived from another model.
Variants and derived
None recorded.
Siblings
None recorded.
All ancestors
None.
All descendants
None.
Variants
Nemotron-4-340B-Base
Same dates as Nemotron-4 340B.
- Variant Inline variant in this record.Base completion model, pre-trained on 9 trillion tokens. Primary source[2]Model Overview [3]Abstract
Also known as: nvidia/Nemotron-4-340B-Base
Nemotron-4-340B-Instruct
Same dates as Nemotron-4 340B.
- Variant Inline variant in this record.Fine-tuned from the Base model for English single- and multi-turn chat; alignment used SFT, DPO and NVIDIA's Reward-aware Preference Optimization. Primary source[5]Model Overview [3]Abstract
Also known as: nvidia/Nemotron-4-340B-Instruct
Related AI Radar coverage
AI Radar coverage starts in June 2026; no coverage linked yet.