// Workers AI · dad joke modeWhat did the Nemotron say? "I'm attracted to data.
| Nemotron | |
|---|---|
| Developer | Nvidia |
| Release | November 15, 2023 |
| Type | Large language model |
| License | Various Nvidia Open Model License[1] |
| Website | developer |
| Repository | github |
Nemotron is a family of artificial intelligence models developed by Nvidia. It includes large language models and multimodal models intended for reasoning, computer programming, information retrieval and agentic AI applications. Nvidia has released open model weights, training data, software and training methods for parts of the family.[2][3]
History
[edit]Nvidia introduced the Nemotron-3 8B models in November 2023 for enterprise generative artificial intelligence applications.[4] Nemotron-4 340B followed in June 2024 and included base, instruction-tuned and reward models designed partly for generating synthetic data.[5]
Nvidia later introduced Llama Nemotron, a series of reasoning models derived from Meta's Llama models.[6] In December 2025, Nvidia announced the Nemotron 3 generation, beginning with Nemotron 3 Nano.[2] Nemotron 3 Super and Ultra were released in 2026.[7][8]
Models
[edit]| Family | Released | Purpose |
|---|---|---|
| Nemotron-3 | 2023 | Enterprise applications[4] |
| Nemotron-4 | 2024 | Synthetic-data generation[5] |
| Llama Nemotron | 2025 | Llama-based reasoning[6] |
| Nemotron 3 | 2025 | Agentic AI[2] |
The Nemotron 3 models use a hybrid architecture combining Mamba, Transformer, and mixture of experts components.[3]
| Model | Parameters | Intended use |
|---|---|---|
| Nano | 30B; 3B active | Efficient agents[2] |
| Super | 120B; 12B active | Agentic reasoning[7][9] |
| Ultra | 550B; 55B active | Complex reasoning[8] |
| Nano Omni | 30B; 3B active | Multimodal agents[10] |
Nano Omni supports text, image, video and audio input.[10] Nemotron models can be downloaded for local deployment or accessed through Nvidia's application programming interfaces and NIM inference services.[11][12]
See also
[edit]References
[edit]- ↑ "NVIDIA Nemotron Open Model License". Nvidia. December 15, 2025. Retrieved July 29, 2026.
- 1 2 3 4 Lee, Jane Lanhee (December 15, 2025). "Nvidia unveils new open-source AI models amid boom in Chinese offerings". Reuters. Retrieved July 29, 2026.
- 1 2 Knight, Will (December 15, 2025). "Nvidia Becomes a Major Model Maker With Nemotron 3". Wired. Retrieved July 29, 2026.
- 1 2 "NVIDIA AI Foundation Models: Build Custom Enterprise Chatbots and Co-Pilots With Production-Ready LLMs". Nvidia Technical Blog. Nvidia. November 15, 2023. Retrieved July 29, 2026.
- 1 2 Nuñez, Michael (June 14, 2024). "Nvidia's Nemotron-4 340B model redefines synthetic data generation, rivals GPT-4". VentureBeat. Retrieved July 29, 2026.
- 1 2 Bercovich, Akhiad; Levy, Itay; Golan, Izik; Dabbah, Mohammad; et al. (May 2, 2025). "Llama-Nemotron: Efficient Reasoning Models". arXiv:2505.00949 [cs.CL].
- 1 2 Knight, Will (March 11, 2026). "Nvidia Will Spend $26 Billion to Build Open-Weight AI Models, Filings Show". Wired. Retrieved July 29, 2026.
- 1 2 "NVIDIA Nemotron 3 Ultra". Nvidia Research. Nvidia. June 4, 2026. Retrieved July 29, 2026.
- ↑ "Introducing Nemotron 3 Super: An Open Hybrid Mamba-Transformer MoE for Agentic Reasoning". Nvidia Technical Blog. Nvidia. March 11, 2026. Retrieved July 31, 2026.
- 1 2 "NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model". Nvidia Technical Blog. Nvidia. April 28, 2026. Retrieved July 29, 2026.
- ↑ Sirodot, Bertrand; Bronzati, Fabricio (November 7, 2025). "Introduction to NVIDIA Inference Microservices, aka NIM". Dell Technologies Info Hub. Dell Technologies. Retrieved July 29, 2026.
- ↑ "NVIDIA Nemotron". Nvidia. Nvidia. Retrieved July 29, 2026.