Edge Rewrite
// HTMLRewriter · presentation

This page was redesigned at the edge.

Cloudflare fetched the original article and streamed it through HTMLRewriter to apply an entirely new visual system without rebuilding the source page.

// request.cf · coarse context

A page that knows where it met you.

Only coarse request metadata is shown. This demo does not display or persist visitor IP addresses.

Country
US
Cloudflare location
CMH
Connection
HTTP/2
Language
Not provided

Ray ID: a23ee8950ac01dfc

Jump to content

// Workers AI · dad joke modeWhat did the Nemotron say? "I'm attracted to data.

From Wikipedia, the free encyclopedia
Nemotron
DeveloperNvidia
ReleaseNovember 15, 2023; 2 years ago (2023-11-15)
TypeLarge language model
LicenseVarious
Nvidia Open Model License[1]
Websitedeveloper.nvidia.com/topics/ai/nemotron
Repositorygithub.com/NVIDIA-NeMo/Nemotron

Nemotron is a family of artificial intelligence models developed by Nvidia. It includes large language models and multimodal models intended for reasoning, computer programming, information retrieval and agentic AI applications. Nvidia has released open model weights, training data, software and training methods for parts of the family.[2][3]

History

[edit]

Nvidia introduced the Nemotron-3 8B models in November 2023 for enterprise generative artificial intelligence applications.[4] Nemotron-4 340B followed in June 2024 and included base, instruction-tuned and reward models designed partly for generating synthetic data.[5]

Nvidia later introduced Llama Nemotron, a series of reasoning models derived from Meta's Llama models.[6] In December 2025, Nvidia announced the Nemotron 3 generation, beginning with Nemotron 3 Nano.[2] Nemotron 3 Super and Ultra were released in 2026.[7][8]

Models

[edit]
Major model families
Family Released Purpose
Nemotron-3 2023 Enterprise applications[4]
Nemotron-4 2024 Synthetic-data generation[5]
Llama Nemotron 2025 Llama-based reasoning[6]
Nemotron 3 2025 Agentic AI[2]

The Nemotron 3 models use a hybrid architecture combining Mamba, Transformer, and mixture of experts components.[3]

Nemotron 3 models
Model Parameters Intended use
Nano 30B; 3B active Efficient agents[2]
Super 120B; 12B active Agentic reasoning[7][9]
Ultra 550B; 55B active Complex reasoning[8]
Nano Omni 30B; 3B active Multimodal agents[10]

Nano Omni supports text, image, video and audio input.[10] Nemotron models can be downloaded for local deployment or accessed through Nvidia's application programming interfaces and NIM inference services.[11][12]

See also

[edit]

References

[edit]
  1. "NVIDIA Nemotron Open Model License". Nvidia. December 15, 2025. Retrieved July 29, 2026.
  2. 1 2 3 4 Lee, Jane Lanhee (December 15, 2025). "Nvidia unveils new open-source AI models amid boom in Chinese offerings". Reuters. Retrieved July 29, 2026.
  3. 1 2 Knight, Will (December 15, 2025). "Nvidia Becomes a Major Model Maker With Nemotron 3". Wired. Retrieved July 29, 2026.
  4. 1 2 "NVIDIA AI Foundation Models: Build Custom Enterprise Chatbots and Co-Pilots With Production-Ready LLMs". Nvidia Technical Blog. Nvidia. November 15, 2023. Retrieved July 29, 2026.
  5. 1 2 Nuñez, Michael (June 14, 2024). "Nvidia's Nemotron-4 340B model redefines synthetic data generation, rivals GPT-4". VentureBeat. Retrieved July 29, 2026.
  6. 1 2 Bercovich, Akhiad; Levy, Itay; Golan, Izik; Dabbah, Mohammad; et al. (May 2, 2025). "Llama-Nemotron: Efficient Reasoning Models". arXiv:2505.00949 [cs.CL].
  7. 1 2 Knight, Will (March 11, 2026). "Nvidia Will Spend $26 Billion to Build Open-Weight AI Models, Filings Show". Wired. Retrieved July 29, 2026.
  8. 1 2 "NVIDIA Nemotron 3 Ultra". Nvidia Research. Nvidia. June 4, 2026. Retrieved July 29, 2026.
  9. "Introducing Nemotron 3 Super: An Open Hybrid Mamba-Transformer MoE for Agentic Reasoning". Nvidia Technical Blog. Nvidia. March 11, 2026. Retrieved July 31, 2026.
  10. 1 2 "NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model". Nvidia Technical Blog. Nvidia. April 28, 2026. Retrieved July 29, 2026.
  11. Sirodot, Bertrand; Bronzati, Fabricio (November 7, 2025). "Introduction to NVIDIA Inference Microservices, aka NIM". Dell Technologies Info Hub. Dell Technologies. Retrieved July 29, 2026.
  12. "NVIDIA Nemotron". Nvidia. Nvidia. Retrieved July 29, 2026.
[edit]