NVIDIA Nemotron
High-efficiency, multimodal, open models for long-running AI agents.
Overview & Benefits
Open Models Built for Reasoning Workloads
NVIDIA Nemotron is an open family of reasoning models, optimized for accuracy and compute efficiency across every stage of an agentic workflow — plus multimodal capabilities for vision, speech, retrieval, and safety.
Nemotron Nano
Superior accuracy and efficiency for specialized sub-agents that run cost-effectively at the edge or at scale.
Nemotron Super
Highest accuracy, high-throughput reasoning, and tool calling for production multi-agent systems.
Nemotron Ultra
Best-in-class reasoning for mission-critical, multi-step workflows that demand deep planning.
Visual Understanding
Reason across video, audio, image, and text to power multimodal agents and document understanding.
Speech
Ultra-low-latency ASR, TTS, and neural machine translation for real-time voice experiences.
Retrieval-Augmented Generation
Nemotron Retriever delivers high-accuracy embeddings and reranking to ground answers in your data.
Safety
Real-time content safety, jailbreak protection, and multilingual guardrails for trusted deployments.
Models
Choose the Right Nemotron Model
Every Nemotron model is available to download and deploy from build.nvidia.com — pick the tier that matches your accuracy, latency, and cost targets.
Nemotron Nano
Compact, efficient reasoning for sub-agents, edge deployments, and high-volume pipelines.
View on build.nvidia.comNemotron Super
Balanced accuracy and throughput with strong tool calling for multi-agent production systems.
View on build.nvidia.comNemotron Ultra
Maximum reasoning depth for complex, mission-critical, multi-step workflows.
View on build.nvidia.comTechnology
Building Blocks for Agentic AI
Combine Nemotron models with the NVIDIA software stack to build, optimize, and deploy AI agents in production.
NVIDIA NeMo
An end-to-end platform to build, customize, and deploy generative AI models with your own data.
Learn moreNVIDIA NIM
Optimized inference microservices for fast, secure deployment of Nemotron models anywhere.
Learn moreNVIDIA Blueprints
Reference workflows that show how to assemble Nemotron models into real agentic applications.
Learn moreGet Started
Ways to Get Started
Start Prototyping for Free
Explore and test Nemotron models instantly in your browser with the NVIDIA API catalog.
Start prototypingRun on Inference Providers
Deploy Nemotron across leading inference providers and your own infrastructure.
View providersAdopters
Enterprises Using Nemotron
Leading organizations build production AI agents and copilots on NVIDIA Nemotron.
FAQs
Frequently Asked Questions
NVIDIA Nemotron models are truly open source. NVIDIA publishes the training datasets, the techniques used to build them, and the model weights so the community can fully understand and reproduce them.
Ready to Get Started?
Build, customize, and deploy open NVIDIA Nemotron models for your next AI agent.
Try Nemotron live in the chat above — no signup required.