All AI models · NVIDIA
NVIDIA model list
All 20 notable NVIDIA models tracked in 2026, out of 279 across every lab. Sorted newest first, with context length and capabilities. Data from the open models.dev database.
20
Models
12
Reasoning
5
Vision
20
Open weights
| Model | Context | Capabilities |
|---|---|---|
|
Nemotron 3 Ultra 550B A55B
nvidia/nemotron-3-ultra-550b-a55b
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
|
1000K | Reasoning Tools Open |
|
Nemotron 3.5 Content Safety
nvidia/nemotron-3.5-content-safety
Safety model for policy screening, moderation, and risk-aware routing workflows
|
128K | Reasoning Vision Open |
|
Nemotron 3 Nano Omni 30B A3B Reasoning
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Open Nemotron omni model combining reasoning with text, vision, and audio
|
256K | Reasoning Tools Vision Open |
|
Nemotron 3 Content Safety
nvidia/nemotron-3-content-safety
Safety model for policy screening, moderation, and risk-aware routing workflows
|
128K | Open |
|
Llama Nemotron Rerank VL 1B v2
nvidia/llama-nemotron-rerank-vl-1b-v2
Reranking model for improving retrieval quality in search and recommendation systems
|
128K | Vision Open |
|
Nemotron Cascade 2 30B A3B
nvidia/nemotron-cascade-2-30b-a3b
Nemotron model for efficient reasoning, coding, and specialized AI agents
|
256K | Reasoning Tools Open |
|
Nemotron VoiceChat
nvidia/nemotron-voicechat
Nemotron multimodal model for visual reasoning and agentic AI workflows
|
128K | Tools Open |
|
Nemotron 3 Super 120B A12B
nvidia/nemotron-3-super-120b-a12b
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
|
262K | Reasoning Tools Open |
|
Llama Nemotron Embed VL 1B v2
nvidia/llama-nemotron-embed-vl-1b-v2
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
|
32K | Vision Open |
|
Nemotron Content Safety Reasoning 4B
nvidia/nemotron-content-safety-reasoning-4b
Safety model for policy screening, moderation, and risk-aware routing workflows
|
128K | Reasoning Open |
|
Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
|
262K | Reasoning Tools Open |
|
Nemotron Nano 12B v2 VL
nvidia/nemotron-nano-12b-v2-vl
Nemotron multimodal model for visual reasoning and agentic AI workflows
|
128K | Reasoning Tools Vision Open |
|
Llama 3.1 Nemotron Safety Guard 8B v3
nvidia/llama-3.1-nemotron-safety-guard-8b-v3
Safety model for policy screening, moderation, and risk-aware routing workflows
|
128K | Open |
|
Nemotron Nano 9B v2
nvidia/nemotron-nano-9b-v2
Compact Nemotron model for efficient reasoning and deployable AI agents
|
131K | Reasoning Tools Open |
|
Llama 3.3 Nemotron Super 49B v1.5
nvidia/llama-3.3-nemotron-super-49b-v1.5
Nemotron model for efficient reasoning, coding, and specialized AI agents
|
131K | Reasoning Tools Open |
|
Mistral Nemotron
nvidia/mistral-nemotron
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
|
128K | Tools Open |
|
Llama 3.1 Nemotron 70B Instruct
nvidia/llama-3.1-nemotron-70b-instruct
Nemotron model for efficient reasoning, coding, and specialized AI agents
|
128K | Tools Open |
|
Llama 3.3 Nemotron Super 49B v1
nvidia/llama-3.3-nemotron-super-49b-v1
Nemotron model for efficient reasoning, coding, and specialized AI agents
|
131K | Reasoning Tools Open |
|
Llama 3.1 Nemotron Ultra 253B
nvidia/llama-3.1-nemotron-ultra-253b
Flagship Nemotron model for high-throughput reasoning and complex agents
|
128K | Reasoning Tools Open |
|
Nemotron Mini 4B Instruct
nvidia/nemotron-mini-4b-instruct
Compact Nemotron model for efficient reasoning and deployable AI agents
|
128K | Tools Open |
Which of these can I call for free?
This page lists what exists, not what is free. For the models you can call right now with one free key and no credit card, see the live API catalog or read the free LLM API guide. Open-weight models can also be run on your own GPU, and the VRAM calculator tells you whether yours is big enough.
Model lists by lab
OpenAI (56)
Google (43)
Alibaba (Qwen) (33)
Anthropic (23)
NVIDIA (20)
Mistral (18)
Zhipu AI (14)
Cohere (14)
MiniMax (7)
Moonshot AI (7)
xAI (6)
Xiaomi (6)
DeepSeek (5)
Meta (4)
Poolside (4)
DeepReinforce (4)
Perplexity (3)
StepFun (3)
Sakana AI (2)
Tencent (2)
Sarvam AI (2)
Microsoft (1)
Thinking Machines (1)
Meituan (1)
Model metadata sourced from the open models.dev database. See the full 279-model index for a filterable view across every lab.