← CookbookView source on GitHub ↗

NVIDIA Nemotron Models

Browse NVIDIA Nemotron models available on Nebius Token Factory.


Available Models

Model Provider Parameters Context Highlights
Nemotron-3-Nano-30B-A3B

▶ Try it @ TF
NVIDIA 30B total
3B active
262 K Compact MoE model optimized for efficient reasoning, chat, and coding with strong multilingual support and long-context RAG/agent workflows
Nemotron-3-Nano-Omni

▶ Try it @ TF
NVIDIA 30 B total
3 B active
262 K The most open, efficient, and accurate omni-modal reasoning model for agentic AI
Nemotron-3-Super-120B-A12B

▶ Try it @ TF
NVIDIA 120B total
12B active
256 K hybrid MoE model optimized for efficient multi-agent AI and complex reasoning tasks.
Nemotron-3-Ultra-550B-A55B

▶ Try it @ TF
NVIDIA 550B total
55B active
256 K Flagship hybrid MoE model optimized for the most demanding multi-agent AI and complex reasoning tasks
Llama-3.1-Nemotron-Ultra-253B-v1

▶ Try it @ TF
NVIDIA / Meta 253 B 128 K NVIDIA-tuned Llama variant built for high-efficiency reasoning, safety, and enterprise-grade performance