NVIDIA Nemotron Models
Browse NVIDIA Nemotron models available on Nebius Token Factory.
Available Models
| Model | Provider | Parameters | Context | Highlights |
|---|---|---|---|---|
| Nemotron-3-Nano-30B-A3B ▶ Try it @ TF |
NVIDIA | 30B total 3B active |
262 K | Compact MoE model optimized for efficient reasoning, chat, and coding with strong multilingual support and long-context RAG/agent workflows |
| Nemotron-3-Nano-Omni ▶ Try it @ TF |
NVIDIA | 30 B total 3 B active |
262 K | The most open, efficient, and accurate omni-modal reasoning model for agentic AI |
| Nemotron-3-Super-120B-A12B ▶ Try it @ TF |
NVIDIA | 120B total 12B active |
256 K | hybrid MoE model optimized for efficient multi-agent AI and complex reasoning tasks. |
| Nemotron-3-Ultra-550B-A55B ▶ Try it @ TF |
NVIDIA | 550B total 55B active |
256 K | Flagship hybrid MoE model optimized for the most demanding multi-agent AI and complex reasoning tasks |
| Llama-3.1-Nemotron-Ultra-253B-v1 ▶ Try it @ TF |
NVIDIA / Meta | 253 B | 128 K | NVIDIA-tuned Llama variant built for high-efficiency reasoning, safety, and enterprise-grade performance |