Back
All NVIDIA models
LLMS • CURATED • UPDATED JUN 4, 2026

NVIDIA Nemotron 3 Super

120B open-weight hybrid MoE for efficient multi-agent reasoning

A 120B total / 12B active parameter hybrid Mamba-Transformer MoE language model with LatentMoE, multi-token prediction, and native NVFP4 pretraining. Optimized for complex multi-agent applications with a 1M-token context window and up to 5× higher throughput than the previous Nemotron Super.

Pricing Free
Platforms api, local
API Yes
Open Source Yes
Modalities LLMs
Best For Best for Multi-Agent Efficiency
Date Added 2026-06-04

Fills the gap between Nano and Ultra with a strong efficiency-to-accuracy ratio for agentic orchestration and latency-sensitive serving.

Claude Opus 4.6 Claude Opus 5 NotebookLM GPT-5.6 Sol Kimi K3
Free Completely free
📚

Uncensored AI Models: What Abliteration Actually Does To Them

Abliteration removes a model's ability to refuse by deleting one direction from its activations. It ...

Which LLMs Actually Deliver in 2026?

Comprehensive comparison of the best large language models in 2026 including ChatGPT, Claude, Gemini...

Large Language Models Explained: How They Actually Work

Complete guide to Large Language Models (LLMs). Learn how LLMs work, their architecture, training pr...

Choosing the Right LLM: Decision Framework That Actually Works

Complete guide to choosing the right large language model for your needs. Compare ChatGPT, Claude, G...

LLM Mastery: Practical Techniques That Get Results

Complete guide to using large language models effectively. Learn prompt engineering, API integration...

View NVIDIA Nemotron 3 Super Alternatives (2026) →

Compare NVIDIA Nemotron 3 Super with 5+ similar llms AI tools.

Q

Is NVIDIA Nemotron 3 Super free?

A

Yes, NVIDIA Nemotron 3 Super is completely free to use.

Q

Does NVIDIA Nemotron 3 Super have an API?

A

Yes, NVIDIA Nemotron 3 Super offers an API for programmatic integration.

Q

What is NVIDIA Nemotron 3 Super best for?

A

NVIDIA Nemotron 3 Super is best for Best for Multi-Agent Efficiency. A 120B total / 12B active parameter hybrid Mamba-Transformer MoE language model with LatentMoE, multi-token prediction, and native NVFP4 pretraining. Fills the gap between Nano and Ultra with a strong efficiency-to-accuracy ratio for agentic orchestration and latency-sensitive serving.

Q

What platforms does NVIDIA Nemotron 3 Super support?

A

NVIDIA Nemotron 3 Super supports api, local.

Q

Is NVIDIA Nemotron 3 Super open source?

A

Yes, NVIDIA Nemotron 3 Super is open source. You can access the source code, contribute, and deploy it on your own infrastructure.

Q

How do I use NVIDIA Nemotron 3 Super?

A

NVIDIA Nemotron 3 Super is a large language model for text generation, analysis, and conversation. Use the API for programmatic access. Enter prompts or questions to get responses.

🏷️

Work on NVIDIA Nemotron 3 Super? You're hand-reviewed in our directory. Add this badge to your site — it links back to this profile.

Featured on CuratedAI