Find a model

Search models and open their dedicated profile. Use the arrow keys to navigate and Enter to open.

nvidia-nemotron-3-super-120b-a12b

Unknown · 1 configuration

Intelligence—
Input / 1M$0.10
Output / 1M$0.50
Context262.1K

Tariff: Requesty · Source: models-dev · USD per million tokens

Thinking configuration

What does more thinking buy you?

Compare this model’s measured thinking configurations.

◌

No pair of measurements available

Change the filters or pick another benchmark. Missing values are not estimated.

Selected: thinking default · Click a point to switch configuration

NVIDIA Nemotron 3 Super is a hybrid Mixture-of-Experts (MoE) model engineered for highest compute efficiency and accuracy in multi-agent applications and specialized agentic systems. It is optimized to run many collaborating agents per application on a single GPU, delivering high accuracy for reasoning, tool use, and instruction following.

This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.

Specifications from the catalog

Inputtext
Outputtext
Tool callingYes
Structured outputYes
SpeedNot available
First tokenNot available
Maximum output262.1K
Catalog listing date2026-03-11