nemotron-nano-v2-12b
NVIDIA · 1 configuration
Intelligence—
Input / 1M$0.24
Output / 1M$0.71
Context128K
Tariff: Cortecs · Source: models-dev · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
◌
No pair of measurements available
Change the filters or pick another benchmark. Missing values are not estimated.
Selected: thinking default · Click a point to switch configuration
NVIDIA Nemotron Nano v2 12B is a 12-billion-parameter multimodal reasoning model designed for advanced video understanding, document intelligence, and visual reasoning, built with a hybrid Transformer-Mamba architecture for high efficiency and low latency.
This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.
Specifications from the catalog
Inputtext, image
Outputtext
Tool callingYes
Structured outputYes
SpeedNot available
First tokenNot available
Maximum output128K
Catalog listing date2025-10-31