Find a model

Search models and open their dedicated profile. Use the arrow keys to navigate and Enter to open.

nemotron-3.5-lightning-30b-a3b

NVIDIA · 1 configuration

Intelligence—
Input / 1M—
Output / 1M—
Context1M

Tariff: Requesty · Source: models-dev · USD per million tokens

Thinking configuration

What does more thinking buy you?

Compare this model’s measured thinking configurations.

◌

No pair of measurements available

Change the filters or pick another benchmark. Missing values are not estimated.

Selected: thinking default · Click a point to switch configuration

NVIDIA Nemotron 3.5 Lightning 30B-A3B is a hybrid Mamba-2 + MoE + Attention model with 30B total and 3B active parameters, pre-trained on over 20T tokens with an NVFP4 recipe and Multi-Token Prediction for fast generation. Up to 1M token context for long-running autonomous agents, sub-agent workhorse deployments, and agentic workflows. Supports reasoning and tool calling. English and coding languages plus Spanish, French, German, Italian, and Japanese. Open weights under the OpenMDW License Agreement v1.1. Part of the NVIDIA Nemotron family.

This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.

Specifications from the catalog

Inputtext
Outputtext
Tool callingYes
Structured outputNo
SpeedNot available
First tokenNot available
Maximum output65.5K
Catalog listing date2026-08-11