Llama 4 Maverick 17B FP8
Meta · 1 configuration
Intelligence—
Input / 1M$0.20
Output / 1M$0.80
Context1M
Tariff: Deep Infra · Source: models-dev · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
◌
No pair of measurements available
Change the filters or pick another benchmark. Missing values are not estimated.
Selected: thinking default · Click a point to switch configuration
Open multimodal Llama model for strong reasoning and fast responses
This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.
Specifications from the catalog
Inputtext, image
Outputtext
Tool callingNo
Structured outputYes
SpeedNot available
First tokenNot available
Maximum output16.4K
Catalog listing date2025-04-05