Gemini 3.1 Pro Preview
Google · 2 thinking configurations
Intelligence29.7
Input / 1M$2.00
Output / 1M$12.00
Context1M
Tariff: Google · Source: artificial-analysis · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
Scroll the chart ↔ · tap a point for details
Efficient frontier: best measured score at each cost1 measured configurations · 1 without both values
Selected: thinking thinking · Click a point to switch configuration
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
Intelligence IndexArtificial Analysis composite index, not a percentage. Different versions of the index are not directly comparable.
29.7 pointsGPQA DiamondGraduate-level scientific questions. Measures specialist reasoning rather than web-search quality.
94.1 %Humanity’s Last ExamDifficult questions across many disciplines, compared under the same evaluation setup.
47.0 %SciCodeCode generation for scientific research problems.
58.7 %Terminal-Bench 2.1Agent tasks in a terminal environment. Results also depend on the evaluation harness.
73.8 %τ²-BenchTool use and interaction in assistance scenarios. Does not replace an evaluation of your integration.
95.6 %IFBenchVerifiable instruction following. A proxy for precision rather than a benchmark of creative style.
77.1 %MMMU-ProMultimodal understanding of academic problems that include images.
82.4 %Omniscience accuracyFactual accuracy in Artificial Analysis's Omniscience evaluation.
54.9 %Specifications from the catalog
Inputaudio, file, image, text, video
Outputtext
Tool callingYes
Structured outputYes
Speed113.9 tok/s
First token27.84 s
Maximum output65.5K
Catalog listing date2026-02-19