DeepSeek V3.1 Terminus
DeepSeek · 2 thinking configurations
Intelligence13.9
Input / 1M$0.27
Output / 1M$1.00
Context128K
Tariff: DeepSeek · Source: artificial-analysis · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
Scroll the chart ↔ · tap a point for details
Efficient frontier: best measured score at each cost2 measured configurations · 0 without both values
Selected: thinking none · Click a point to switch configuration
DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
Intelligence IndexArtificial Analysis composite index, not a percentage. Different versions of the index are not directly comparable.
13.9 pointsGPQA DiamondGraduate-level scientific questions. Measures specialist reasoning rather than web-search quality.
75.1 %Humanity’s Last ExamDifficult questions across many disciplines, compared under the same evaluation setup.
8.7 %τ²-BenchTool use and interaction in assistance scenarios. Does not replace an evaluation of your integration.
37.1 %IFBenchVerifiable instruction following. A proxy for precision rather than a benchmark of creative style.
41.2 %Omniscience accuracyFactual accuracy in Artificial Analysis's Omniscience evaluation.
23.5 %Specifications from the catalog
Inputtext
Outputtext
Tool callingYes
Structured outputYes
SpeedNot available
First tokenNot available
Maximum output65.5K
Catalog listing date2025-09-22