DeepSeek V3 0324
DeepSeek · 1 configuration
Intelligence9.7
Input / 1M$0.24
Output / 1M$0.90
Context128K
Tariff: DeepSeek · Source: artificial-analysis · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
Scroll the chart ↔ · tap a point for details
Efficient frontier: best measured score at each cost1 measured configurations · 0 without both values
Selected: thinking default · Click a point to switch configuration
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the DeepSeek V3 model and performs really well...
Intelligence IndexArtificial Analysis composite index, not a percentage. Different versions of the index are not directly comparable.
9.7 pointsGPQA DiamondGraduate-level scientific questions. Measures specialist reasoning rather than web-search quality.
65.5 %Humanity’s Last ExamDifficult questions across many disciplines, compared under the same evaluation setup.
4.7 %SciCodeCode generation for scientific research problems.
39.0 %Terminal-Bench 2.1Agent tasks in a terminal environment. Results also depend on the evaluation harness.
13.9 %τ²-BenchTool use and interaction in assistance scenarios. Does not replace an evaluation of your integration.
47.1 %IFBenchVerifiable instruction following. A proxy for precision rather than a benchmark of creative style.
41.0 %Omniscience accuracyFactual accuracy in Artificial Analysis's Omniscience evaluation.
24.3 %Specifications from the catalog
Inputtext
Outputtext
Tool callingYes
Structured outputYes
SpeedNot available
First tokenNot available
Maximum output147.5K
Catalog listing date2025-03-24