AI Benchmark Model

DeepSeek V3.1 Terminus (Reasoning)

Benchmark scores, pricing, speed, and model comparisons for DeepSeek V3.1 Terminus (Reasoning).

Key scores

Intelligence14.8Index score
Coding43.5Index score
Math89.7Index score
Blended Cost$1.91Per 1M tokens

Review benchmark scores, pricing, performance data, and generated comparisons for this AI model.

Benchmark results

10 tracked benchmark rows for DeepSeek V3.1 Terminus (Reasoning).

AIME 2025 Math
89.7%
MMLU Pro Knowledge
85.1%
LiveCodeBench Code
79.8%
GPQA Reasoning
79.2%
LCR Other
69.3%
IFBench Other
57.0%
SciCode Other
38.0%
TAU2 Other
37.1%
TerminalBench Hard Other
30.3%
HLE Other
16.4%