AI Benchmark Model

Hermes 4 - Llama-3.1 405B (Reasoning)

Benchmark scores, pricing, speed, and model comparisons for Hermes 4 - Llama-3.1 405B (Reasoning).

Key scores

Intelligence7.5Index score
Math69.7Index score
Speed39.3Tokens/sec
Blended Cost$1.50Per 1M tokens

Review benchmark scores, pricing, performance data, and generated comparisons for this AI model.

Benchmark results

9 tracked benchmark rows for Hermes 4 - Llama-3.1 405B (Reasoning).

MMLU Pro Knowledge
82.9%
GPQA Reasoning
72.7%
AIME 2025 Math
69.7%
LiveCodeBench Code
68.6%
IFBench Other
32.7%
LCR Other
22.3%
TAU2 Other
22.2%
TerminalBench Hard Other
11.4%
HLE Other
10.9%