AI Benchmark Model

Hermes 4 - Llama-3.1 70B (Reasoning)

Benchmark scores, pricing, speed, and model comparisons for Hermes 4 - Llama-3.1 70B (Reasoning).

Key scores

Intelligence7.9Index score
Math68.7Index score

Review benchmark scores, pricing, performance data, and generated comparisons for this AI model.

Benchmark results

9 tracked benchmark rows for Hermes 4 - Llama-3.1 70B (Reasoning).

MMLU Pro Knowledge
81.1%
GPQA Reasoning
69.9%
AIME 2025 Math
68.7%
LiveCodeBench Code
65.3%
IFBench Other
31.3%
TAU2 Other
22.5%
LCR Other
9.7%
HLE Other
8.8%
TerminalBench Hard Other
4.5%