AI Benchmark Model

Hermes 4 - Llama-3.1 70B (Non-reasoning)

Benchmark scores, pricing, speed, and model comparisons for Hermes 4 - Llama-3.1 70B (Non-reasoning).

Key scores

Intelligence6.7Index score
Math11.3Index score

Review benchmark scores, pricing, performance data, and generated comparisons for this AI model.

Benchmark results

9 tracked benchmark rows for Hermes 4 - Llama-3.1 70B (Non-reasoning).

MMLU Pro Knowledge
66.4%
GPQA Reasoning
49.1%
IFBench Other
29.0%
LiveCodeBench Code
26.9%
TAU2 Other
21.6%
AIME 2025 Math
11.3%
LCR Other
4.7%
HLE Other
3.6%
TerminalBench Hard Other
0.0%