AI Benchmark Model
Hermes 4 - Llama-3.1 405B (Reasoning)
Benchmark scores, pricing, speed, and model comparisons for Hermes 4 - Llama-3.1 405B (Reasoning).
Key scores
Review benchmark scores, pricing, performance data, and generated comparisons for this AI model.
Benchmark results
9 tracked benchmark rows for Hermes 4 - Llama-3.1 405B (Reasoning).
MMLU Pro Knowledge
82.9% GPQA Reasoning
72.7% AIME 2025 Math
69.7% LiveCodeBench Code
68.6% IFBench Other
32.7% LCR Other
22.3% TAU2 Other
22.2% TerminalBench Hard Other
11.4% HLE Other
10.9%