Provider logo

Hermes 4 Large

nousresearch/hermes-4-405b
Provider logo

Hermes 4 Large

nousresearch/hermes-4-405b

Advanced reasoning model built on Llama-3.1-405B with hybrid thinking modes. Features internal deliberation capabilities, excels at math, code, STEM, and logical reasoning while supporting structured outputs with improved steerability and neutral alignment.

Added Aug 26, 2025

Model weights

Context Window

128.0K

Max Output

8.2K

Input Price (Auto)

$0.30/1M

Output Price (Auto)

$1.20/1M

Cache Read (Auto)

$0.15/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

8.8

Better than 29% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

53.6%

Better than 33% of models compared

HLE

Humanity's Last Exam

4.2%

Better than 17% of models compared

IFBench

Instruction-following benchmark

34.8%

Better than 27% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

26.6%

Better than 32% of models compared

AA-LCR

Long context reasoning evaluation

20.0%

Better than 34% of models compared

Coding

SciCode

Python programming for scientific computing

34.6%

Better than 57% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

9.8%

Better than 45% of models compared

LiveCodeBench

Contamination-free coding benchmark

54.6%

Better than 62% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

15.3%

Better than 19% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

72.9%

Better than 42% of models compared

Last updated Jun 28, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Hermes 4 Large with similar models from the same provider or model family.