Nous Research logo

Hermes 3 70B

Nous Research

Hermes 3 70B is the flagship instruction-following model by Nous Research, fine-tuned for advanced reasoning, creative writing, and complex task execution. It excels at instruction following and delivers high performance across a wide range of domains.

Key Specifications

Parameters
70.0B
Context
-
Release Date
August 14, 2024
Average Score
68.4%

Timeline

Key dates in the model's history
Announcement
August 14, 2024
Last Update
January 29, 2026
Today
October 5, 2026

Technical Specifications

Parameters
70.0B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Fine-tuned from
llama-3.1-70b-instruct
Capabilities
MultimodalZeroEval

Benchmark Results

Model performance metrics across various tests and benchmarks

General Knowledge

Tests on general knowledge and understanding
HellaSwag
Accuracy • Self-reported
88.0%
Winogrande
Accuracy • Self-reported
83.0%
MMLU
Accuracy • Self-reported
79.1%
TruthfulQA
Self-reported by the model provider • Self-reported
63.3%

Mathematics

Mathematical problems and computations
MATH
Self-reported by the model provider • Self-reported
20.8%

Reasoning

Logical reasoning and analysis
GPQA
Self-reported by the model provider • Self-reported
66.1%

Other Tests

Specialized benchmarks
MT-Bench
Score 100 • Self-reported
89.9%
BoolQ
Accuracy • Self-reported
88.0%
PIQA
Accuracy • Self-reported
84.0%
ARC-E
Accuracy • Self-reported
83.0%
AGIEval
Self-reported by the model provider • Self-reported
56.2%
ARC-C
Self-reported by the model provider • Self-reported
65.5%
BBH
Self-reported by the model provider • Self-reported
67.8%
IFBench
Self-reported by the model provider • Self-reported
81.2%
MMLU-Pro
Self-reported by the model provider • Self-reported
47.2%
MuSR
Self-reported by the model provider • Self-reported
50.7%
OpenBookQA
Self-reported by the model provider • Self-reported
49.4%

License & Metadata

License
apache-2.0
Announcement Date
August 14, 2024
Last Updated
January 29, 2026

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.