Key Specifications
Parameters
1.0T
Context
1.0M
Release Date
April 27, 2026
Average Score
71.8%
Timeline
Key dates in the model's history
Announcement
April 27, 2026
Last Update
August 27, 2026
Today
September 11, 2026
Technical Specifications
Parameters
1.0T
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$0.43
Output (per 1M tokens)
$0.87
Max Input Tokens
1.0M
Max Output Tokens
131.1K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
General Knowledge
Tests on general knowledge and understanding
HellaSwag
Base model, 10-shot • Self-reported
MMLU
Base model, 5-shot • Self-reported
Winogrande
Base model, 5-shot • Self-reported
Programming
Programming skills tests
SWE-Bench Verified
• Self-reported
Mathematics
Mathematical problems and computations
GSM8k
Base model, 8-shot • Self-reported
MATH
Base model, 4-shot • Self-reported
Reasoning
Logical reasoning and analysis
DROP
Base model, 3-shot • Self-reported
GPQA
Base model, 5-shot • Self-reported
Other Tests
Specialized benchmarks
AIME
Base model, AIME 24&25, 2-shot • Self-reported
ARC-C
Base model, 25-shot • Self-reported
BBH
Base model, 3-shot • Self-reported
C-Eval
Base model, 5-shot • Self-reported
Claw-Eval
Hugging Face evaluation result, General • Self-reported
CMMLU
Base model, 5-shot • Self-reported
Finance Agent v2
• Verified
Global-MMLU
Base model, 5-shot • Self-reported
GraphWalks
Parents 1M F1 • Self-reported
HumanEval+
Base model, 1-shot • Self-reported
Humanity's Last Exam
No tools • Self-reported
LiveCodeBench v6
Base model, 1-shot • Self-reported
MBPP+
Base model, 3-shot • Self-reported
MiMo Coding Bench
• Self-reported
MMLU-Pro
Base model, 5-shot • Self-reported
MMLU-Redux
Base model, 5-shot • Self-reported
SWE-Bench Pro
• Self-reported
SWE-bench Verified (Agentless)
Base model, 3-shot • Self-reported
TAU3-Bench
• Self-reported
Terminal-Bench 2.0
• Self-reported
TriviaQA
Base model, 5-shot • Self-reported
WildClawBench
Hugging Face evaluation result, Overall • Self-reported
License & Metadata
License
mit
Announcement Date
April 27, 2026
Last Updated
August 27, 2026
Similar Models
All ModelsMiMo-V2-Flash
Xiaomi
309.0B
Best score:0.8 (GPQA)
Released:Dec 2025
GLM-5
Zhipu AI
744.0B
Best score:0.9 (TAU)
Released:Feb 2026
Kimi K2-Thinking-0905
Moonshot AI
1.0T
Best score:0.8 (GPQA)
Released:Sep 2025
Price:$0.60/1M tokens
GLM-5.1
Zhipu AI
754.0B
Best score:0.9 (GPQA)
Released:Apr 2026
Price:$1.40/1M tokens
GLM-4.7
Zhipu AI
358.0B
Best score:0.9 (TAU)
Released:Dec 2025
Price:$0.60/1M tokens
Kimi K2 Instruct
Moonshot AI
1.0T
Best score:0.9 (HumanEval)
Released:Jan 2025
Price:$0.57/1M tokens
Nemotron 3 Ultra (550B A55B)
NVIDIA
550.0B
Best score:0.9 (GPQA)
Released:Jun 2026
Kimi K2-Instruct-0905
Moonshot AI
1.0T
Best score:0.9 (MMLU)
Released:Sep 2025
Price:$0.60/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.