Claude Opus 4.7
MultimodalClaude Opus 4.7 is Anthropic's latest Opus-class model, a direct upgrade to Opus 4.6 with notable improvements in advanced software engineering, particularly on the most difficult tasks. It handles complex, long-running agentic workflows with rigor and consistency, follows instructions more literally and precisely, and verifies its own outputs before reporting back.
Key Specifications
Parameters
-
Context
1.0M
Release Date
April 16, 2026
Average Score
68.3%
Timeline
Key dates in the model's history
Announcement
April 16, 2026
Last Update
August 27, 2026
Technical Specifications
Parameters
-
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$5.00
Output (per 1M tokens)
$25.00
Max Input Tokens
1.0M
Max Output Tokens
128.0K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
Programming
Programming skills tests
SWE-Bench Verified
Agentic coding evaluation. Memorization screens flag a subset of problems; excluding those, Opus 4.7's margin over Opus 4.6 holds. • Self-reported
Reasoning
Logical reasoning and analysis
GPQA
GPQA Diamond. Graduate-level reasoning evaluation. • Self-reported
Other Tests
Specialized benchmarks
BrowseComp
Agentic search evaluation. • Self-reported
CharXiv-R
CharXiv Reasoning. Visual reasoning with tools: 91.0%. Without tools: 82.1%. • Self-reported
CyberGym
Cybersecurity vulnerability reproduction. Opus 4.6's score was updated from the originally reported 66.6 to 73.8 after harness parameter updates to better elicit cyber capability. • Self-reported
Finance Agent
Agentic financial analysis evaluation (Finance Agent v1.1). State-of-the-art score at release. • Self-reported
Finance Agent v2
• Self-reported
FrontierCode 1.1
FrontierCode 1.1 current leaderboard; mergeability score at max effort: 38.5%. • Self-reported
FrontierSWE
Claude Code • Self-reported
Humanity's Last Exam
Multidisciplinary reasoning. With tools: 54.7%. Without tools: 46.9%. • Self-reported
Legal Agent Benchmark
All-pass rate (Harvey held-out set) • Self-reported
LiveBench
2026-01-08, Thinking xHigh Effort • Self-reported
MCP Atlas
Scaled tool use evaluation. Opus 4.6's MCP-Atlas score was updated to reflect revised grading methodology from Scale AI. • Self-reported
MMMLU
Multilingual Q&A. • Self-reported
OSWorld-Verified
Agentic computer use evaluation. • Self-reported
SWE-Bench Pro
Agentic coding evaluation. • Self-reported
Terminal-Bench 2.0
Terminus-2 harness with thinking disabled. 1× guaranteed / 3× ceiling resource allocation averaged over five attempts per task. • Self-reported
License & Metadata
License
proprietary
Announcement Date
April 16, 2026
Last Updated
August 27, 2026
Similar Models
All ModelsClaude Opus 4.6
Anthropic
MM
Best score:1.0 (TAU)
Released:Feb 2026
Price:$5.00/1M tokens
Claude Sonnet 4.6
Anthropic
MM
Best score:0.9 (GPQA)
Released:Feb 2026
Price:$3.00/1M tokens
Claude Opus 4.8
Anthropic
MM
Best score:0.9 (GPQA)
Released:May 2026
Price:$5.00/1M tokens
Claude Opus 4.5
Anthropic
MM
Best score:0.9 (TAU)
Released:Nov 2025
Price:$5.00/1M tokens
Claude Sonnet 4.5
Anthropic
MM
Best score:0.9 (TAU)
Released:Sep 2025
Price:$3.00/1M tokens
Claude Sonnet 4
Anthropic
MM
Best score:0.8 (GPQA)
Released:May 2025
Price:$3.00/1M tokens
Claude Sonnet 5
Anthropic
MM
Released:Jun 2026
Price:$2.00/1M tokens
Claude Opus 5
Anthropic
MM
Released:Jul 2026
Price:$5.00/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.