Claude 3.5 Sonnet vs DeepSeek-V3.1: Specs & Benchmark Comparison

CharacteristicClaude 3.5 SonnetDeepSeek-V3.1
CompanyAnthropicDeepSeek
Release DateJune 21, 2024January 9, 2025
Parameters671B
MultimodalYesNo
Context (input)200K164K
Context (output)200K164K
Input Price / 1M$3.00$0.27
Output Price / 1M$15.00$1.00
Average Score0.80.8
Benchmarks
GPQA0.60.8
MMLU-Pro0.80.8

Visual Benchmark Comparison

Claude 3.5 Sonnet
DeepSeek-V3.1
GPQA0.6 vs 0.8
0.6
0.8
MMLU-Pro0.8 vs 0.8
0.8
0.8

Verdict

DeepSeek-V3.1 leads in 2 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Claude 3.5 Sonnet — 0.8, DeepSeek-V3.1 — 0.8.

API Cost

DeepSeek-V3.1 is 14.2x cheaper: input $0.27/1M vs $3.00/1M tokens.

Context Window

Claude 3.5 Sonnet supports a larger context: 200K vs 164K tokens.

Recency

DeepSeek-V3.1 is newer: released 1/9/2025 vs 6/21/2024.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Claude 3.5 Sonnet or DeepSeek-V3.1?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Claude 3.5 Sonnet or DeepSeek-V3.1?
DeepSeek-V3.1 is cheaper for input: $0.27 per 1M tokens vs $3.00.
Which has a larger context window — Claude 3.5 Sonnet or DeepSeek-V3.1?
Claude 3.5 Sonnet supports a larger context: 200,000 tokens vs 163,840.

The Claude 3.5 Sonnet and DeepSeek-V3.1 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Claude 3.5 Sonnet or DeepSeek-V3.1 page. See also the complete list of AI model comparisons.