Grok 4.5 vs Hy3: Specs & Benchmark Comparison
Grok 4.5 is developed by xAI, while Hy3 comes from Tencent. Both were released in July 2026. Hy3 has a published size of about 295 billion parameters; xAI has not disclosed the parameter count of Grok 4.5.
The two models share 4 published benchmarks. Grok 4.5 leads on 4 of them. The widest gaps are on DeepSWE, where Grok 4.5 scores 53.0% against 28.0%; Terminal-Bench 2.1, where Grok 4.5 scores 83.0% against 71.7%. Averaged across everything we track, Grok 4.5 sits at 52.9% and Hy3 at 57.1%.
Grok 4.5 is available through an API at $2 per million input tokens with a 500K-token context window. We do not track a hosted API for Hy3, so pricing and context cannot be compared directly.
| Characteristic | Grok 4.5 | Hy3 |
|---|---|---|
| Company | xAI | Tencent |
| Release Date | July 16, 2026 | July 6, 2026 |
| Parameters | — | 295B |
| Multimodal | Yes | No |
| Context (input) | 500K | — |
| Input Price / 1M | $2.00 | — |
| Output Price / 1M | $6.00 | — |
| Average Score | 52.9% | 57.1% |
| Benchmarks | ||
| DeepSWE | 53.0% | 28.0% |
| Terminal-Bench 2.1 | 83.0% | 71.7% |
| SWE-Bench Pro | 65.0% | 57.9% |
| GPQA | 93.0% | 90.4% |
Visual Benchmark Comparison
Verdict
Both models show equal results — the choice depends on your specific use case.
Both models show comparable average scores: Grok 4.5 — 0.5, Hy3 — 0.6.
Both models were released around the same time: 7/16/2026 and 7/6/2026.
More About These Models
Related Comparisons
Frequently Asked Questions
Which is better for coding — Grok 4.5 or Hy3?
Which model is cheaper — Grok 4.5 or Hy3?
Which has a larger context window — Grok 4.5 or Hy3?
The Grok 4.5 and Hy3 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Grok 4.5 or Hy3 page. See also the complete list of AI model comparisons.