Compare / head-to-head

Hy3vsHy3 preview

Hy3 leads 5 of 5 shared benchmarks. Both offer a 256K-token context window.

Benchmarks from cited public sources; pricing from official pages; status from official provider feeds.

Shared benchmarks
5 – 0
Hy3 leads
Cheaper per token
—
list price, input + output
Larger context
Tie
both 256K tokens
Providers
Tencent Hunyuan
same provider
Hy3
Tencent Hunyuan · released 2026-07-06
text
Context
256K
Max out
—
Input /1M
$0.132 ↗
Output /1M
$0.528 ↗
Cached /1M
$0.033
Scores
45 · 10 core
Hy3 preview
Tencent Hunyuan · released 2026-04-23
text
Context
256K
Max out
—
Input /1M
—
Output /1M
—
Cached /1M
—
Scores
15 · 3 core
Quality

Benchmark matrix

BenchmarkHy3Hy3 previewΔEdge
Reported by both · 5
SWE-bench Verified78% ↗74.4% ↗+3.6 ptHy3
BrowseComp84.2% ↗67.1% ↗+17.1 ptHy3
CL-bench23.8% ↗22.8% ↗+1 ptHy3
CL-bench Life17% ↗15.7% ↗+1.3 ptHy3
WideSearch76.4% ↗70.2% ↗+6.2 ptHy3
Only Hy3 reports · 39
GPQA Diamond90.4% ↗not reported——
Humanity's Last Exam (no tools)37% ↗not reported——
LMArena Elo1440.6 ↗not reported——
AA-LCR73.4% ↗not reported——
Apex-Agent25.6% ↗not reported——
ArxivMath52.2% ↗not reported——
ClawEval68.5% ↗not reported——
CMT-Benchmark37.8% ↗not reported——
DeepSearchQA91% ↗not reported——
DeepSWE28% ↗not reported——
e-bench50.2% ↗not reported——
FrontierScience-Olympiad74.8% ↗not reported——
FrontierScience-Research21.3% ↗not reported——
HLE (with tools, text-only)53.2% ↗not reported——
HorizonMath7.1% ↗not reported——
Hy-Backend 2.025% ↗not reported——
Hy-CompanyBench41.7% ↗not reported——
Hy-Euler Pro24.2% ↗not reported——
Hy-FinModelBench69% ↗not reported——
Hy-Math60.9% ↗not reported——
Hy-SkillsWorld45.8% ↗not reported——
Hy-SWE Max49% ↗not reported——
IMOAnswerBench90% ↗not reported——
LMArena Agent (LMArena)-0.0538 ↗not reported——
LMArena WebDev (LMArena)1507.4 ↗not reported——
MathArena Apex38.7% ↗not reported——
MCP Atlas79.1% ↗not reported——
NL2Repo45.6% ↗not reported——
PHYBench77.4% ↗not reported——
ProdBench23% ↗not reported——
SkillsBench55.3% ↗not reported——
SuperChem54.9% ↗not reported——
SWE-bench Multilingual75.8% ↗not reported——
SWE-bench Pro57.9% ↗not reported——
Terminal-Bench 2.171.7% ↗not reported——
Toolathlon48.5% ↗not reported——
Toolathlon-Verified (HKUST)56.8% ↗not reported——
USAMO 202672% ↗not reported——
WildClawBench53.6% ↗not reported——
Only Hy3 preview reports · 10
AdvancedIF (CIF&CC subsets)not reported79.5% ↗——
CHSBO 2025 (text-only, pass^3)not reported87.8% ↗——
ClawEval (pass^3)not reported55% ↗——
FrontierScience Olympiadnot reported70% ↗——
GPQA-Diamondnot reported87.2% ↗——
HLE (text-only)not reported30% ↗——
IMO Answer Benchnot reported84.3% ↗——
LongBench v2not reported65.4% ↗——
Terminal-Bench 2.0not reported54.4% ↗——
WildClawBench (text-only)not reported45.3% ↗——
Scores tagged "epoch" or "matharena" are independent runs, used only where the lab has not published its own; ⚠ marks rows MathArena flags as released after the competition. Higher is better on every row. Δ is Hy3 minus Hy3 preview in the benchmark's own unit. "Not reported" means the lab has not published that figure; it is not a zero. ↗ opens the source.
Specs & pricing

Side by side

SpecHy3Hy3 previewEdge
Context window256K tokens256K tokensTie
Max output———
Input price / 1M$0.132 ↗——
Output price / 1M$0.528 ↗——
Cached input / 1M$0.033 ↗——
Input + output / 1M
Lower is cheaper. List prices; batch, tool and regional fees excluded.
$0.66——
ModalitiestexttextTie
Released2026-07-062026-04-23—
Cited benchmark scores4515—
Reliability

Tencent Hunyuan status

All providers →
More matchups

Hy3 preview vs …

Built by Respan
Which one wins on your data?

Public benchmarks are a starting point. Run Hy3 and Hy3 preview on your own prompts with Respan evals, or route to either through one gateway key with automatic failover.