hy3 vs grok-4.20-beta-0309-reasoning Benchmark Comparison
Direct benchmark comparison between hy3 and grok-4.20-beta-0309-reasoning based on LMArena Elo and the latest 2026 API pricing.
Direct Technical & Pricing Comparison
*These models represent the Pareto Frontier (optimal cost-to-performance).*
Comparison Summary: hy3 is the more capable model in this pair, leading by 54 Elo points. Specifically, hy3 is currently on the Pareto line, suggesting it offers better systemic value for its performance tier.