❯ Tencent open-sources Hy4 preview: 770B parameters, 1M context, self-scored just past GLM-5.3
model specsPer Tencent’s announcement, the company released and open-sourced Hunyuan Hy4 preview, a mixture-of-experts model with 770 billion total parameters and 49 billion active, and a context window above 1 million tokens. Weights are on Hugging Face, and the model is wired into Tencent’s own Yuanbao, ima and CodeBuddy, with API access via Tencent Cloud TokenHub and OpenRouter.
read the benchmark carefullyTencent says an internal blind evaluation had 163 experts score 203 engineering tasks, with Hy4 preview averaging 2.99 out of 4.00, edging GLM-5.3 at 2.92 and Kimi K3 at 2.94; against GLM-5.3 the win/tie/loss split was 46.8%, 12.8%, 40.4%. The three scores sit close enough to read as rough parity rather than a win, and Tencent designed and ran the evaluation itself, with no third-party benchmark replicating it on the same terms.
the field has bunched upWithin a single week, GLM-5.3, Kimi K3 and Hy4 have crowded into the same capability band. When rankings rest on self-evaluation and the gaps land in decimals, an engineering team’s selection criteria fall back to license terms, deployment cost and inference reliability — things you can see. Leaderboards stop carrying information at this range.
▪ SIGNALA 2.99-to-2.92 gap says the contest among Chinese open-weight models has moved from benchmarks to engineering delivery.