Tencent's Hy4 model climbs open-source AI rankings via ecosystem training
Tencent Holdings has made significant strides with its Hy4 preview model, achieving a score of 64.3 on the DeepSWE benchmark, which outperforms Alibaba's Qwen-3.8 Max (56.6) and DeepSeek-V4 Pro (62.7). Released on Friday, Hy4 secured the eighth position globally on Code Arena's WebDev leaderboard, trailing Anthropic's Claude Fable 5 and preceding Alibaba's Qwen 3.8-Flash-Next at ninth. This represents a remarkable leap from Tencent's earlier model, Hy3, which was ranked 34th. Analysts, including Goldman Sachs' Ronald Keung, noted that Hy4's success stems from Tencent's unique product-plus-model strategy, leveraging user data from its extensive ecosystem for continuous training, particularly enhancing productivity and coding tasks in the evolving agentic AI landscape.
Key facts
- Hy4 preview scored 64.3 on DeepSWE benchmark, surpassing Alibaba's Qwen-3.8 Max (56.6) and DeepSeek-V4 Pro (62.7).
- Hy4 preview ranked eighth globally on Code Arena's WebDev leaderboard, behind Anthropic's Claude Fable 5 and ahead of Alibaba's Qwen 3.8-Flash-Next.
- Tencent's previous model Hy3 ranked 34th on the same benchmark.
- Hy4 preview was released on Friday.
- Tencent uses its product ecosystem to train models, collecting user data for iterative training.
- Goldman Sachs analysts led by Ronald Keung commented on the strategy.
- Hy4 preview brings Tencent's Hunyuan series back to top tier of open-source models.
- Notable gains in coding capabilities over Hy3.
Entities
Institutions
- Tencent Holdings
- Alibaba Group Holding
- DeepSeek
- Goldman Sachs
- Anthropic
- Code Arena