๐Ÿฆ™Freshcollected in 3h

Tencent Gray-Tests Flagship Hunyuan Hy4

Tencent Gray-Tests Flagship Hunyuan Hy4
PostLinkedIn
๐Ÿฆ™Read original on Reddit r/LocalLLaMA

๐Ÿ’กHy4 may be Tencent's next flagship, with early signs of expert-level reasoning and tool use.

โšก 30-Second TL;DR

What Changed

Hy4 reportedly appeared in Tencent Yuanbao's model selector for gray testing.

Why It Matters

If the listing reflects a genuine rollout, Hy4 could strengthen Tencent's position in China's increasingly competitive general-purpose and reasoning-model market. Its apparent tool-use and multimodal focus may also make it relevant for agent and enterprise application builders.

What To Do Next

Monitor the Yuanbao App model list and test Hy4 on tool-calling and multimodal workflows when access becomes available.

Who should care:Founders & Product Leaders

Key Points

  • โ€ขHy4 reportedly appeared in Tencent Yuanbao's model selector for gray testing.
  • โ€ขThe interface labels Hy4 as an expert-level model and highlights tool-use capabilities.
  • โ€ขHy4 is positioned above Hy3 and alongside DeepSeek in the selection list.
  • โ€ขTencent previously confirmed that a larger-parameter Hy4 with stronger multimodal performance was coming soon.

๐Ÿง  Deep Insight

Background and context from public sources โ€” not the original article. 26 sources cited.

๐Ÿ”‘ Enhanced Key Takeaways

  • โ€ขTencent's AI strategy for 2025-2026 pivoted from integrating AI into existing applications to a focused build-out of massive data center infrastructure to support its proprietary Hunyuan large language model.
  • โ€ขThe company committed to more than doubling its investment in AI products and models in 2026, including Hunyuan and the Yuanbao application, compared to its 2025 spending of 18 billion yuan (US$2.6 billion).
  • โ€ขThe Hunyuan Hy3, officially released in July 2026, is an open-weight flagship large language model built on a Mixture-of-Experts (MoE) architecture with 295 billion total parameters and 21 billion active parameters, supporting a 256K token context length.
  • โ€ขTencent also offers Hunyuan 3D, a generative model launched globally in November 2025, which creates textured 3D meshes from text or reference images and exports standard GLB and OBJ files.
  • โ€ขIn March 2026, Tencent Cloud implemented significant price increases for its Hunyuan 2.0 Instruct and Hunyuan 2.0 Think models, with input prices surging by approximately 463% and output prices by over 456%.
๐Ÿ“Š Competitor Analysisโ–ธ Show
Feature/MetricTencent Hunyuan Hy3 (Flagship)DeepSeek (General LLM)
ArchitectureMixture-of-Experts (MoE)Mixture-of-Experts (MoE)
Total Parameters295 Billion671 Billion
Active Parameters21 Billion per token37 Billion per task
Context LengthUp to 256K tokensUp to 128K tokens (DeepSeek) / 1M tokens (DeepSeek V4)
Key Benchmarks (Hy3)SWE-bench Verified: 78, GPQA Diamond: 90.4HumanEval (coding): 73.78%, GSM8K (problem-solving): 84.1%
Pricing (Hy3 Preview)$0.27 per 1M tokens (input + output combined)Costs 95% less per token than GPT-4 (as of April 2026)
LicenseApache 2.0 (open-source)Open-source (for 7B/67B Base and Chat)

๐Ÿ› ๏ธ Technical Deep Dive

  • Hunyuan Hy3: This flagship model utilizes a Mixture-of-Experts (MoE) architecture, integrating both fast and slow thinking capabilities. It has a total of 295 billion parameters with 21 billion active parameters per token. Hy3 supports a context length of up to 256K tokens and is released under the commercially friendly Apache 2.0 license. It features a reasoning_effort parameter to control response latency and depth (no_think, low, high). Deployment is supported via vLLM, SGLang, OpenAI-compatible APIs, FP8 quantized models, and speculative decoding using Multi-Token Prediction (MTP).
  • Hunyuan-Large (Hunyuan-MoE-A52B): An earlier open-source MoE model from Tencent, it features a total of 389 billion parameters with 52 billion active parameters. It is capable of handling up to 256K tokens and employs high-quality synthetic data, KV Cache Compression (using Grouped Query Attention and Cross-Layer Attention), and Expert-Specific Learning Rate Scaling.
  • Hunyuan 3D: This generative model transforms text prompts or reference images into textured 3D meshes, exporting standard GLB and OBJ files. It supports Physically-Based Rendering (PBR) textures for realistic materials.
  • HunyuanVideo: Part of the broader Hunyuan AI platform, this model is designed for AI-powered video generation from text prompts and still images, focusing on coherent motion, realistic environments, and scene-level consistency.

Note: Specific technical details for Hunyuan Hy4, beyond its expected larger parameter scale and enhanced multimodal capabilities, are not yet publicly available as it is currently in gray testing.

๐Ÿ”ฎ Future ImplicationsAI analysis grounded in cited sources

Tencent will intensify its focus on AI infrastructure development and the monetization of its AI services.
The company has pivoted its AI strategy towards building massive data centers and aims to convert these substantial infrastructure investments into profitable, high-efficiency AI services to avoid potential price wars.
The Yuanbao app is positioned to become a central platform for Tencent's advanced AI model deployment and user interaction.
The gray testing of the flagship Hy4 model within Yuanbao, alongside the existing integration of Hy3, indicates Yuanbao's strategic role as a key consumer-facing interface for Tencent's cutting-edge AI capabilities.
Tencent's Hunyuan series will continue to leverage Mixture-of-Experts (MoE) architectures and prioritize advanced agentic capabilities.
The Hy3 model is built on an MoE architecture with a strong emphasis on agent capabilities, and the 'expert-level' label for Hy4 in Yuanbao suggests a continued focus on sophisticated task execution and tool use.

โณ Timeline

2023-09
Tencent officially debuts its proprietary foundation model, Hunyuan, for enterprises in China via Tencent Cloud.
2024-05
Tencent launches an upgraded, open-source version of its Hunyuan text-to-image LLM, enhancing performance by 20%.
2024-11
Tencent open-sources Hunyuan-Large (Hunyuan-MoE-A52B), a Transformer-based MoE model with 389 billion total parameters.
2025-11
Tencent launches the Hunyuan 3D engine globally, enabling users to generate 3D assets from text, images, or sketches.
2026-03
Tencent Cloud significantly increases pricing for its Hunyuan 2.0 Instruct and Think models by over 450%.
2026-07
Tencent officially releases Hunyuan Hy3, an open-weight flagship LLM built on a MoE architecture with 295 billion total parameters.
๐Ÿ“ฐ

Weekly AI Recap

Read this week's curated digest of top AI events โ†’

๐Ÿ‘‰Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/LocalLLaMA โ†—

This is a summary, not the original. Read the source, or get the weekly briefing.

Weekly AI briefing

One email a week. Unsubscribe anytime.