【AI前沿】Tencent Hunyuan Hy3 Launch: Significant Improvement in Agent Capabilities and Product Experience

2026-07-06

AI NEWSLatest AI NewsArticleTencent Hunyuan Hy3 Launch: Significant Improvement in Agent Capabilities and Product ExperiencePublished in Latest AI NewsTime :Jul 6, 2026Read :13minuteOn July 6, Tencent Hunyuan Hy3 was officially released. Compared to the preview version, it demonstrates significantly stronger intelligence than models of the same size and is comparable to (2-5 times larger) flagship models, with further reduced pricing, and a significant improvement in overall stability and cost-effectiveness. Hy3 has been integrated into multiple businesses such as WorkBuddy/CodeBuddy, Yuanbao, Marvis, and ima. The API is now available on Tencent Cloud TokenHub, and multiple overseas API platforms will be added in the near future.The model’s intelligence has made a comprehensive leap, achieving a qualitative transformation in the practicality of intelligent agentsHy3 is a model that combines fast and slow thinking, using the MoE architecture, with a total parameter count of 295B and an activated parameter count of 21B, supporting a context length of up to 256K. The Hy3 preview released on April 23 was the first version after the reconstruction of Hunyuan. It achieved a qualitative transformation compared to Hy2 in complex reasoning, instruction following, context learning, code generation, and agent capabilities. Hy3 continues to show a clear and steep growth curve in its abilities, further improving post-training computational scale and data quality and diversity, resulting in a further enhancement in various tasks compared to the Hy3 preview, achieving performance comparable to large-scale flagship models for the first time with a smaller size.The capabilities of the Hunyuan model are rapidly improvingHy3 has made particularly significant progress in productivity tasks such as software development, office production, financial modeling, front-end design, and game creation, making it a reliable and cost-effective choice. In a blind test involving 270 experts based on real work scenarios, Hy3 (average score 2.67 / 4) outperformed GLM5.1 (average score 2.51 / 4), especially showing significant advantages in categories such as front-end development, data and storage, and CI/CD.Proven through real-world user scenarios, significantly enhancing user experienceHy3 has been tested in the use of global developers and in Tencent’s extensive real business scenarios. Since the release of the preview, daily token consumption has increased by 20 times, indicating that the market is beginning to recognize the positioning of high cost-effectiveness and practical models.As the most popular AI office intelligent agent in China, WorkBuddy has real and complex scenario requirements such as automated script generation and workflow orchestration, which have provided high-value directions for the iteration of the Hunyuan model. Since its release, the number of users who have autonomously selected the Hy3 preview on WorkBuddy has grown six times. In internal evaluations of the WorkBuddy office scenario, the task success rate of Hy3 increased from 72% to 90%, and the average time was shortened by 34%. It makes solving various problems more convenient and efficient.Yuanbao’s dialogue interaction scenarios have also provided valuable feedback for the model. Addressing issues like hallucinations in long text and AI search scenarios, Hy3 has learned to produce reliable outputs under complex evidence through fine-grained data cleaning and training constraints. In evaluations based on real logs, the common sense error rate of Hy3 decreased by half compared to the preview version, and the hallucination rate decreased by more than half. Based on the significantly improved agent capabilities of Hy3, Yuanbao has also launched agent functions. In its internal evaluation, Hy3 has exceeded many domestic excellent models such as GLM 5.1 in both comprehensive office and life service scenarios, and is sufficient to stably support real business processes. Users can input their needs in daily conversations, and Yuanbao can directly execute complex tasks and deliver files such as PPT, Word, Excel, PDF, and HTML, meeting daily office needs, and all for free.ima evaluated Hy3 based on two core scenarios: online knowledge base Q&A and agent capabilities. In agent tasks, Hy3 performed excellently, with system stability reaching 95.1%. Its tool orchestration capabilities were particularly outstanding, with a significant reduction in ineffective operations such as blind retries and failure to stop. This allows for more accurate planning of complex office tasks, completing them in one step. The knowledge base Q&A scenario also showed significant improvements, with a net increase in reasoning quality of nearly 19%, leading to more systematic thinking and broader information coverage. The structural completeness and usability of long-form writing and solution generation have also significantly improved.In the core scenarios of file editing/creation, file mana