Tencent Hy icon

Tencent Hy

Hy4 Preview

Tencent's unified AI brand. Flagship Hy3-preview is an open-weight 295B MoE LLM (21B active) with 256K context and limited-time free Tencent Cloud API.

Content updated 3 days ago·Hy4 Preview released 4 days ago

Pricing:Free + from ¥1/per use
Try for Free
Jump to section
Tencent Hy official model page

More tools to compare

Omniwork icon

Omniwork

Coasty icon

Coasty

MakersClaw icon

MakersClaw

Freebuff icon

Freebuff

AutoClaw icon

AutoClaw

Blocks.ai icon

Blocks.ai

Pros & Cons

Pros

  • 770B total / 49B active MoE profile gives Tencent a much larger open flagship than Hy3.
  • 1M context is useful for long codebases, research sets, documents, and agent traces.
  • Apache 2.0 weights make commercial review simpler than Hy3's custom community license.
  • TokenHub and OpenRouter access give teams hosted routes in addition to self-hosting.
  • WorkBuddy and CodeBuddy launch access gives users a short free evaluation window.

Cons

  • Hy4 Preview is an early release, not a stable GA model.
  • Tencent notes known issues around over-long reasoning and over-verification.
  • Self-hosting a 770B MoE model still requires serious infrastructure, even with FP8 weights.
  • Independent third-party evaluation coverage is still limited at launch.
  • Teams outside the Tencent ecosystem may need extra integration and compliance review.

Overview

Tencent Hy is Tencent's global foundation-model brand, formerly known as Tencent Hunyuan. The lineup spans language, reasoning, agent, image, video, and 3D models. Its current flagship language release is Hy4 Preview, released and open-sourced on August 28, 2026.

Hy4 Preview is a Mixture-of-Experts language model with 770 billion total parameters, 49 billion active parameters per token, and a 1M-token context window. Tencent positions it for real productivity tasks across coding, office work, game development, finance, security, and scientific research. It is available through Tencent products, Tencent Cloud TokenHub, OpenRouter, and self-hosted model weights.

For teams evaluating an AI agent or coding model, the important decision is not only benchmark score. Hy4 Preview combines Apache 2.0 open weights, a much longer context window than Hy3, hosted API access, and Tencent product integration. It is most relevant to developers who want to test a large open frontier model while keeping Hy3 as the safer stable baseline.

Key Features

  • 770B MoE architecture - Hy4 Preview uses a mixture-of-experts design with 770B total parameters and 49B active parameters. Sparse activation is intended to improve serving efficiency compared with activating an equivalent dense model for every token.

  • 1M-token context window - The model supports long documents, codebases, agent traces, research corpora, and multi-file reasoning tasks that exceed ordinary chat limits. This is useful for coding agents, long-form analysis, and enterprise knowledge work.

  • Productivity-focused training - Tencent says Hy4 Preview was developed with training data co-created by internal experts across software engineering, game development, finance, security, and related domains.

  • Tencent ecosystem integration - Hy4 Preview is available through WorkBuddy, CodeBuddy, Yuanbao, ima, and other Tencent products. This gives enterprise users more routes to evaluate the model beyond a standalone model card.

  • TokenHub and OpenRouter API access - Hy4 Preview can be reached through Tencent Cloud TokenHub and OpenRouter. Tencent's launch announcement lists API pricing at USD 0.834 per million input tokens, USD 2.501 per million output tokens, and USD 0.042 per million cache-hit tokens.

  • Apache 2.0 open weights - Hy4 Preview and Hy4 Preview-FP8 weights are published through Tencent's official model channels under Apache License 2.0, with vLLM and SGLang deployment recipes.

How It Compares

Hy4 Preview competes with Chinese and international frontier models used for reasoning, coding, and agent workflows. Its strongest differentiators are the combination of large MoE scale, 1M context, Apache 2.0 weights, Tencent product access, and hosted API routes.

Tencent Hy4 Preview Tencent Hy3 GPT-5.6 Sol Claude Opus 4.8 GLM-5
Primary positioning Open flagship productivity model Stable Tencent Hy release Flagship OpenAI model Closed frontier assistant model Chinese frontier model
Parameters 770B total, 49B active 295B total, 21B active Undisclosed Undisclosed Undisclosed
Context Up to 1M Up to 256K Up to 1M in published evals Up to 1M in some surfaces Up to 1M in some surfaces
Hosted API Tencent Cloud TokenHub, OpenRouter Tencent Cloud TokenHub OpenAI API Anthropic API Varies by provider
Open weights Yes, Apache 2.0 Yes, custom Tencent license No No Partial/provider-dependent

The main tradeoff is release maturity. OpenAI and Anthropic have broader global developer tooling and third-party integrations, while Hy3 has a more stable Tencent release profile. Tencent Hy4 Preview may be more attractive when cost, 1M context, Apache 2.0 weights, Chinese-language ecosystem alignment, or Tencent Cloud deployment matter more.

Pricing & Plans

Tencent Hy has multiple access routes: product integrations, TokenHub API access, OpenRouter access, and model materials for self-hosted evaluation. API pricing is usage-based rather than a flat subscription.

Tencent's Hy4 Preview launch announcement lists:

Token type Official listed price
Input tokens USD 0.834 per million tokens
Output tokens USD 2.501 per million tokens
Cache hits USD 0.042 per million tokens

At launch, Hy4 Preview is free on WorkBuddy and CodeBuddy for two weeks, and free Hy3 access on both platforms has been extended until September 30, 2026. If you are buying through Tencent Cloud, OpenRouter, or another provider, verify the current region, currency, quota, cache policy, and billing unit before estimating production cost.

Best For

  • Tencent Cloud or OpenRouter users who want a hosted route to Hy4 Preview.
  • Developers building coding agents, office agents, search agents, or long-document workflows.
  • Research and ML teams evaluating Apache 2.0 open frontier-class MoE models.
  • Enterprises in supported Tencent regions that need Chinese ecosystem alignment.
  • Cost-sensitive teams comparing hosted reasoning models across providers.

FAQ

What is Tencent Hy?

Tencent Hy is Tencent's global foundation-model brand, formerly known as Tencent Hunyuan. It covers language, reasoning, agent, image, video, and 3D models.

What is Hy4 Preview?

Hy4 Preview is Tencent's August 2026 language-model release built on a 770B-parameter MoE architecture with 49B active parameters and up to 1M context.

Is Hy4 Preview different from Hy3?

Yes. Hy4 Preview is larger than Hy3, raises context from 256K to 1M tokens, and uses Apache 2.0 weights. Hy3 remains the more stable official July 2026 release baseline.

How much does the Tencent Hy4 Preview API cost?

Tencent's launch announcement lists Hy4 Preview at USD 0.834 per million input tokens, USD 2.501 per million output tokens, and USD 0.042 per million cache-hit tokens.

Is Tencent Hy4 Preview open source?

Hy4 Preview is released under Apache License 2.0, with model weights available through official Tencent channels including Hugging Face, ModelScope, GitCode, and CNB.

What context length does Hy4 Preview support?

Tencent states Hy4 Preview supports up to 1M tokens of context. That is a major increase over Hy3's 256K-token context window.

Who should consider Tencent Hy?

Tencent Hy is most relevant for teams that need long-context reasoning, coding-agent capability, Tencent Cloud access, Chinese ecosystem alignment, or an open-weight alternative to closed API-only models.

Version History

Hy4 Preview

Current Version

Released on August 28, 2026

View Update
+What's new
3 updates
  • Evaluate Tencent's next-generation open flagship with 770B total parameters, 49B active parameters, and a 1M-token context window for coding, office, game, and scientific workflows
  • Deploy Hy4 Preview through Tencent products, Tencent Cloud TokenHub, OpenRouter, or self-hosted Apache 2.0 weights with BF16 and FP8 checkpoints plus vLLM and SGLang recipes
  • Compare the preview's new productivity training, internal engineering-task evaluations, 31.8% inference-throughput optimization claim, and early known limitations before migration

Hy3

Released on July 6, 2026

View Update
+What's new
3 updates
  • Improve production reliability and cost efficiency over Hy3 Preview while keeping the 295B MoE architecture, 21B active parameters, and 256K context window
  • Deploy official Hy3 across Tencent products and TokenHub API access, replacing the earlier preview checkpoint as the current release
  • Run fast and slow thinking in one model for reasoning, coding, instruction following, and agent workflows

Hy3 Preview

Released on April 23, 2026

View Update
+What's new
3 updates
  • Run Hy3 Preview as Tencent's current open flagship LLM, using a 295B-parameter MoE model with 21B active parameters and an added 3.8B MTP layer for faster generation
  • Analyze long codebases, documents and agent sessions with up to 256K tokens of context, supported by 80 transformer layers and grouped-query attention
  • Switch between quick answers and deeper reasoning per task, with reported results of 74.4 on SWE-bench Verified, 54.4 on Terminal-Bench 2.0 and 67.1 on BrowseComp

Hunyuan 2.0

Released on December 5, 2025

+What's new
2 updates
  • Access HY 2.0 in Yuanbao, ima and Tencent Cloud APIs with 406B total parameters, 32B active MoE, 256K context, and stronger reasoning, coding and instruction following
  • Mark the international naming shift from Tencent Hunyuan to Tencent HY, while keeping Hunyuan as the technical family name used in repositories, APIs and earlier documentation

Hunyuan Compact Series (0.5B / 1.8B / 4B / 7B)

Released on July 30, 2025

+What's new
2 updates
  • Deploy compact Hunyuan models in 0.5B, 1.8B, 4B and 7B sizes, with Pretrain and Instruct checkpoints available on Hugging Face for local and edge use
  • Fine-tune smaller Hunyuan models for consumer GPUs, PCs, mobile and embedded environments, while retaining long-context, hybrid-reasoning and agent-oriented capabilities

Hunyuan-A13B

Released on June 25, 2025

+What's new
2 updates
  • Run Hunyuan-A13B as an open MoE model with 80B total parameters, 13B active parameters and 256K context for efficient reasoning, coding and agent workloads
  • Switch between fast and slow thinking modes in a smaller open model, giving developers a lower-cost bridge between T1-style reasoning and later compact dense releases

Hunyuan-T1 (Official)

Released on March 21, 2025

+What's new
2 updates
  • Use Hunyuan-T1 as the official deep-reasoning model built on TurboS, improving complex math, logic, science and coding tasks over the earlier preview version
  • Access T1 through Tencent Cloud and Tencent Yuanbao for long-context reasoning workflows, rather than relying only on the earlier Yuanbao preview experience

Hunyuan TurboS

Released on February 27, 2025

+What's new
3 updates
  • Use Hunyuan TurboS for fast-answer workloads that need lower latency, with public claims of doubled output speed and 44% lower first-token latency versus the prior generation
  • Run long-context and reasoning workloads on the Hybrid-Transformer-Mamba MoE base that later powered Hunyuan-T1, while reducing sequence-processing overhead
  • Build on the TurboS base for adaptive fast and slow reasoning, choosing quicker replies for simple prompts and deeper reasoning for complex tasks

Hunyuan T1-Preview

Released on February 17, 2025

+What's new
2 updates
  • Try Hunyuan T1-Preview in Tencent Yuanbao as Hunyuan's first consumer-facing deep-reasoning model before the official T1 release
  • Test early reasoning behavior on math, coding and complex problem-solving tasks, giving Tencent feedback before the TurboS-based official version arrived in March 2025

Hunyuan-Large (389B MoE)

Released on November 5, 2024

+What's new
3 updates
  • Open-source the largest Transformer-based MoE LLM at the time — 389B total parameters with 52B active parameters and 256K context — under the Tencent Hunyuan Community License Agreement
  • Download Pretrain, Instruct and Instruct-FP8 checkpoints with the technical report, covering math, code, multilingual and long-context benchmark evaluations
  • Outperform Llama 3.1 70B and 405B on multiple academic benchmarks at launch, establishing Hunyuan-Large as a reference open MoE baseline for the research community

Hunyuan Turbo

Released on September 5, 2024

+What's new
3 updates
  • Launch the next-generation Hunyuan flagship at the Tencent Global Digital Ecosystem Summit as the first Hunyuan model to fully adopt a Mixture-of-Experts architecture for production serving
  • Double training efficiency, double inference throughput and cut inference cost by roughly 50% versus Hunyuan 1.0, with input and output API pricing reduced to half of the previous generation
  • Apply the Turbo service to language understanding, text generation, math and code tasks while relying on Tencent's published benchmark comparisons for quality claims

Hunyuan 1.0

Released on September 7, 2023

+What's new
3 updates
  • Access Tencent Hunyuan as Tencent's first proprietary foundation model on Tencent Cloud, launched with more than 100B parameters and over 2T pretraining tokens
  • Connect Hunyuan-backed capabilities across 50+ Tencent products at launch, including Tencent Cloud, Tencent Meeting, Tencent Docs, Weixin Search and QQ Browser
  • Open the model to enterprises in China for testing and app building through Tencent Cloud, marking Tencent's formal entry into the Chinese frontier LLM race

Track Tencent Hy in ToolWorthy Weekly

Important tool updates, better alternatives, and selected AI signals in one weekly brief.

Weekly only. Unsubscribe anytime.