Lab release history
Last updated Aug 28, 2026
Tencent Hunyuan model releases
Tencent's Hunyuan foundation-model team, releasing large MoE and multimodal models. This page collects the lab's model releases, lifecycle events, source links, and model metadata in one crawlable record.
8 models
Hy4 preview
AvailableTencent Hunyuan's fourth-generation flagship, released and open-sourced under Apache-2.0 on Aug 28 2026 (weights and an FP8 variant on Hugging Face as tencent/Hy4-preview). A Mixture-of-Experts model with 770B total parameters and ~49B activated per token: a 78-layer backbone with 256 routed experts plus a shared expert in most layers (8 routed experts selected per token) and a native multi-token-prediction layer for speculative decoding. Carries a 1M-token (1,048,576) context and is aimed at long-horizon software engineering (planning, debugging, verification across extended tasks with tool calling), office and financial analysis, cross-document work, game-development prototyping, and scientific workloads. Available via Tencent Cloud TokenHub and OpenRouter (model id tencent/hy4-preview, tool calling + structured outputs) or self-hosted on vLLM/SGLang; OpenRouter listed launch pricing of $0.834 input / $2.501 output / $0.042 cache-read per 1M tokens (Aug 28 snapshot). In Tencent's internal blind test (163 experts over 203 engineering tasks) Hy4 averaged 2.99 vs 2.92 for GLM-5.3 and 2.94 for Kimi K3 — narrow margins, and Tencent flags preview-stage tendencies to over-reason and over-verify. All figures vendor-reported and unverified by independent labs at launch.
Hy-MT2-30B-A3B
AvailableThe flagship of Tencent Hunyuan's Hy-MT2 family of 'fast-thinking' multilingual machine-translation models, open-weighted on Hugging Face on Aug 20 2026. A Mixture-of-Experts model with 30B total and ~3B active parameters covering 33 language pairs plus five Chinese-dialect and minority-language pairs, with workflows for structured/delimiter-based, contextual, glossary-based, and style-guided translation. Runs a short 8,192-token context with up to 4,096 output tokens and is small enough to run locally. Tencent reports it outperforming open heavyweights such as DeepSeek-V4-Pro and Kimi K2.6 on translation quality, with even the smaller 1.8B sibling (Hy-MT2-1.8B, released the same day) beating commercial APIs from Microsoft and Doubao — vendor figures, unverified at launch. A specialized translation LLM (text in/out), included as in-scope; the smaller 1.8B and FP8 variants are not tracked separately.
Hunyuan Hy3
AvailableThe general-availability release of Tencent's third-generation Hunyuan (Hunyuan 3.0), officially launched and open-sourced on July 6, 2026 after April's "Hy3 preview". A 295B-total / 21B-active Transformer MoE with an additional 3.8B multi-token-prediction (MTP) layer and a 256K-token context, offering three selectable inference modes that blend fast and slow thinking. Positioned as a leading open model for its size and cost efficiency, with standout results in coding, search, and scientific reasoning: Tencent reports it rivals flagship open models such as GLM-5.2 and DeepSeek-V4 (at 2-5x the active parameters) and matches or surpasses GPT-5.5 on several science benchmarks. Vendor-reported scores include 78.0 on SWE-bench Verified and 57.9 on SWE-bench Pro. Now Apache-2.0 licensed (the preview used Tencent's community license), with weights on Hugging Face (tencent/Hy3) and ModelScope and a free two-week API route on OpenRouter (tencent/hy3:free) through July 21, 2026. Deeply integrated into WeChat and Tencent's core products.
Hunyuan Hy3-preview
AvailableTencent's third-generation Hunyuan, rebuilt from scratch in ~90 days and open-sourced as the "Hy3 preview". A 295B-total / 21B-active Transformer MoE (80 layers, 192 experts with top-8 routing, plus a 3.8B multi-token-prediction layer) with a 256K-token context, positioned as a leading open reasoning-and-agent model for its size with strong cost efficiency. Vendor-reported results: 74.4 on SWE-bench Verified, 54.4 on Terminal-Bench 2.0, and 70.2 on WideSearch, with strong STEM-olympiad performance. Open weights on GitHub and Hugging Face under Tencent's community license.
Hunyuan-A13B-Instruct
AvailableTencent Hunyuan open-weight fine-grained MoE model with 80B total parameters and 13B active parameters, optimized for agentic tool use.
Hunyuan-Large
AvailableTencent's 389B total / 52B active open-weight Transformer MoE, released with a 256K pretraining context and 128K instruct context.
Hunyuan Turbo
AvailableTencent's faster, lower-cost Hunyuan update before the open Hunyuan-Large model card.
Hunyuan
AvailableTencent's first Hunyuan foundation model release, introduced as a general-purpose Chinese enterprise model.