Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Hy-MT2-30B-A3B

Hy-MT2-30B-A3B is Tencent's open-source translation model, the second generation of the Hy-MT lineage, designed for mutual translation across 33 languages including Chinese and selected ethnic minority languages. Tencent frames it as Hy Translation 2.0 and reports that it significantly improves instruction-following over the previous generation, handling structured, delimiter-based, contextual, glossary-based, and style-adapted translation tasks that require more than literal word-for-word rendering. The model is positioned for both specialized domains and real business scenarios where translation must obey precise formatting and stylistic rules.

Underneath, the architecture is a sparsely activated Transformer with 30 billion total parameters but only 3 billion active per token, organized across 48 layers where layer 0 is dense and layers 1 through 47 are Mixture-of-Experts. Each MoE layer routes tokens to 8 of 128 routed experts with a sigmoid gating function, plus one shared expert, using Grouped Query Attention with per-head QK RMSNorm and RoPE position encoding. NVIDIA's coverage documents a 256K-token context window, and Tencent cites leading results on open-source translation benchmarks such as Flores200 and WMT25, making the model a strong fit for long-context, multilingual translation workloads where instruction adherence and broad language coverage matter more than general-purpose conversation.

OpenRoutertencent/hy-mt2-30b-a3bHy

Quick Info

Powered by
Provider
OpenRouter
Model key
tencent/hy-mt2-30b-a3b
Release date
Aug 20, 2026
Last updated
Aug 20, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.074
Output token cost
$0.295

Limits

Output tokens
4,096 tokens
Context window
8,192 tokens

Transparent token rates

Compare Hy-MT2-30B-A3B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Hy-MT2-30B-A3B

OpenRouter

Coverage

The Hy-MT2 paper on alphaXiv describes a family of fast-thinking multilingual translation models from Tencent's Hunyuan team, with three sizes — 1.8B, 7B, and 30B-A3B MoE — all supporting translation across 33 languages and following translation instructions in multiple languages. The authors position the family as han The paper reports that the 7B and 30B variants outperform open-source models such as DeepSeek-V4-Pro and Kimi K2.6 in fast-thinking mode, while the lightweight 1.8B surpasses mainstream commercial translation APIs from Microsoft and Doubao overall. For on-device deployment, AngelSlim's 1.25-bit extreme quantization com

OpenRouter

CoverageRelease Notes

The NVIDIA NeMo AutoModel 0.5.0 release notes (dated 26.06) add native model support and a supervised fine-tuning recipe for HY-MT2-30B-A3B. The same release broadens coverage to other large MoE and dense models — DeepSeek-V4 Flash, ERNIE 4.5, MiMo-V2-Flash, Ling 2.0, Hy3-preview, Falcon H1, MiniCPM5-1B, and Nemotron-3 For Hy-MT2-30B-A3B specifically, NeMo AutoModel 0.5.0 ships a native model and checkpoint adapter plus an SFT recipe, enabling teams to fine-tune the 30B-A3B translation MoE within NeMo's training pipelines. The release also extends discrete-diffusion LLM training to LLaDA2 and Nemotron-Labs-Diffusion, adds diffusion f

OpenRouter

CoveragePreview

ThursdAI's Tencent releases index tracks 19 Hunyuan-team releases since January 2025, 14 shipped with open weights, including the Hy-MT2 translation family on May 28, 2026 under Apache 2.0. The same Hunyuan portfolio covers 3D generation, video, avatars, OCR, and portrait models, with more recent entries such as the AU Earlier 2026 entries include the Hy3-preview 295B/21B MoE reasoning model and the Hunyuan3D WorldClaw text-to-3D game-world paper (Aug 13, 2026). The index positions Hy-MT2 within Tencent's broader pattern of releasing specialized, open-weight models for distinct modalities — translation, speech, 3D, reasoning — with H

OpenRouter

Coverage

Tencent Hunyuan has open-sourced Hy-MT2, a multilingual translation model family, and launched a companion "Tencent Hy Translation" mini program based on it. The family ships in three sizes — a lightweight 1.8B variant, a 7B "sweet spot" model, and the flagship Hy-MT2-30B-A3B — all supporting bidirectional translation Hy-MT2-30B-A3B is the family's first mixture-of-experts design, expanding total parameter count while keeping activated parameters per inference in check for professional-scenario translation. Tencent's published evaluations report the 1.8B/7B/30B-A3B models reaching 88.1%, 96.9%, and 98.1% of Gemini 3.1 Pro's level on

Videos about Hy-MT2-30B-A3B

More models around Hy-MT2-30B-A3B