Tencent's Hunyuan team released the Hy4 preview open-weight artificial intelligence model using a Mixture-of-Experts architecture. The model features 770 billion total parameters, 49 billion active parameters per token, and a context window of 1 million tokens. The system identified its own inference bottlenecks during research and increased end-to-end throughput 31.8% via operator fusion and communication optimizations.

Designed for coding and productivity, the model employs 256 routed experts as part of a trend of open releases from Chinese developers including DeepSeek, Qwen, and Kimi. Tencent distributed the weights via HuggingFace and Github to allow developers and researchers to build on the architecture.

Sign in to suggest edits

Key sources

  1. SOURCE@tencenthunyuan“lifted e2e throughput 31.8% via operator fusion and comms opts”x.com
  2. SUPPORT@sudeepsriv“activates only around 49B for each token”x.com
Markdown