Physix Frontier · News Briefing Card (IT Home · Sep 7, 2026)

Tencent Hunyuan Hy4 Preview Cuts Token Usage

KEY FACTS

  • Tencent’s Hunyuan team and WorkBuddy have fully deployed an optimized version of the Hy4 preview model.
  • The update specifically addresses long thinking chains and excessive self-verification in complex tasks.
  • Post-optimization, the model significantly reduces task rounds and token consumption without compromising performance.
  • Released on August 28, the model features 77 billion total parameters with 4.9 billion active.

KEY DATA

770 BTotal Parameters
49 BActive Parameters
1MContext Length

PHYSIX OBSERVATION

This update directly targets a critical pain point in LLM applications: prohibitive inference costs. By algorithmically suppressing wasteful internal processing to compress tokens without sacrificing performance, it lowers deployment barriers for enterprises. This signals a shift in open-source competition from raw parameter scaling to balancing inference efficiency with economic viability, favoring real-world implementation.

Source: IT Home report