Tencent Opens Hy3 to Global Users, Claims Top Spot on OpenRouter Within a Week

Tencent announced global availability of its Hy3 large language model on August 5, expanding access beyond China through three channels: the WorkBuddy AI workspace, the Miora creative studio, and the Tencent Cloud TokenHub model-as-a-service platform. The rollout follows Hy3's initial release on July 6 and comes with a free access period on WorkBuddy through August 31. Hy3 uses a hybrid fast-and-slow-thinking Mixture-of-Experts architecture with 295 billion total parameters and 21 billion activ

2 min
Tencent Opens Hy3 to Global Users, Claims Top Spot on OpenRouter Within a Week

Tencent announced global availability of its Hy3 large language model on August 5, expanding access beyond China through three channels: the WorkBuddy AI workspace, the Miora creative studio, and the Tencent Cloud TokenHub model-as-a-service platform. The rollout follows Hy3's initial release on July 6 and comes with a free access period on WorkBuddy through August 31.

Hy3 uses a hybrid fast-and-slow-thinking Mixture-of-Experts architecture with 295 billion total parameters and 21 billion active parameters, supporting a context length of up to 256,000 tokens. Tencent claims the model performs comparably to flagship models with two to five times as many active parameters on reasoning, instruction following, code generation, and agent tasks.

Usage metrics suggest strong early demand. Tencent reports Hy3 generated more than 68 times the API call volume of its predecessor and reached the top position on OpenRouter's global LLM usage leaderboard within one week of launch. The model is available under the Apache 2.0 license and has been distributed through Hugging Face, ModelScope, and third-party developer platforms including Cline, Kilo, and OpenCode.

Pricing on OpenRouter starts at $0.1288 per million input tokens and $0.5336 per million output tokens, positioning Hy3 as a cost-competitive option against comparable models from OpenAI, Anthropic, and Google.

Tencent is pairing the model with its existing product ecosystem. WorkBuddy, which Tencent describes as China's most widely used AI agent workspace, achieved a task success rate above 90 percent in internal evaluations when running on Hy3, while reducing average task completion time by 34 percent compared to the previous model generation. Miora, Tencent's AI-native creative studio, connects Hy3's reasoning capabilities into workflows spanning graphics, video, 3D, and UI design.

On the enterprise side, Tencent Cloud TokenHub serves as a multi-model gateway with intelligent routing, letting organizations switch between Hy3 and third-party models through a single API. Regional partners including South Korea's Cafe24 and Japan's Metelix are integrating Hy3 into their respective AI platform services.

The global expansion of Hy3 adds another major Chinese model to an increasingly crowded international market, following recent launches from Alibaba's Qwen, ByteDance's Seed, and Moonshot AI's Kimi series.

Written by

More to read

  • OpenAI Flags Astra Model as Potentially Reaching Critical Cybersecurity Risk Level

    # OpenAI Flags Astra Model as Potentially Reaching "Critical" Cybersecurity Risk Level OpenAI has paused parts of development on its upcoming Astra model after internal evaluations indicated it could reach the highest risk tier — "Critical" — in the company's Preparedness Framework for cybersecurity capabilities. This is the first time OpenAI has flagged one of its own models as potentially reaching this level. ## Key Points - Internal tests of Astra showed "significant advancements in agenti

    1 min
  • ByteDance Trains 10 Trillion-Parameter AI Model to Rival Anthropic's Mythos

    ByteDance is pretraining a large model with up to 10 trillion parameters, a scale the Financial Times reports could put it in the same class as Anthropic's most advanced systems. The model, still in early pretraining, would be more than three times the size of Moonshot AI's Kimi K3, currently the largest Chinese model at 2.8 trillion parameters. Three people familiar with the project told the FT the model is in pretraining, a phase that typically lasts three to six months before full training a

    1 min