Model card
Hy3-preview is a specialized Mixture-of-Experts (MoE) model from Tencent, engineered specifically for agentic workflows and high-throughput production environments. Unlike standard monolithic LLMs, Hy3 offers a unique architectural advantage through configurable reasoning modes. Developers can toggle between disabled, low, and high reasoning levels, providing a granular way to balance inference latency against complex problem-solving capabilities. This makes it particularly effective for multi-step agent loops where cost and speed are as critical as logic. With a substantial 262k context window, it handles long-form documentation and extensive conversation histories without significant degradation. For teams integrating via API, Hy3 serves as a scalable middle ground between lightweight chat models and heavy-duty reasoning engines, offering a predictable way to scale compute resources based on the specific complexity of the incoming task.
Model files and versions
Download this model
How to use
- 01Step 1
Read the model card and source information.
- 02Step 2
Start with a small, non-sensitive evaluation.
- 03Step 3
Review quality, licensing and usage limits.
- 04Step 4
Adopt it only after validation.
Discussions
Use this space to keep checking source information, usage experience and maintenance status.
Open source page