Hy3-preview
About Hy3-preview
Hy3 preview is a 295B-parameter Mixture-of-Experts (MoE) language model from Tencent Hunyuan, built for production-grade agent workloads. With only 21B parameters activated per token and native 256K context support, it handles complex tasks like cross-file code refactoring, long-document analysis, and multi-step tool use, rather than just generating fluent dialogue. Hy3 scores near state-of-the-art on SWE-bench Verified and advanced STEM benchmarks, while offering three inference modes (no_think, think_low, think_high) to dynamically trade off latency and reasoning depth. Its sparse activation architecture delivers competitive intelligence at a significantly lower token cost.
Leverage Hy3-preview’s 295B-parameter MoE architecture and 256K context for production-grade agentic workflows and deep reasoning.
Repository-Scale Refactoring
Execute complex architectural changes across massive codebases using native 256K context support.
Use Case Example:
"Migrated a distributed Go backend from REST to gRPC, updating service definitions and client libraries across 40+ repositories in a single pass."
Autonomous DevOps Agents
Power agents that navigate environments, use multi-step tools, and solve infrastructure issues autonomously.
Use Case Example:
"Deployed an agent to resolve a memory leak in a Rust-based embedded system by analyzing core dumps and applying a validated firmware patch."
PhD-Level STEM Reasoning
Tackle advanced scientific challenges in math and biology using high-depth reasoning modes.
Use Case Example:
"Formulated a formal proof for a fluid dynamics theorem using 'think_high' mode to validate complex boundary conditions for a research paper."
Intelligent Document Auditing
Analyze lengthy technical or legal documents to detect logical gaps and hidden risks with high precision.
Use Case Example:
"Scanned a 150-page semiconductor schematic and its manual to identify a power-sequencing logic error before the fabrication phase."
Metadata
Specification
State
Deprecated
Architecture
Mixture-of-Experts
Calibrated
No
Mixture of Experts
Yes
Total Parameters
80B
Activated Parameters
21B
Reasoning
No
Precision
FP8
Context length
262K
Max Tokens
Compare with Other Models
See how this model stacks up against others.

Tencent
chat
Hunyuan-MT-7B
Total Context:
33K
Max output:
33K
Input:
$
0.0
/ M Tokens
Output:
$
0.0
/ M Tokens

Tencent
chat
Hunyuan-A13B-Instruct
Total Context:
131K
Max output:
131K
Input:
$
0.14
/ M Tokens
Output:
$
0.57
/ M Tokens

Tencent
chat
Hy3
Total Context:
262K
Max output:
262K
Input:
$
0.132
/ M Tokens
Output:
$
0.528
/ M Tokens

Tencent
chat
Hy3-preview
Total Context:
262K
Max output:
Input:
$
0.066
/ M Tokens
Output:
$
0.26
/ M Tokens
