정보에 대해서Hy3-preview
Hy3 preview is a 295B-parameter Mixture-of-Experts (MoE) language model from Tencent Hunyuan, built for production-grade agent workloads. With only 21B parameters activated per token and native 256K context support, it handles complex tasks like cross-file code refactoring, long-document analysis, and multi-step tool use, rather than just generating fluent dialogue. Hy3 scores near state-of-the-art on SWE-bench Verified and advanced STEM benchmarks, while offering three inference modes (no_think, think_low, think_high) to dynamically trade off latency and reasoning depth. Its sparse activation architecture delivers competitive intelligence at a significantly lower token cost.
Leverage Hy3-preview’s 295B-parameter MoE architecture and 256K context for production-grade agentic workflows and deep reasoning.
Repository-Scale Refactoring
Execute complex architectural changes across massive codebases using native 256K context support.
Use Case Example:
"Migrated a distributed Go backend from REST to gRPC, updating service definitions and client libraries across 40+ repositories in a single pass."
Autonomous DevOps Agents
Power agents that navigate environments, use multi-step tools, and solve infrastructure issues autonomously.
Use Case Example:
"Deployed an agent to resolve a memory leak in a Rust-based embedded system by analyzing core dumps and applying a validated firmware patch."
PhD-Level STEM Reasoning
Tackle advanced scientific challenges in math and biology using high-depth reasoning modes.
Use Case Example:
"Formulated a formal proof for a fluid dynamics theorem using 'think_high' mode to validate complex boundary conditions for a research paper."
Intelligent Document Auditing
Analyze lengthy technical or legal documents to detect logical gaps and hidden risks with high precision.
Use Case Example:
"Scanned a 150-page semiconductor schematic and its manual to identify a power-sequencing logic error before the fabrication phase."
메타데이터
사양
주
Deprecated
건축
Mixture-of-Experts
교정된
아니요
전문가의 혼합
네
총 매개변수
80B
활성화된 매개변수
21B
추론
아니요
Precision
FP8
콘텍스트 길이
262K
Max Tokens
다른 모델과 비교
이 Model이 다른 것들과 어떻게 비교되는지 보세요.

Tencent
chat
Hunyuan-MT-7B
Total Context:
33K
Max output:
33K
Input:
$
0.0
/ M Tokens
Output:
$
0.0
/ M Tokens

Tencent
chat
Hunyuan-A13B-Instruct
Total Context:
131K
Max output:
131K
Input:
$
0.14
/ M Tokens
Output:
$
0.57
/ M Tokens

Tencent
chat
Hy3
Total Context:
262K
Max output:
262K
Input:
$
0.132
/ M Tokens
Output:
$
0.528
/ M Tokens

Tencent
chat
Hy3-preview
Total Context:
262K
Max output:
Input:
$
0.066
/ M Tokens
Output:
$
0.26
/ M Tokens
