정보에 대해서Hy4-preview
Hy4 preview is Tencent Hy's new-generation productivity flagship model, built on Gated DeepSeek Sparse Attention with iHC residual design, with approximately 770B total and 49B activated parameters per token. It natively supports a 1M context window with up to 64k output and a native MTP layer for speculative decoding to boost generation throughput. It demonstrates strong comprehension, planning, and sustained execution on complex tasks, and natively supports tool calls, structured output, and context caching. Deeply optimized for Agent, Coding, and productivity scenarios, it suits long-horizon planning and code workflows.
사용 가능한 Serverless
쿼리를 즉시 실행하고 사용한 만큼만 지불하세요.
1M 토큰당 (Input/Output)
$
0.834
/ M Tokens
1M 토큰당 (Input/Output)
$
0.042
/ M Tokens
1M 토큰당 (Input/Output)
$
2.501
/ M Tokens
메타데이터
사양
주
Available
건축
교정된
아니요
전문가의 혼합
아니요
총 매개변수
770B
활성화된 매개변수
추론
아니요
Precision
FP8
콘텍스트 길이
1049K
Max Tokens
262K
지원됨 기능
Serverless
지원됨
Serverless LoRA
지원하지 않음
Fine-tuning
지원하지 않음
Embedding
지원하지 않음
Rerankers
지원하지 않음
지원 Image Input
지원하지 않음
JSON Mode
지원됨
구조화된 Outputs
지원하지 않음
도구
지원됨
Fim Completion
지원하지 않음
Chat Prefix Completion
지원됨
다른 모델과 비교
이 Model이 다른 것들과 어떻게 비교되는지 보세요.

Tencent
chat
Hunyuan-MT-7B
Total Context:
33K
Max output:
33K
Input:
$
0.0
/ M Tokens
Output:
$
0.0
/ M Tokens

Tencent
chat
Hunyuan-A13B-Instruct
Total Context:
131K
Max output:
131K
Input:
$
0.14
/ M Tokens
Output:
$
0.57
/ M Tokens

Tencent
chat
Hy4-preview
Total Context:
1049K
Max output:
262K
Input:
$
0.834
/ M Tokens
Output:
$
2.501
/ M Tokens

Tencent
chat
Hy3
Total Context:
262K
Max output:
262K
Input:
$
0.132
/ M Tokens
Output:
$
0.528
/ M Tokens

Tencent
chat
Hy3-preview
Total Context:
262K
Max output:
Input:
$
0.066
/ M Tokens
Output:
$
0.26
/ M Tokens
