
Comparación de modelos
DeepSeek-R1-Distill-Qwen-32B
contra
Precios
Input
$
0.18
/ M Tokens
$
/ M Tokens
Output
$
0.18
/ M Tokens
$
/ M Tokens
Metadatos
Especificación
Estado
Deprecated
Arquitectura
Dense Transformer
Calibrado
No
Sí
Mezcla de expertos
No
Sí
Parámetros totales
32B
Parámetros activados
32B
Razonamiento
No
Sí
Precisión
FP8
Longitud del contexto
131K
Tokens máximos
131K
Funcionalidades admitidas
Serverless
Compatible
Compatible
Serverless LoRA
No compatible
Compatible
Ajuste fino
No compatible
Compatible
Embeddings
No compatible
No compatible
Rerankers
No compatible
Compatible
Soporte para Input de Image
No compatible
No compatible
JSON Mode
Compatible
Compatible
Outputs estructuradas
No compatible
Compatible
Herramientas
Compatible
Compatible
Finalización de Fim
Compatible
Compatible
Completado de prefijo de Chat
No compatible
Compatible
DeepSeek-R1-Distill-Qwen-32B en comparación
Ver cómo DeepSeek-R1-Distill-Qwen-32B se compara con otros modelos populares en dimensiones clave.
vs.

GLM-4.6V
vs.

Qwen3-VL-32B-Instruct
vs.

Qwen3-VL-32B-Thinking
vs.

Qwen3-VL-8B-Instruct
vs.

Qwen3-VL-8B-Thinking
vs.

Qwen3-VL-30B-A3B-Instruct
vs.

Qwen3-VL-30B-A3B-Thinking
vs.

Qwen3-Omni-30B-A3B-Instruct
vs.

Ring-flash-2.0
vs.

Ling-flash-2.0
vs.

Qwen3-Omni-30B-A3B-Captioner
vs.

Qwen3-Omni-30B-A3B-Thinking
vs.

Qwen3-Next-80B-A3B-Instruct
vs.

Qwen3-Next-80B-A3B-Thinking
vs.

Ling-mini-2.0
vs.

Hunyuan-MT-7B
vs.
gpt-oss-120b
vs.
gpt-oss-20b
vs.

Qwen3-Coder-30B-A3B-Instruct
vs.

Qwen3-30B-A3B-Thinking-2507
