Qwen3.8-2.4T-A95B
Acerca de Qwen3.8-2.4T-A95B
Desarrollado sobre la base arquitectónica de Qwen3.5, Qwen3.8 ofrece mejoras sustanciales en programación, trabajo profesional, investigación y tareas agénticas de largo recorrido. Más allá de responder a preguntas más difíciles, Qwen3.8 está diseñado para llevar a cabo tareas complejas de múltiples pasos hasta su finalización con una mayor fiabilidad.
Serverless disponible
Ejecute consultas de forma inmediata y pague solo por el uso
Precio de Input
$
2.0
/ M Tokens
Lectura de caché
$
0.25
/ M Tokens
Precio de Output
$
6.0
/ M Tokens
Metadatos
Especificación
Estado
Available
Arquitectura
Hybrid MoE
Calibrado
No
Mezcla de expertos
Sí
Parámetros totales
2400B
Parámetros activados
95B
Razonamiento
No
Precisión
FP8
Longitud del contexto
1049K
Tokens máximos
131K
Funcionalidades admitidas
Serverless
Compatible
Serverless LoRA
No compatible
Ajuste fino
No compatible
Embeddings
No compatible
Rerankers
No compatible
Soporte para Input de Image
No compatible
JSON Mode
Compatible
Outputs estructuradas
No compatible
Herramientas
Compatible
Finalización de Fim
No compatible
Completado de prefijo de Chat
Compatible
Comparar con otros modelos
Descubre cómo se compara este modelo con otros.

Qwen
chat
Qwen3-VL-32B-Instruct
Contexto total:
262K
Output máx.:
262K
Input:
$
0.2
/ M Tokens
Output:
$
0.6
/ M Tokens

Qwen
chat
Qwen3-VL-32B-Thinking
Contexto total:
262K
Output máx.:
262K
Input:
$
0.2
/ M Tokens
Output:
$
1.5
/ M Tokens

Qwen
chat
Qwen3-VL-8B-Instruct
Contexto total:
262K
Output máx.:
262K
Input:
$
0.18
/ M Tokens
Output:
$
0.68
/ M Tokens

Qwen
chat
Qwen3-VL-8B-Thinking
Contexto total:
262K
Output máx.:
262K
Input:
$
0.18
/ M Tokens
Output:
$
2.0
/ M Tokens

Qwen
chat
Qwen3-VL-235B-A22B-Instruct
Contexto total:
262K
Output máx.:
262K
Input:
$
0.3
/ M Tokens
Output:
$
1.5
/ M Tokens

Qwen
chat
Qwen3-VL-235B-A22B-Thinking
Contexto total:
262K
Output máx.:
262K
Input:
$
0.45
/ M Tokens
Output:
$
3.5
/ M Tokens

Qwen
chat
Qwen3-VL-30B-A3B-Instruct
Contexto total:
262K
Output máx.:
262K
Input:
$
0.29
/ M Tokens
Output:
$
1.0
/ M Tokens

Qwen
chat
Qwen3-VL-30B-A3B-Thinking
Contexto total:
262K
Output máx.:
262K
Input:
$
0.29
/ M Tokens
Output:
$
1.0
/ M Tokens

Qwen
image-to-video
Wan2.2-I2V-A14B
$
0.29
/ Video
