

Comparación de modelos
GLM-4.1V-9B-Thinking
contra
Qwen3-8B

Precios
Input
$
0.035
/ M Tokens
$
0.06
/ M Tokens
Output
$
0.14
/ M Tokens
$
0.06
/ M Tokens
Metadatos
Especificación
Estado
Deprecated
Available
Arquitectura
Vision-Language Model
Causal Decoder-only
Calibrado
No
No
Mezcla de expertos
No
Sí
Parámetros totales
9B
8B
Parámetros activados
9B
8B
Razonamiento
No
No
Precisión
FP8
FP8
Longitud del contexto
66K
131K
Tokens máximos
66K
131K
Funcionalidades admitidas
Serverless
Compatible
Compatible
Serverless LoRA
No compatible
No compatible
Ajuste fino
No compatible
No compatible
Embeddings
No compatible
No compatible
Rerankers
No compatible
No compatible
Soporte para Input de Image
No compatible
No compatible
JSON Mode
No compatible
Compatible
Outputs estructuradas
No compatible
No compatible
Herramientas
No compatible
Compatible
Finalización de Fim
No compatible
No compatible
Completado de prefijo de Chat
No compatible
No compatible
GLM-4.1V-9B-Thinking en comparación
Ver cómo GLM-4.1V-9B-Thinking se compara con otros modelos populares en dimensiones clave.
vs.

Qwen3-VL-32B-Instruct
vs.

Qwen3-VL-32B-Thinking
vs.

Qwen3-VL-8B-Instruct
vs.

Qwen3-VL-8B-Thinking
vs.

Qwen3-VL-30B-A3B-Instruct
vs.

Qwen3-VL-30B-A3B-Thinking
vs.

Qwen3-Omni-30B-A3B-Instruct
vs.

Qwen3-Omni-30B-A3B-Captioner
vs.

Qwen3-Omni-30B-A3B-Thinking
vs.

Qwen3-Next-80B-A3B-Instruct
vs.

Qwen3-Next-80B-A3B-Thinking
vs.

Ling-mini-2.0
vs.

Hunyuan-MT-7B
vs.
gpt-oss-20b
vs.

Qwen3-Coder-30B-A3B-Instruct
vs.

Qwen3-30B-A3B-Thinking-2507
vs.

Qwen3-30B-A3B-Instruct-2507
vs.

Hunyuan-A13B-Instruct
vs.

Qwen3-14B
vs.

Qwen3-30B-A3B
