
Comparación de modelos
contra
Qwen3-30B-A3B-Thinking-2507

Precios
Input
0.09
Output
0.3
Metadatos
Especificación
Estado
Deprecated
Arquitectura
Mixture of Experts
Calibrado
Sí
No
Mezcla de expertos
Sí
Sí
Parámetros totales
30B
Parámetros activados
3.3B
Razonamiento
Sí
No
Precisión
FP8
Longitud del contexto
262K
Tokens máximos
131K
Funcionalidades admitidas
Serverless
Compatible
Compatible
Serverless LoRA
Compatible
No compatible
Ajuste fino
Compatible
No compatible
Embeddings
Compatible
Compatible
Rerankers
Compatible
No compatible
Soporte para Input de Image
No compatible
No compatible
JSON Mode
Compatible
Compatible
Outputs estructuradas
Compatible
No compatible
Herramientas
Compatible
Compatible
Finalización de Fim
Compatible
No compatible
Completado de prefijo de Chat
Compatible
No compatible
en comparación
Ver cómo se compara con otros modelos populares en dimensiones clave.
vs.

Step-3.5-Flash
vs.

Qwen3-VL-32B-Instruct
vs.

Qwen3-VL-32B-Thinking
vs.

Qwen3-VL-30B-A3B-Instruct
vs.

Qwen3-VL-30B-A3B-Thinking
vs.

Qwen3-VL-235B-A22B-Instruct
vs.

Qwen3-VL-235B-A22B-Thinking
vs.

Qwen3-VL-235B-A22B-Instruct
vs.

Qwen3-VL-235B-A22B-Thinking
vs.

Qwen3-Omni-30B-A3B-Instruct
vs.

Ring-flash-2.0
vs.

Ring-flash-2.0
vs.

Qwen3-Omni-30B-A3B-Captioner
vs.

Qwen3-Omni-30B-A3B-Thinking
vs.

Qwen3-Next-80B-A3B-Instruct
vs.

Qwen3-Next-80B-A3B-Thinking
vs.
gpt-oss-120b
vs.
gpt-oss-120b
vs.

Qwen3-Coder-30B-A3B-Instruct
vs.

Qwen3-30B-A3B-Thinking-2507
