

Comparación de modelos
Qwen3-30B-A3B-Instruct-2507
contra
Ring-flash-2.0

Precios
Input
$
0.09
/ M Tokens
$
0.14
/ M Tokens
Output
$
0.3
/ M Tokens
$
0.57
/ M Tokens
Metadatos
Creado el
Licencia
APACHE-2.0
MIT LICENSE
Proveedor
Qwen
inclusionAI
Especificación
Estado
Available
Deprecated
Arquitectura
Mixture of Experts
MoE architecture
Calibrado
No
Sí
Mezcla de expertos
Sí
Sí
Parámetros totales
30B
100B
Parámetros activados
3.3B
6.1B
Razonamiento
No
No
Precisión
FP8
FP8
Longitud del contexto
262K
131K
Tokens máximos
262K
131K
Funcionalidades admitidas
Serverless
Compatible
Compatible
Serverless LoRA
No compatible
No compatible
Ajuste fino
No compatible
No compatible
Embeddings
No compatible
No compatible
Rerankers
No compatible
No compatible
Soporte para Input de Image
No compatible
No compatible
JSON Mode
Compatible
No compatible
Outputs estructuradas
No compatible
No compatible
Herramientas
Compatible
No compatible
Finalización de Fim
No compatible
No compatible
Completado de prefijo de Chat
No compatible
Compatible
Qwen3-30B-A3B-Instruct-2507 en comparación
Ver cómo Qwen3-30B-A3B-Instruct-2507 se compara con otros modelos populares en dimensiones clave.
vs.

Qwen3.6-27B
vs.

Qwen3.6-35B-A3B
vs.
gemma-4-26B-A4B-it
vs.
gemma-4-31B-it
vs.

Qwen3.5-27B
vs.

Qwen3.5-35B-A3B
vs.

Qwen3-VL-32B-Instruct
vs.

Qwen3-VL-32B-Thinking
vs.

Qwen3-VL-8B-Instruct
vs.

Qwen3-VL-8B-Thinking
vs.

Qwen3-VL-30B-A3B-Instruct
vs.

Qwen3-VL-30B-A3B-Thinking
vs.

Qwen3-Omni-30B-A3B-Instruct
vs.

Ring-flash-2.0
vs.

Qwen3-Omni-30B-A3B-Captioner
vs.

Qwen3-Omni-30B-A3B-Thinking
vs.

Qwen3-Next-80B-A3B-Instruct
vs.

Qwen3-Next-80B-A3B-Thinking
vs.
gpt-oss-120b
vs.
gpt-oss-20b
