gemma-4-31B-it
About gemma-4-31B-it
Gemma 4 31B is Google DeepMind's latest open-source model, built on a 31B dense architecture from the same research foundation as Gemini 3. Purpose-built for advanced reasoning and agentic workflows, it ranks #3 among all open models on the Arena AI leaderboard — outperforming models up to 20x its size — with native function-calling, 256K context, and full Apache 2.0 licensing.
Available Serverless
Run queries immediately, pay only for usage
Input Price
$
0.13
/ M Tokens
Output Price
$
0.4
/ M Tokens
Metadata
Specification
State
Available
Architecture
Dense
Calibrated
No
Mixture of Experts
No
Total Parameters
31B
Activated Parameters
30.7B
Reasoning
No
Precision
FP8
Context length
262K
Max Tokens
262K
Supported Functionality
Serverless
Supported
Serverless LoRA
Not supported
Fine-tuning
Not supported
Embeddings
Not supported
Rerankers
Not supported
Support image input
Supported
JSON Mode
Supported
Structured Outputs
Not supported
Tools
Supported
Fim Completion
Not supported
Chat Prefix Completion
Not supported
Compare with Other Models
See how this model stacks up against others.
