Gemma 4 31B IT

Large 31B Gemma 4 instruction-tuned model with reasoning.

gemma-4-31b-it
STABLEGet StartedView uptime
262,144 context
Released April 2, 2026
Starting at $0.10/M input tokens
Starting at $0.30/M output tokens
Streaming
Vision
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

All Providers for Gemma 4 31B IT

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

SCX.aiUp to 4x faster
Context: 131.1kQuant: bf16
Input
$0.3
/M tokens
Cache Read
/M tokens
Output
$0.91
/M tokens
Get Started
DeepInfra
Context: 262.1kQuant: fp8
Input
$0.13
/M tokens
Cache Read
/M tokens
Output
$0.38
/M tokens
Get Started
Runware
Context: 262.1k
Input
$0.102
/M tokens
Cache Read
$0.012
/M tokens
Output
$0.297
/M tokens
Get Started
NovitaAI
Context: 262.1kQuant: bf16
Input
$0.13
/M tokens
Cache Read
/M tokens
Output
$0.38
/M tokens
Get Started
Together AI
Context: 262.1k
Input
$0.13
/M tokens
Cache Read
/M tokens
Output
$0.38
/M tokens
Get Started
Cerebras
Context: 131.1k
Input
$0.99
/M tokens
Cache Read
/M tokens
Output
$1.49
/M tokens
Get Started

Frequently asked questions

What is Gemma 4 31B IT?

Large 31B Gemma 4 instruction-tuned model with reasoning. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Gemma 4 31B IT cost?

Pricing for Gemma 4 31B IT on LLM Gateway starts at $0.10 per million input tokens and $0.30 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Gemma 4 31B IT?

Gemma 4 31B IT supports a context window of up to 262,144 tokens on its largest provider deployment.

Which providers serve Gemma 4 31B IT?

Gemma 4 31B IT is served by DeepInfra, Runware, NovitaAI, Together AI, Cerebras, SCX.ai through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Gemma 4 31B IT support tool calling and structured outputs?

Yes. Gemma 4 31B IT supports both tool (function) calling and structured JSON outputs through LLM Gateway.

When was Gemma 4 31B IT released?

Gemma 4 31B IT was released on April 2, 2026.