Gemma 3 12B IT

Medium 12B Gemma 3 model balancing size and capability.

gemma-3-12b-it
STABLEModel DeactivatedGet StartedView uptime
1,000,000 context
Released March 10, 2025
Starting at $0.07/M input tokens
Starting at $0.30/M output tokens
Streaming
No ratings yetSign in to rate

Select Provider

All Providers for Gemma 3 12B IT

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Google AI Studio
Context: 1M
Deactivated since Apr 30, 2026
Input
$0.075
/M tokens
Cache Read
/M tokens
Output
$0.3
/M tokens
Get Started

Frequently asked questions

What is Gemma 3 12B IT?

Medium 12B Gemma 3 model balancing size and capability. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Gemma 3 12B IT cost?

Pricing for Gemma 3 12B IT on LLM Gateway starts at $0.07 per million input tokens and $0.30 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Gemma 3 12B IT?

Gemma 3 12B IT supports a context window of up to 1,000,000 tokens on its largest provider deployment.

Which providers serve Gemma 3 12B IT?

Gemma 3 12B IT is served by Google AI Studio through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Gemma 3 12B IT support tool calling and structured outputs?

No. Gemma 3 12B IT does not currently support tool calling or structured JSON outputs through LLM Gateway.

When was Gemma 3 12B IT released?

Gemma 3 12B IT was released on March 10, 2025.