GPT OSS 120B

Open-source 120B parameter model with reasoning capabilities via Groq inference.

gpt-oss-120b
STABLEGet StartedView uptime
131,072 context
Released August 5, 2025
Starting at $0.03/M input tokens
Starting at $0.14/M output tokens
Streaming
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

All Providers for GPT OSS 120B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

SCX.aiUp to 4x faster
Context: 131.1kQuant: fp8
Input
$0.17
/M tokens
Cache Read
/M tokens
Output
$0.55
/M tokens
Get Started
Groq
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.75
/M tokens
Get Started
Cerebras
Context: 131.1k
Input
$0.35
/M tokens
Cache Read
/M tokens
Output
$0.75
/M tokens
Get Started
NanoGPT
Context: 131.1k
Input
$0.05
/M tokens
Cache Read
/M tokens
Output
$0.25
/M tokens
Get Started
ByteDance
Context: 128k
Input
$0.1
/M tokens
Cache Read
$0.02
/M tokens
Output
$0.5
/M tokens
Get Started
Nebius AI
Context: 131.1kQuant: fp4
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Together AI
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Azure
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Runware
Context: 131.1k
Input
$0.032
/M tokens
Cache Read
$0.032
/M tokens
Output
$0.14
/M tokens
Get Started

Frequently asked questions

What is GPT OSS 120B?

Open-source 120B parameter model with reasoning capabilities via Groq inference. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does GPT OSS 120B cost?

Pricing for GPT OSS 120B on LLM Gateway starts at $0.03 per million input tokens and $0.14 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of GPT OSS 120B?

GPT OSS 120B supports a context window of up to 131,072 tokens on its largest provider deployment.

Which providers serve GPT OSS 120B?

GPT OSS 120B is served by Groq, Cerebras, NanoGPT, ByteDance, Nebius AI, Together AI, Azure, Runware, SCX.ai through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does GPT OSS 120B support tool calling and structured outputs?

Yes. GPT OSS 120B supports both tool (function) calling and structured JSON outputs through LLM Gateway.

When was GPT OSS 120B released?

GPT OSS 120B was released on August 5, 2025.