Qwen3 Max 2026-01-23

Latest Qwen 3 flagship model with integrated thinking mode and tool support.

qwen3-max-2026-01-23
STABLEModel DeactivatedGet StartedView uptime
262,144 context
Released January 23, 2026
Starting at $0.36/M input tokens (tiered)
Starting at $1.43/M output tokens (tiered)
Streaming
Vision
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

Select Provider

All Providers for Qwen3 Max 2026-01-23

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Alibaba Cloud
Context: 262.1k
Deactivated since Jul 8, 2026
Input
$1.2
/M tokens
Cache Read
$0.24
/M tokens
Output
$6
/M tokens
Cache Write 5m
$1.5
/M tokens
Cache Write 1h
$1.5
/M tokens
Get Started

Frequently asked questions

What is Qwen3 Max 2026-01-23?

Latest Qwen 3 flagship model with integrated thinking mode and tool support. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Qwen3 Max 2026-01-23 cost?

Pricing for Qwen3 Max 2026-01-23 on LLM Gateway starts at $0.36 per million input tokens and $1.43 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Qwen3 Max 2026-01-23?

Qwen3 Max 2026-01-23 supports a context window of up to 262,144 tokens on its largest provider deployment.

Which providers serve Qwen3 Max 2026-01-23?

Qwen3 Max 2026-01-23 is served by Alibaba Cloud through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Qwen3 Max 2026-01-23 support tool calling and structured outputs?

Yes. Qwen3 Max 2026-01-23 supports both tool (function) calling and structured JSON outputs through LLM Gateway.

When was Qwen3 Max 2026-01-23 released?

Qwen3 Max 2026-01-23 was released on January 23, 2026.