Llama 3 8B Instruct

Llama 3 8B instruction-following model.

llama-3-8b-instruct
STABLEModel DeactivatedGet StartedView uptime
8,192 context
Released April 3, 2025
Starting at $0.04/M input tokens
Starting at $0.04/M output tokens
Streaming
JSON Output
No ratings yetSign in to rate

Select Provider

All Providers for Llama 3 8B Instruct

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

NovitaAI
Context: 8.2k
Deactivated since Jul 11, 2026
Input
$0.04
/M tokens
Cache Read
/M tokens
Output
$0.04
/M tokens
Get Started

Frequently asked questions

What is Llama 3 8B Instruct?

Llama 3 8B instruction-following model. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Llama 3 8B Instruct cost?

Pricing for Llama 3 8B Instruct on LLM Gateway starts at $0.04 per million input tokens and $0.04 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Llama 3 8B Instruct?

Llama 3 8B Instruct supports a context window of up to 8,192 tokens on its largest provider deployment.

Which providers serve Llama 3 8B Instruct?

Llama 3 8B Instruct is served by NovitaAI through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Llama 3 8B Instruct support tool calling and structured outputs?

Llama 3 8B Instruct supports structured JSON outputs, but not tool calling.

When was Llama 3 8B Instruct released?

Llama 3 8B Instruct was released on April 3, 2025.