Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    is now on LLM Gateway — 30% off open-source modelsends in 19d 11:25:30
    LLM Gateway
    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    1.5k
    Log InGet Started

    AI Models Directory

    Browse and compare 200+ AI models from OpenAI, Anthropic, Google, and 40+ providers — filter by capabilities, pricing, and context size.

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    248
    Models
    47
    Providers
    129
    Vision Models
    157
    Tool-enabled
    1
    Free Models
    Features
    Nebius AI
    nemotron-3-nano-omni
    $0.06$0.24—
    Nebius AI
    nemotron-3-nano-30b
    $0.06$0.24—
    Nebius AI
    nemotron-3-super-120b
    $0.30$0.90—
    DeepInfra
    nemotron-3-ultra-550b
    $0.50$2.50$0.15
    Nebius AI
    nemotron-3-ultra-550b
    $1.00$3.00—
    DeepInfra
    hy3
    $0.14$0.58$0.04
    NovitaAI
    hy3
    $0.14$0.58$0.04
    Sakana AI
    fugu-ultra
    $5.00$30.00$0.50
    Reve
    reve-create
    $0.024/req——
    Nebius AI
    hermes-4-70b
    $0.13$0.40—
    Nebius AI
    hermes-4-405b
    $1.00$3.00—
    ByteDance
    seedream-5-0-pro
    $0.090/req——
    ByteDance
    seedream-5-0-lite
    $0.035/req——
    ByteDance
    seedream-4-5
    $0.045/req——
    ByteDance
    seedance-1-5-pro
    $0.02592/sec$0.1166/sec—
    ByteDance
    seedance-2-0-mini
    $0.0378/sec$0.0756/sec—
    ByteDance
    seedance-2-0-fast
    $0.121/sec$0.2722/sec—
    ByteDance
    seedance-2-0
    $0.1512/sec$0.3402/sec—
    ByteDance
    seedream-4-0
    $0.035/req——
    ByteDance
    seed-1-8-251228
    $0.25$2.00$0.05
    ByteDance
    seed-1-6-flash-250715
    $0.07$0.30$0.01
    ByteDance
    seed-1-6-250915
    $0.25$2.00$0.05
    ByteDance
    seed-1-6-250615
    $0.25$2.00$0.05
    DeepInfra
    bge-m3
    $0.01$0.00—
    AtlasCloud
    kling-v3-0-turbo
    $0.168/sec$0.21/sec—
    AtlasCloud
    kling-v3-0
    $0.084/sec$0.42/sec—
    DeepInfra
    qwen3-reranker-0.6b
    $0.01$0.00—
    DeepInfra
    qwen3-reranker-4b
    $0.03$0.00—
    Alibaba Cloud
    qwen-audio-3.0-tts-flash
    $0.015/1K chars——
    DeepInfra
    qwen3-reranker-8b
    $0.05$0.00—
    Alibaba Cloud
    qwen-audio-3.0-tts-plus
    $0.02/1K chars——
    Alibaba Cloud
    wan-2-6-t2v
    $0.00$0.00—
    DeepInfra
    qwen3-embedding-8b
    $0.01$0.00—
    Nebius AI
    qwen3-embedding-8b
    $0.01$0.00—
    Alibaba Cloud(singapore)
    qwen3.6-flash
    $0.25$1.50$0.05
    Alibaba Cloud
    qwen3.6-flash
    $0.25$1.50$0.05
    Alibaba Cloud(cn-beijing)
    qwen3.6-flash
    $0.17$0.99$0.03
    Alibaba Cloud(us-virginia)
    qwen3.6-flash
    $0.17$0.99$0.03
    Alibaba Cloud
    qwen3.6-35b-a3b
    $0.25$1.48—
    NovitaAI
    qwen3.6-35b-a3b
    $0.25$1.48—
    Alibaba Cloud(singapore)
    qwen3.6-35b-a3b
    $0.25$1.48—
    Alibaba Cloud(singapore)
    qwen3.6-plus
    $0.50$3.00$0.05
    Alibaba Cloud
    qwen3.6-plus
    $0.50$3.00$0.05
    Alibaba Cloud(singapore)
    qwen3.6-max-preview
    $1.30$7.80$0.13
    Alibaba Cloud
    qwen3.6-max-preview
    $1.30$7.80$0.13
    Alibaba Cloud
    qwen-image-edit-max
    $0.080/req——
    Alibaba Cloud
    qwen-image-edit-plus
    $0.040/req——
    Alibaba Cloud
    qwen2-5-vl-32b-instruct
    $1.40$4.20—
    NovitaAI
    qwen3-vl-235b-a22b-thinking
    $0.98$3.95—
    DeepInfra
    qwen3-vl-235b-a22b-instruct
    $0.20$0.88$0.11
    Page 2 of 11

    Browse models by use case

    • Best models for coding
    • Reasoning models
    • Best models for roleplay
    • Creative writing models
    • Translation models
    • Best models for math
    • Long context models
    • Cheapest models
    • Premium models
    • Open source models
    • Vision models
    • Tool-calling models
    • Web search models
    • Embedding models
    • Text generation models
    • Text-to-image models
    • Image editing models
    • Video generation models
    • Discounted models

    How to choose an AI model

    Start from the capability you need — reasoning, vision, tool calling, or long context — then compare price per million tokens and context window. The filters above narrow the directory, and each model's page lists provider availability, live pricing, and uptime.

    Compare AI model pricing

    Prices are shown per million input and output tokens, exactly as providers publish them. Sort by price to find the cheapest models, or estimate a monthly bill for your traffic with the token cost calculator.

    Try a model before you integrate

    Every model here is callable through one OpenAI-compatible API — switch models by changing a single string. Chat with any of them first in the Lounge to compare quality, speed, and cost side by side.

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Lounge
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Contact Us

    Community

    • Twitter
    • Discord

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • GDPR
    • SOC 2 Type II
    • Status

    Compare

    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Glacier
    • Iceberg
    • Granite
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Quartz
    • Avalanche
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Z AI
    • Moonshot AI
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai
    • Custom
    • NanoGPT
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Sakana AI
    • Tundra
    • Xiaomi
    • DeepInfra
    • Reve
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI

    © 2026 LLM Gateway. All rights reserved.