Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    is now on LLM Gateway — 30% off open-source modelsends in 19:37:16
    LLM Gateway
    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started

    AI Models Directory

    Browse and compare 200+ AI models from OpenAI, Anthropic, Google, and 40+ providers — filter by capabilities, pricing, and context size.

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Price per unit
    Output Price ($/M tokens)
    Context Size (tokens)
    129/263
    Models
    38/59
    Providers
    94
    Vision Models (filtered)
    124
    Tool-enabled (filtered)
    0
    Free Models (filtered)

    GPT-6 Astra

    openaiPremium
    gpt-6-astra
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    Structured JSON
    Web Search
    Azure
    Context: 1.1M
    Input
    $10.00
    /M tokens
    Cached
    $1.00
    /M tokens
    Output
    $50.00
    /M tokens
    Tiered Pricing
    IN
    CACHED
    OUT
    ≤272K tokens
    $10.00
    $1.00
    $50.00
    >272K tokens
    $20.00
    $2.00
    $75.00
    + $0.010 per search
    Get Started

    Muse Spark 1.2 Contributor

    meta
    muse-spark-1.2-contributor
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    Structured JSON
    Meta Contributor
    Context: 1.0M
    Input
    $0.10
    /M tokens
    Cached
    $0.00
    /M tokens
    Output
    $0.20
    /M tokens
    Get Started

    Muse Spark 1.3 Contributor

    meta
    muse-spark-1.3-contributor
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    Structured JSON
    Meta Contributor
    Context: 1.0M
    Input
    $0.10
    /M tokens
    Cached
    $0.00
    /M tokens
    Output
    $0.20
    /M tokens
    Get Started

    Muse Spark 1.3

    meta
    muse-spark-1.3
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    Structured JSON
    Meta
    Context: 1.0M
    Input
    $1.25
    /M tokens
    Cached
    $0.15
    /M tokens
    Output
    $4.25
    /M tokens
    Get Started

    Qwen3.8 27B

    Qwen3.8 27B
    Qwen3.8-27B
    Streaming
    Vision
    Tools
    Reasoning
    Consensus Protocol
    Context: 32.8k
    Input
    $0.20
    /M tokens
    Cached
    $0.05
    /M tokens
    Output
    $2.00
    /M tokens
    Get Started

    Gemini 3.8 Flash

    google
    gemini-3.8-flash
    Streaming
    Vision
    Tools
    Reasoning
    Reasoning Budget
    JSON Output
    Structured JSON
    Web Search
    Google AI Studio
    Context: 1.0M
    Input
    $0.75
    /M tokens
    Cached
    $0.07
    /M tokens
    Output
    $3.75
    /M tokens
    + $0.014 per search
    Get Started

    Claude Fable 5.1

    anthropicPremium
    claude-fable-5-1
    Streaming
    Vision
    Tools
    Reasoning
    Structured JSON
    Anthropic
    Context: 1M
    Input
    $10.00
    /M tokens
    Cached
    $0.25
    /M tokens
    Output
    $50.00
    /M tokens
    Get Started

    Qwen3.8 27B

    alibaba
    qwen3.8-27b
    Streaming
    Tools
    Reasoning
    JSON Output
    NovitaAI
    Context: 1M
    Input
    $0.42
    /M tokens
    Cached
    $0.08
    /M tokens
    Output
    $3.00
    /M tokens
    Get Started

    Qwen3.8 Flash

    alibaba
    qwen3.8-flash
    Streaming
    Tools
    Reasoning
    JSON Output
    Structured JSON
    Vision
    Reasoning Budget
    Web Search
    NovitaAI
    Context: 1M
    Input
    $0.15
    /M tokens
    Cached
    $0.02
    /M tokens
    Output
    $0.47
    /M tokens
    Get Started

    DeepSeek V4 Flash Vision Exp

    deepseek
    deepseek-v4-flash-vision-exp
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    DeepSeek
    Context: 1.1M
    Time-based pricing
    Input
    $0.44
    /M tokens
    Cached
    $0.01
    /M tokens
    Output
    $1.32
    /M tokens

    Peak: Monday–Friday, 09:00–12:00 and 14:00–18:00 Beijing time.

    Off-peak: Monday–Friday outside those hours, plus all day Saturday and Sunday.

    Get Started

    Muse Image 1.0

    meta
    muse-image-1.0
    Vision
    Reasoning
    Image Generation
    Meta
    Context: —
    Per image
    $0.01000
    Input
    $0.00
    /M tokens
    Cached
    —
    /M tokens
    Output
    $0.00
    /M tokens
    Get Started

    GLM-5.3 Flash

    zai50% off
    glm-5.3-flash
    Streaming
    Vision
    Tools
    Reasoning
    JSON Output
    Structured JSON
    NovitaAI
    Context: 1.0MQuant: fp850% off
    Input
    $0.15$0.07
    -50% off
    /M tokens
    Cached
    $0.03$0.01
    -50% off
    /M tokens
    Output
    $0.50$0.25
    -50% off
    /M tokens
    Get Started
    Page 1 of 11

    Browse models by use case

    • Best models for coding
    • Reasoning models
    • Best models for roleplay
    • Creative writing models
    • Translation models
    • Best models for math
    • Long context models
    • Cheapest models
    • Premium models
    • Open source models
    • Vision models
    • Tool-calling models
    • Web search models
    • Embedding models
    • Text generation models
    • Text-to-image models
    • Image editing models
    • Video generation models
    • Discounted models

    How to choose an AI model

    Start from the capability you need — reasoning, vision, tool calling, or long context — then compare price per million tokens and context window. The filters above narrow the directory, and each model's page lists provider availability, live pricing, and uptime. Not sure where to start? See which models developers actually run in production in the live rankings.

    Compare AI model pricing

    Prices are shown per million input and output tokens, exactly as providers publish them. Sort by price to find the cheapest models, or estimate a monthly bill for your traffic with the token cost calculator.

    Try a model before you integrate

    Every model here is callable through one OpenAI-compatible API — switch models by changing a single string. Chat with any of them first in the Lounge to compare quality, speed, and cost side by side.

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Partners
    • Lounge
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Developer resources
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • About
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Azure Anthropic
    • Z AI
    • Moonshot AI
    • Baidu
    • Perplexity
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Meta Contributor
    • Sakana AI
    • Xiaomi
    • DeepInfra
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI
    • Consensus Protocol

    © 2026 LLM Gateway. All rights reserved.

    1.6k