Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    AI Models Directory

    Browse and compare 200+ AI models from OpenAI, Anthropic, Google, and 40+ providers — filter by capabilities, pricing, and context size.

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    257
    Models
    55
    Providers
    133
    Vision Models
    162
    Tool-enabled
    0
    Free Models
    Features
    DeepInfra
    qwen3-vl-30b-a3b-instruct
    $0.15$0.60—
    NovitaAI
    qwen3-vl-30b-a3b-instruct
    $0.20$0.70—
    NovitaAI
    qwen3-235b-a22b-fp8
    $0.20$0.80—
    Moonshot AI
    kimi-k2.5
    $0.60$3.00$0.10
    EmberCloud
    kimi-k2.5
    $0.40$1.98$0.22
    Alibaba Cloud(cn-beijing)
    kimi-k2.5
    $0.57$3.01—
    DeepInfra
    kimi-k2.5
    $0.45$2.25$0.07
    Alibaba Cloud(us-virginia)
    kimi-k2.5
    $0.57$3.01—
    Alibaba Cloud(eu-frankfurt)
    kimi-k2.5
    $0.57$3.01—
    Alibaba Cloud
    kimi-k2.5
    $0.57$3.01—
    NovitaAI
    llama-3.2-3b-instruct
    $0.03$0.05—
    NovitaAI
    llama-3-70b-instruct
    $0.51$0.74—
    Alibaba Cloud
    qwen-image-edit-max
    $0.080/req——
    Alibaba Cloud
    qwen-image-edit-plus
    $0.040/req——
    NovitaAI
    qwen3-vl-235b-a22b-thinking
    $0.98$3.95—
    DeepInfra
    qwen3-vl-235b-a22b-instruct
    $0.20$0.88$0.11
    NovitaAI
    qwen3-vl-235b-a22b-instruct
    $0.30$1.50—
    Alibaba Cloud(singapore)
    qwen3-vl-flash
    $0.05$0.40$0.01
    Alibaba Cloud(cn-beijing)
    qwen3-vl-flash
    $0.02$0.21$0.00
    Alibaba Cloud(us-virginia)
    qwen3-vl-flash
    $0.02$0.21$0.00
    Alibaba Cloud
    qwen3-vl-flash
    $0.05$0.40$0.01
    Alibaba Cloud(eu-frankfurt)
    qwen3-vl-plus
    $0.20$1.60$0.04
    Alibaba Cloud(singapore)
    qwen3-vl-plus
    $0.20$1.60$0.04
    Alibaba Cloud(us-virginia)
    qwen3-vl-plus
    $0.14$1.43$0.03
    Alibaba Cloud(cn-beijing)
    qwen3-vl-plus
    $0.14$1.43$0.03
    Alibaba Cloud
    qwen3-vl-plus
    $0.20$1.60$0.04
    Alibaba Cloud(cn-beijing)
    qwen3-coder-flash
    $0.14$0.57$0.03
    Alibaba Cloud(eu-frankfurt)
    qwen3-coder-flash
    $0.30$1.50$0.06
    Alibaba Cloud
    qwen3-coder-flash
    $0.30$1.50$0.06
    Alibaba Cloud(singapore)
    qwen3-coder-flash
    $0.30$1.50$0.06
    Alibaba Cloud(us-virginia)
    qwen3-coder-flash
    $0.14$0.57$0.03
    Alibaba Cloud(cn-beijing)
    qwen-coder-plus
    $0.50$1.00—
    Alibaba Cloud
    qwen-coder-plus
    $0.50$1.00—
    Alibaba Cloud(singapore)
    qwen-coder-plus
    $0.50$1.00—
    MiniMax
    minimax-text-01
    $0.20$1.10—
    MiniMax
    minimax-m2.1-lightning
    $0.12$0.48—
    EmberCloud
    glm-4.7-flash
    $0.06$0.40$0.01
    Z AI
    glm-4.7-flashx
    $0.07$0.40$0.01
    Z AI
    glm-image
    $0.015/req——
    ByteDance
    seedream-4-5
    $0.045/req——
    ByteDance
    seedream-4-0
    $0.035/req——
    ByteDance
    seed-1-8-251228
    $0.25$2.00$0.05
    ByteDance
    seed-1-6-flash-250715
    $0.07$0.30$0.01
    ByteDance
    seed-1-6-250915
    $0.25$2.00$0.05
    ByteDance
    seed-1-6-250615
    $0.25$2.00$0.05
    OpenAI
    gpt-4o-mini-search-preview
    $0.15$0.60—
    OpenAI
    gpt-4o-search-preview
    $2.50$10.00—
    Z AI
    cogview-4
    $0.010/req——
    Alibaba Cloud
    qwen-image-max-2025-12-30
    $0.075/req——
    Alibaba Cloud
    qwen-image
    $0.035/req——
    Page 8 of 12

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Partners
    • Lounge
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Glacier
    • Iceberg
    • Granite
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Quartz
    • Avalanche
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Z AI
    • Moonshot AI
    • Baidu
    • Permafrost
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • Custom
    • NanoGPT
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Sakana AI
    • Tundra
    • Xiaomi
    • DeepInfra
    • Reve
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI

    © 2026 LLM Gateway. All rights reserved.

    Browse models by use case

    • Best models for coding
    • Reasoning models
    • Best models for roleplay
    • Creative writing models
    • Translation models
    • Best models for math
    • Long context models
    • Cheapest models
    • Premium models
    • Open source models
    • Vision models
    • Tool-calling models
    • Web search models
    • Embedding models
    • Text generation models
    • Text-to-image models
    • Image editing models
    • Video generation models
    • Discounted models

    How to choose an AI model

    Start from the capability you need — reasoning, vision, tool calling, or long context — then compare price per million tokens and context window. The filters above narrow the directory, and each model's page lists provider availability, live pricing, and uptime. Not sure where to start? See which models developers actually run in production in the live rankings.

    Compare AI model pricing

    Prices are shown per million input and output tokens, exactly as providers publish them. Sort by price to find the cheapest models, or estimate a monthly bill for your traffic with the token cost calculator.

    Try a model before you integrate

    Every model here is callable through one OpenAI-compatible API — switch models by changing a single string. Chat with any of them first in the Lounge to compare quality, speed, and cost side by side.

    is now on LLM Gateway — 30% off open-source modelsends in 7d 00:50:18
    LLM Gateway
    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started
    1.6k