Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    Best Models for Coding

    Frontier and open-weight models for code generation, review, and agentic coding — compared by price and context window

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    41/297
    Models
    35/64
    Providers
    28
    Vision Models (filtered)
    39
    Tool-enabled (filtered)
    0
    Free Models (filtered)
    Features
    Google AI Studio
    gemini-pro-latest
    $2.00$12.00$0.20
    Azure
    gpt-5.3-codex
    $1.75$14.00$0.17
    OpenAI
    gpt-5.3-codex
    $1.75$14.00$0.17
    Azure
    gpt-5.2-codex
    $1.75$14.00$0.17
    Azure
    gpt-5.4
    $2.50$15.00$0.25
    OpenAI
    gpt-5.4
    $2.50$15.00$0.25
    Mistral AI
    devstral-2512
    $0.40$2.00—
    Mistral AI
    codestral-2508
    $0.30$0.90—
    Quartz
    gemini-3.1-pro-preview
    $2.00$12.00$0.20
    Google Vertex AI
    gemini-3.1-pro-preview
    $2.00$12.00$0.20
    Google AI Studio
    gemini-3.1-pro-preview
    $2.00$12.00$0.20
    AWS Bedrock(eu-west-2)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    AWS Bedrock(jp)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    AWS Bedrock(au)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    AWS Bedrock(eu)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    AWS Bedrock(us)
    claude-sonnet-4-6
    $3.30$16.50$0.33
    AWS Bedrock(global)
    claude-sonnet-4-6
    $3.00$15.00$0.30
    Vertex AI (Anthropic)
    claude-sonnet-4-6
    $3.00$15.00$0.30
    AWS Bedrock
    claude-sonnet-4-6
    $3.00$15.00$0.30
    Anthropic
    claude-sonnet-4-6
    $3.00$15.00$0.30
    Alibaba Cloud(eu-frankfurt)
    qwen3-coder-flash
    $0.30$1.50$0.06
    Alibaba Cloud(cn-beijing)
    qwen3-coder-flash
    $0.14$0.57$0.03
    Alibaba Cloud(us-virginia)
    qwen3-coder-flash
    $0.14$0.57$0.03
    Alibaba Cloud(singapore)
    qwen3-coder-flash
    $0.30$1.50$0.06
    Alibaba Cloud
    qwen3-coder-flash
    $0.30$1.50$0.06
    Vertex AI (OpenAI-compatible)
    glm-4.7
    $0.60$2.20—
    Together AI
    glm-4.7
    $0.45$2.00—
    EmberCloud
    glm-4.7
    $0.38$1.98$0.19
    ByteDance
    glm-4.7
    $0.60$2.20$0.11
    NovitaAI
    glm-4.7
    $0.60$2.20$0.11
    Cerebras
    glm-4.7
    $2.25$2.75—
    Z AI
    glm-4.7
    $0.60$2.20$0.11
    Vertex AI (OpenAI-compatible)
    deepseek-v3.2
    $0.56$1.68$0.06
    DeepInfra
    deepseek-v3.2
    $0.26$0.38$0.13
    NovitaAI
    deepseek-v3.2
    $0.27$0.40$0.13
    ByteDance
    deepseek-v3.2
    $0.28$0.42$0.06
    Azure
    gpt-5.1-codex-mini
    $0.25$2.00$0.02
    Azure
    gpt-5.1-codex
    $1.25$10.00—
    AWS Bedrock(jp)
    claude-haiku-4-5
    $1.10$5.50$0.11
    AWS Bedrock(au)
    claude-haiku-4-5
    $1.10$5.50$0.11
    AWS Bedrock(apac)
    claude-haiku-4-5
    $1.00$5.00$0.10
    AWS Bedrock(eu)
    claude-haiku-4-5
    $1.10$5.50$0.11
    AWS Bedrock(us)
    claude-haiku-4-5
    $1.10$5.50$0.11
    AWS Bedrock(global)
    claude-haiku-4-5
    $1.00$5.00$0.10
    Vertex AI (Anthropic)
    claude-haiku-4-5
    $1.00$5.00$0.10
    AWS Bedrock
    claude-haiku-4-5
    $1.00$5.00$0.10
    Anthropic
    claude-haiku-4-5
    $1.00$5.00$0.10
    NovitaAI
    qwen3-coder-30b-a3b-instruct
    $0.07$0.27—
    Vertex AI (OpenAI-compatible)
    qwen3-coder-480b-a35b-instruct
    $0.22$1.80$0.02
    NovitaAI
    qwen3-coder-480b-a35b-instruct
    $0.38$1.55—
    Page 4 of 5

    Newsletter

    Stay ahead of the curve

    Insights on LLM routing, new model launches and cost optimization, sent to your inbox.

    • New models and providers, rounded up
    • Tips to cut LLM costs
    • Major product launches

    No spam. Unsubscribe anytime.

    System status
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

  1. Features
  2. AI Gateway
  3. Observability
  4. Models
  5. Providers
  6. Rankings
  7. Add Provider
  8. Partners
  9. Lounge
  10. Changelog
  11. DevPass
  12. Compare Models
  13. Enterprise
  14. Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Developer resources
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • About
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Microsoft Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • Runpod
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Azure Anthropic
    • Z AI
    • Moonshot AI
    • Baidu
    • Perplexity
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Meta Contributor
    • Sakana AI
    • Xiaomi
    • DeepInfra
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI
    • Consensus Protocol
    • Tencent Cloud
    • Atria
    • TypeSafe AI

    © 2026 LLM Gateway. All rights reserved.

    The best coding models combine strong code generation with reliable tool calling, since modern coding agents lean on function calls to edit files and run commands. This page tracks the models developers actually ship with: Anthropic's Claude series, OpenAI's Codex line, Google's Gemini Pro, and fast-improving open-weight options like Qwen3 Coder, Kimi K2.7 Code, GLM, and DeepSeek.

    Access every one of them through a single OpenAI-compatible API with automatic failover, so your coding agent, IDE plugin, or CI pipeline can switch between frontier and budget models without code changes — and you can track exactly what each tool spends.

    Frequently asked questions

    What is the best LLM for coding?

    Claude Sonnet 5 and Claude Opus 4.8 lead most real-world coding evaluations, with OpenAI's GPT-5.3 Codex and Google's Gemini 3.1 Pro close behind. Among open-weight models, Qwen3 Coder, Kimi K2.7 Code, GLM-5.2, and DeepSeek V4 deliver strong results at a fraction of the price.

    What is the cheapest model that is still good at coding?

    Qwen3 Coder 30B (about $0.07 per million input tokens), GLM-4.7 Flash, and DeepSeek V4 Flash are the standouts for budget coding. They handle everyday completion, refactoring, and code review well and are cheap enough to run on every commit.

    Can I use these models with coding agents like Cline or Aider?

    Yes. Any agent that supports an OpenAI-compatible endpoint or custom base URL can route through LLM Gateway with one API key — that includes Cline, Aider, Roo Code, and devpass-code — so you can mix models per task and see per-agent cost analytics.

    Do coding models need tool calling?

    For agentic workflows, yes. Editing files, running tests, and searching a repo all happen through function calls, so pick a model with reliable tool calling — the capability icons in the list above show which provider mappings support tools. For plain autocomplete or one-shot generation, tool calling is optional.

    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started
    1.7k