Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    Best Models for Translation

    Multilingual models with strong translation quality across major and low-resource languages — compared by price and context

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    17/258
    Models
    29/55
    Providers
    14
    Vision Models (filtered)
    16
    Tool-enabled (filtered)
    0
    Free Models (filtered)
    Features
    Google Vertex AI
    gemini-3.6-flash
    $0.75$3.75$0.07
    Google AI Studio
    gemini-3.6-flash
    $0.75$3.75$0.07
    Iceberg
    gemini-3.6-flash
    $0.75$3.75$0.07
    AWS Mantle(us-west-2)
    gpt-5.6-luna
    $0.22$1.32$0.02
    AWS Mantle(us-east-2)
    gpt-5.6-luna
    $0.22$1.32$0.02
    OpenAI
    gpt-5.6-luna
    $0.20$1.20$0.02
    AWS Mantle
    gpt-5.6-luna
    $0.22$1.32$0.02
    Azure
    gpt-5.6-luna
    $0.20$1.20$0.02
    AWS Mantle(us-east-1)
    gpt-5.6-luna
    $0.22$1.32$0.02
    OpenAI
    gpt-5.6-terra
    $2.00$12.00$0.20
    AWS Mantle(us-east-1)
    gpt-5.6-terra
    $2.20$13.20$0.22
    Azure
    gpt-5.6-terra
    $2.00$12.00$0.20
    AWS Mantle
    gpt-5.6-terra
    $2.20$13.20$0.22
    AWS Mantle(us-east-2)
    gpt-5.6-terra
    $2.20$13.20$0.22
    AWS Mantle(us-west-2)
    gpt-5.6-terra
    $2.20$13.20$0.22
    AWS Bedrock(global)
    claude-sonnet-5
    $2.00$10.00$0.20
    AWS Bedrock
    claude-sonnet-5
    $2.00$10.00$0.20
    AWS Bedrock(us)
    claude-sonnet-5
    $2.20$11.00$0.22
    Vertex AI (Anthropic)
    claude-sonnet-5
    $2.00$10.00$0.20
    Anthropic
    claude-sonnet-5
    $2.00$10.00$0.20
    Cerebras
    gemma-4-31b-it
    $0.99$1.49—
    Together AI
    gemma-4-31b-it
    $0.39$0.97—
    DeepInfra
    gemma-4-31b-it
    $0.13$0.38—
    Runware
    gemma-4-31b-it
    $0.10$0.07
    -30% off
    $0.30$0.21
    -30% off
    $0.01$0.01
    -30% off
    SCX.ai (Turbo)
    gemma-4-31b-it
    $0.30$0.91—
    NovitaAI
    gemma-4-31b-it
    $0.14$0.40—
    Alibaba Cloud(cn-beijing)
    qwen3.7-plus
    $0.40$1.60$0.08
    Alibaba Cloud(singapore)
    qwen3.7-plus
    $0.40$1.60$0.08
    Alibaba Cloud(us-virginia)
    qwen3.7-plus
    $0.40$1.60$0.08
    Alibaba Cloud
    qwen3.7-plus
    $0.40$1.60$0.08
    Alibaba Cloud(eu-frankfurt)
    qwen3.7-plus
    $0.28$1.10$0.06
    NovitaAI
    qwen3.7-max
    $1.25$3.75$0.25
    Granite
    qwen3.7-max
    $2.50$1.25
    -50% off
    $7.50$3.75
    -50% off
    $0.50$0.25
    -50% off
    Alibaba Cloud(us-virginia)
    qwen3.7-max
    $2.50$7.50$0.50
    Alibaba Cloud(singapore)
    qwen3.7-max
    $2.50$7.50$0.50
    Alibaba Cloud(cn-beijing)
    qwen3.7-max
    $1.72$5.17$0.34
    Alibaba Cloud
    qwen3.7-max
    $2.50$7.50$0.50
    Alibaba Cloud(eu-frankfurt)
    qwen3.7-max
    $1.65$4.95$0.33
    Google AI Studio
    gemini-3.1-flash-lite
    $0.25$1.50$0.02
    Google Vertex AI
    gemini-3.1-flash-lite
    $0.25$1.50$0.02
    CanopyWave
    deepseek-v4-flash
    $0.14$0.28$0.03
    Together AI
    deepseek-v4-flash
    $0.14$0.28$0.03
    NovitaAI
    deepseek-v4-flash
    $0.14$0.28$0.03
    Alibaba Cloud(singapore)
    deepseek-v4-flash
    $0.20$0.40$0.04
    Runware
    deepseek-v4-flash
    $0.08$0.05
    -30% off
    $0.15$0.11
    -30% off
    $0.01$0.01
    -30% off
    ByteDance
    deepseek-v4-flash
    $0.14$0.28$0.03
    Gonka24
    deepseek-v4-flash
    $0.05$0.09$0.00
    Baidu
    deepseek-v4-flash
    $0.14$0.28$0.03
    DeepSeek
    deepseek-v4-flash
    $0.14$0.28$0.00
    Alibaba Cloud
    deepseek-v4-flash
    $0.20$0.40$0.04
    Page 1 of 2

    Newsletter

    Stay ahead of the curve

    Join developers who get weekly insights on LLM routing, new model launches, and cost optimization — straight to their inbox.

    • New models & providers as they drop
    • Tips to cut latency & costs
    • Early access to beta features

    No spam. Unsubscribe anytime.

    All systems operational
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

    • Features
    • AI Gateway
    • Observability
    • Models
    • Providers
    • Rankings
    • Add Provider
    • Partners
    • Lounge
    • Changelog
    • DevPass
    • Compare Models
    • Enterprise

    Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Azure AI Foundry
    • Vercel AI Gateway
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Glacier
    • Iceberg
    • Granite
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Quartz
    • Avalanche
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Z AI
    • Moonshot AI
    • Baidu
    • Permafrost
    • Perplexity
    • Nebius AI
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • Custom
    • NanoGPT
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Sakana AI
    • Tundra
    • Xiaomi
    • DeepInfra
    • Reve
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI

    © 2026 LLM Gateway. All rights reserved.

    Modern LLMs now rival dedicated translation engines for most language pairs — and beat them on context awareness, tone, terminology consistency, and formatting. The strongest multilingual models are Google's Gemini line, OpenAI's GPT-5.4, Anthropic's Claude, and Alibaba's Qwen, which is particularly strong on Chinese and other Asian languages.

    Long context windows also change how translation work gets done: instead of translating strings in isolation, you can put an entire document plus a glossary into one prompt and keep terminology consistent throughout. For bulk workloads, budget models like Gemini Flash-Lite and DeepSeek V4 Flash bring the cost per translated word down to fractions of a cent.

    Frequently asked questions

    What is the best LLM for translation?

    Gemini 3.1 Pro and GPT-5.4 deliver the most consistent quality across a broad set of language pairs. Qwen3.7 Max is a top pick for Chinese, Japanese, and Korean, and Claude Sonnet 5 excels when tone and nuance matter. For bulk work, Gemini 2.5 Flash-Lite and DeepSeek V4 Flash offer the best cost per word.

    Are LLMs better than Google Translate or DeepL?

    For most content, yes — LLMs follow style guides, preserve formatting and placeholders, keep terminology consistent across a document, and adapt register on request. Dedicated engines still win on raw speed and per-character price for very simple, high-volume strings.

    How do I translate long documents?

    Use a long-context model and send the whole document in one call: a million-token window fits roughly 750,000 words, and single-call translation keeps names and terminology consistent. If a document exceeds the window, chunk it and include a running glossary in each prompt.

    Which models handle low-resource languages best?

    Coverage drops for languages with little training data. Gemini Pro and GPT-5.4 generally hold up best, but always test with your actual language pair before committing volume — with one API key you can run the same text through several models in minutes and compare.

    is now on LLM Gateway — 30% off open-source modelsends in 7d 12:15:07
    LLM Gateway
    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started
    1.6k