Qwen3.7 Flash

Fast, cost-effective Qwen3.7 model with hybrid reasoning and tool calling for high-volume tasks.

qwen3.7-flash
STABLEGet StartedView uptime
1,000,000 context
Starting at $0.03/M input tokens (tiered)
Starting at $0.13/M output tokens (tiered)
Streaming
Vision
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

Select Provider

All Providers for Qwen3.7 Flash

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Alibaba Cloud
Context: 1M
Input
$0.03
/M tokens
Cache Read
$0.006
/M tokens
Output
$0.13
/M tokens
Cache Write 5m
$0.0375
/M tokens
Cache Write 1h
$0.0375
/M tokens
Tiered Pricing
IN
CACHED
OUT
≤32K tokens
$0.03
$0.006
$0.13
≤256K tokens
$0.1
$0.02
$0.4
≤1,000K tokens
$0.2
$0.04
$0.8
Get Started