DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash is a multimodal mixture-of-experts model for coding, reasoning, visual understanding, and tool-using agents. It processes text and images, generates text, and combines a 1M-token context window with continuously adjustable reasoning effort. Its Causal Encoder-Decoder architecture, sparse attention, and compressed KV cache reduce the memory and computation required for long inputs, making it well suited to large-codebase work, document analysis, research, and sustained agentic tasks.

runware/deepseek-v4-1-flash
STABLEGet Started
Streaming
Tools
Reasoning
JSON Output
Structured JSON
No ratings yetSign in to rate

Select Provider

Runware Pricing for DeepSeek-V4.1-Flash

View detailed pricing and capabilities for this provider.

Context: 1.0M
Input
$0.15
/M tokens
Cache Read
$0.01
/M tokens
Output
$0.6
/M tokens
Get Started