Consensus Protocol API
Consensus Protocol serves open-weight large language models on dedicated GPU hardware it operates, via an OpenAI-compatible inference API.
Data & Privacy
HQ:US
Safety identifier:Not forwarded
API Training:No
Prompt Logging:No
Retention:0 days
Available Models
DeepSeek V4.1 Flash
deepseek
deepseek-v4.1-flashStreaming
Vision
Tools
Reasoning
Consensus Protocol
Context: 1.0M
Input
$0.15
/M tokens
Cached
$0.005
/M tokens
Output
$0.55
/M tokens
Qwen3.8 27B
Qwen3.8 27B
Qwen3.8-27BStreaming
Vision
Tools
Reasoning
Consensus Protocol
Context: 262k
Input
$0.08
/M tokens
Cached
$0.05
/M tokens
Output
$0.35
/M tokens
GLM-5.3 Flash
zai
glm-5.3-flashStreaming
Vision
Tools
Reasoning
Consensus Protocol
Context: 1.0M
Input
$0.1
/M tokens
Cached
$0.02
/M tokens
Output
$0.25
/M tokens
DeepSeek V4 Flash
deepseek
deepseek-v4-flashStreaming
Tools
Reasoning
JSON Output
Consensus Protocol
Context: 524.3k
Input
$0.1
/M tokens
Cached
$0.005
/M tokens
Output
$0.2
/M tokens
Gemma 4 31B IT
google
gemma-4-31b-itStreaming
Tools
Reasoning
Consensus Protocol
Context: 262.1k
Input
$0.1
/M tokens
Cached
$0.01
/M tokens
Output
$0.25
/M tokens
GPT OSS 20B
openai
gpt-oss-20bStreaming
Tools
Reasoning
JSON Output
Consensus Protocol
Context: 128kQuant: fp4
Input
$0.04
/M tokens
Cached
$0.01
/M tokens
Output
$0.19
/M tokens
Create an LLM Gateway API key and point any OpenAI-compatible SDK at https://api.llmgateway.io/v1. Set the model to deepseek-v4.1-flash. On pay-as-you-go keys, consensusprotocol/deepseek-v4.1-flash pins every request to Consensus Protocol; DevPass coding plans do not support provider pinning. You do not need a separate Consensus Protocol account or SDK.