Native VS Code Integration
The official LLM Gateway extension is now on the VS Code Marketplace. Use your PAYG or DevPass key to select gateway models directly in Copilot Chat and agent mode.
Read more about Native VS Code Integration →Coding plans, agent integrations, and account updates for DevPass.
The official LLM Gateway extension is now on the VS Code Marketplace. Use your PAYG or DevPass key to select gateway models directly in Copilot Chat and agent mode.
Read more about Native VS Code Integration →The models directory filters and sorts image, video, and speech models by their real unit prices and shows retirement status as chips; model pages sort providers by price, speed, or context; Seedream 5.0 Pro accepts up to 10 reference images; DevPass shows exact renewal and reset times; and Enterprise licenses warn 90 days before expiry.
Read more about Per-Unit Price Filters, Seedream References & More →Routing learns cache-hit rates and output-to-input proportions from recent project/model usage to compare providers for large prompts and sessions. Available automatically across plans, with workload defaults before enough history exists and explicit overrides on Enterprise.
Read more about Adaptive Cache-Aware Provider Selection →DevPass subscribers can now restrict routing to providers that explicitly state API inputs are not used for training. Unknown policies fail closed, and retries or fallbacks never escape the setting. Available on every active DevPass tier.
Read more about No AI Training for DevPass →DevPass subscribers can now remove their saved card after canceling, or whenever a subscription is inactive. The card details leave Stripe while a privacy-safe fingerprint remains to enforce the one-card-per-account rule.
Read more about Remove Your DevPass Payment Card →Claude now routes through Microsoft Foundry under a new azure-anthropic provider. The models directory gained lifecycle status filters, model pages sort providers by price, speed, or context, and API key lists flag the keys that are near or at their limit.
Read more about Claude on Foundry, Model Status & More →Provider cache writes is now a three-way setting. The new Client-managed mode forwards the cache markers your client sends and never adds any of its own, so one API key can serve a coding agent that manages its own caching alongside traffic that should not pay the cache-write premium.
Read more about Client-Managed Prompt Caching →DevPass no longer has to stop at 100%: opt into pay-as-you-go overflow and, once your monthly allowance is used, requests keep flowing from a credits balance billed at provider rates. Top up from the dashboard with your saved card, set auto-reload so the balance refills itself, and track it all on the new Usage page.
Read more about DevPass Pay-As-You-Go Overflow →The LLM Gateway CLI can now start any supported coding agent pre-wired to the gateway — Claude Code, OpenCode, Codex CLI, DevPass Code, and eight more. One command, one API key, 200+ models, and every request tracked in your dashboard.
Read more about Launch Any Coding Agent from the CLI →DevPass upgrades now roll your unspent allowance into the new tier — or schedule the switch for your next renewal. Plus two new providers including SCX.ai's Turbo inference (up to 4x faster), Gemini 3.6 Flash and 3.5 Flash Lite, Gemini TTS on two providers, and Empryo as a first-class coding agent.
Read more about Upgrade Rollover, New Providers & Gemini TTS →Showing 10 of 17 updates
Load more