Native VS Code Integration
The official LLM Gateway extension is now on the VS Code Marketplace. Use your PAYG or DevPass key to select gateway models directly in Copilot Chat and agent mode.
Read more about Native VS Code Integration →The latest features, improvements, and fixes across LLM Gateway, DevPass, Lounge, and Airside.
The official LLM Gateway extension is now on the VS Code Marketplace. Use your PAYG or DevPass key to select gateway models directly in Copilot Chat and agent mode.
Read more about Native VS Code Integration →Generate and edit images with OpenAI's GPT Image 2.5 Sunburst and Flare. Both add xhigh and max quality settings through the images API, chat completions, and Playground.
Read more about GPT Image 2.5 Is Here →The /v1/realtime WebSocket now opens transcription-only sessions: live speech-to-text with no speech model in the loop, billed per minute or per token against the transcription model alone. Lounge gets a Transcribe mode on the same page.
Read more about Realtime Transcription Sessions →The models directory filters and sorts image, video, and speech models by their real unit prices and shows retirement status as chips; model pages sort providers by price, speed, or context; Seedream 5.0 Pro accepts up to 10 reference images; DevPass shows exact renewal and reset times; and Enterprise licenses warn 90 days before expiry.
Read more about Per-Unit Price Filters, Seedream References & More →The LLM Gateway MCP server gains get-account, get-usage, and get-usage-breakdown, so Claude Code, Codex, Cursor, or any MCP client can check spending limits, request and token totals, costs, trends, and your most-used providers, models, coding apps, and API keys without opening the dashboard.
Read more about Usage Analytics in the MCP Server →Gemini 3.8 Flash lands on Google AI Studio and Vertex AI with a 1M context at $0.75/M input, Meta's Muse Spark 1.3 arrives with a 1M context and a $0.10/M Contributor tier, and Kimi K3, GLM-5.3 Flash, and Qwen3.8 Flash pick up new deployments across Runware, Novita, SCX.ai, and Alibaba Cloud.
Read more about Gemini 3.8 Flash, Muse Spark 1.3 & More Models →The dashboard usage chart now overlays a comparison period, whether the previous period, a chosen week or month, or an exact custom range, and switches between a total view and a token-cost breakdown. Bars are grouped and stacked with an explicit legend, and rankings gain a hover-isolating legend and an all-time top apps panel.
Read more about Compare Usage Across Periods →Sign in to the LLM Gateway CLI from your browser or enterprise SSO instead of typing a password, and pull your organization's shared skills into Claude Code, Codex, OpenCode, Cursor, and other agents with one command. Browser login works on every plan; organization skills are available on the Enterprise plan.
Read more about CLI Browser Login and Organization Skills →Routing learns cache-hit rates and output-to-input proportions from recent project/model usage to compare providers for large prompts and sessions. Available automatically across plans, with workload defaults before enough history exists and explicit overrides on Enterprise.
Read more about Adaptive Cache-Aware Provider Selection →Airside now runs a preflight against your endpoint before a model can be filed: every capability you declare is probed live, and the results become the capability badges developers see. Carriers can also pick the upstream API format per model, invite crew by email, file per-model fares, and rename their carrier under review.
Read more about Airside: Model Verification and Crew Invites →Showing 10 of 110 updates
Load more