Document Reading (PDFs & more)
Send PDFs and text-family documents to Gemini models via the OpenAI-compatible `file` content block.
Read more about Document Reading (PDFs & more) →API, routing, model access, and dashboard updates for LLM Gateway.
Send PDFs and text-family documents to Gemini models via the OpenAI-compatible `file` content block.
Read more about Document Reading (PDFs & more) →ByteDance Seedance video models land in the gateway, Chat gets pinning and cross-org sharing, plus vertex-anthropic, grok-4.20, and a stack of fixes.
Read more about Seedance Video Models, Pinned Chats, Sharing Across Orgs & More →Turn text into vectors for semantic search, clustering, and RAG — through the same gateway you already use for chat.
Read more about OpenAI-Compatible Embeddings →Sessions are now Agents — monitor your AI coding agents, track costs per agent, and drill into individual sessions.
Read more about Sessions Rebranded to Agents →Route requests to regional providers, protect your apps with built-in content moderation, enforce API key rate limits, and explore new models.
Read more about Multi-Region Routing, Content Filters & More →Generate videos via the API, track conversations with sessions, and more — plus new models and providers.
Read more about Video Generation, Sessions & More →Access OpenAI's most capable models — GPT-5.4 for complex professional work and GPT-5.4 Pro for smarter, more precise responses — with 1.05M context windows and reasoning support.
Read more about GPT-5.4 and GPT-5.4 Pro Now Available →A dedicated Image Studio in the Playground for gallery-based generation with multi-model comparison, an OpenAI-compatible /v1/images/edits endpoint, and a wave of image generation improvements.
Read more about Image Studio, Image Edits API & More →When a provider fails, LLMGateway now automatically retries your request on another provider. Every attempt is logged with full routing visibility, so you always know what happened.
Read more about Automatic Retry & Fallback with Full Routing Transparency →New unified reasoning object for precise control over reasoning models. Specify exact token budgets with max_tokens or use effort levels — all in one consistent API.
Read more about Unified Reasoning Configuration →Showing 10 of 97 updates
Load more