Claude Opus 4.8 + a Wave of New Models
Claude Opus 4.8 lands with a 1M context window, Sonnet 4.6 gets 1M context too, and Qwen3.7 Max, Grok Build 0.1, Kimi K2.6, and GLM-5.1 join the gateway.
Read more about Claude Opus 4.8 + a Wave of New Models →API, routing, model access, and dashboard updates for LLM Gateway.
Claude Opus 4.8 lands with a 1M context window, Sonnet 4.6 gets 1M context too, and Qwen3.7 Max, Grok Build 0.1, Kimi K2.6, and GLM-5.1 join the gateway.
Read more about Claude Opus 4.8 + a Wave of New Models →Send PDFs and text-family documents to Gemini models via the OpenAI-compatible `file` content block.
Read more about Document Reading (PDFs & more) →ByteDance Seedance video models land in the gateway, Chat gets pinning and cross-org sharing, plus vertex-anthropic, grok-4.20, and a stack of fixes.
Read more about Seedance Video Models, Pinned Chats, Sharing Across Orgs & More →Turn text into vectors for semantic search, clustering, and RAG — through the same gateway you already use for chat.
Read more about OpenAI-Compatible Embeddings →Sessions are now Agents — monitor your AI coding agents, track costs per agent, and drill into individual sessions.
Read more about Sessions Rebranded to Agents →Route requests to regional providers, protect your apps with built-in content moderation, enforce API key rate limits, and explore new models.
Read more about Multi-Region Routing, Content Filters & More →Generate videos via the API, track conversations with sessions, and more — plus new models and providers.
Read more about Video Generation, Sessions & More →Access OpenAI's most capable models — GPT-5.4 for complex professional work and GPT-5.4 Pro for smarter, more precise responses — with 1.05M context windows and reasoning support.
Read more about GPT-5.4 and GPT-5.4 Pro Now Available →A dedicated Image Studio in the Playground for gallery-based generation with multi-model comparison, an OpenAI-compatible /v1/images/edits endpoint, and a wave of image generation improvements.
Read more about Image Studio, Image Edits API & More →When a provider fails, LLMGateway now automatically retries your request on another provider. Every attempt is logged with full routing visibility, so you always know what happened.
Read more about Automatic Retry & Fallback with Full Routing Transparency →Showing 10 of 98 updates
Load more