AI Gateway adds unified fast mode support
BetaVerifiedAdded Sep 22, 2026
AI Gateway has a unified fast mode abstraction, now in beta. You can now request fast mode the same way for every model on AI Gateway. Set speed to fast , and the gateway serves the fast tier when it's available and falls back to standard speed when it isn't. Fast mode trades a higher per-token cost for lower latency or higher throughput.
Topics: AI agents, LLMs, Pricing, Developer tools, Coding agents
Summaries of vendors' own notes. Product names and logos belong to their owners; logos via logo.dev.
More AI SDK and AI Gateway releases
Every AI SDK and AI Gateway release| Date | Release | Type |
|---|---|---|
| Jul 31 | AI Gateway now supports team and project spend budgets Update | Update |
| Jul 31 | DeepSeek V4 Flash now runs updated weights on AI Gateway Preview | Preview |
| Jul 31 | AI Gateway logs now have a dedicated page Update | Update |
| Jul 31 | Expanded search for workflow runs in Vercel Observability Preview | Preview |
| Jul 31 | 10x more capacity for Laguna S 2.1 on AI Gateway Update | Update |
| Jul 30 | Run multiple isolated agents in a single Sandbox Update | Update |
| Jul 30 | MiniMax H3 now available on AI Gateway Update | Update |
| Jul 30 | Inkling Small from Thinking Machines is now available on AI Gateway Update | Update |