Skip to content

AI Gateway adds unified fast mode support

BetaVerifiedAdded Sep 22, 2026

AI Gateway has a unified fast mode abstraction, now in beta. You can now request fast mode the same way for every model on AI Gateway. Set speed to fast , and the gateway serves the fast tier when it's available and falls back to standard speed when it isn't. Fast mode trades a higher per-token cost for lower latency or higher throughput.

Read Vercel's release notes

https://vercel.com/changelog/ai-gateway-adds-unified-fast-mode-support

Summaries of vendors' own notes. Product names and logos belong to their owners; logos via logo.dev.

More AI SDK and AI Gateway releases

AI SDK and AI Gateway

AI Gateway now supports team and project spend budgets

Update
AI SDK and AI Gateway

DeepSeek V4 Flash now runs updated weights on AI Gateway

Preview
AI SDK and AI Gateway

AI Gateway logs now have a dedicated page

Update
AI SDK and AI Gateway

10x more capacity for Laguna S 2.1 on AI Gateway

Update
AI SDK and AI Gateway

Run multiple isolated agents in a single Sandbox

Update
AI SDK and AI Gateway

MiniMax H3 now available on AI Gateway

Update
AI SDK and AI Gateway

Inkling Small from Thinking Machines is now available on AI Gateway

Update

Also shipped on Jul 31, 2026

Web terminal on serverless GPU compute (AI Runtime) is in Public Preview

Preview
DatabricksLakeflow

OpenAI connector (Beta)

Beta
GitHubGitHub Copilot

Upcoming August 2026 model deprecations in GitHub Copilot

Deprecation

Weekly: the week's data and AI releases, Tuesday mornings.