
Gemini 3.6 Flash and Gemini 3.5 Flash-Lite are now available on AI Gateway
Quick Answer
Gemini 3.6 Flash and 3.5 Flash-Lite are now available on AI Gateway, enhancing coding and web development efficiency while reducing token usage.
Quick Take
The unified API offers features like custom reporting and zero data retention, reflecting provider pricing without markup.
Key Points
- Gemini 3.6 Flash improves coding and web development quality with fewer model calls.
- Gemini 3.5 Flash-Lite enhances agentic capabilities for subagents in larger tasks.
- AI Gateway offers a unified API for model calls, usage tracking, and performance optimizations.
- No markup on provider pricing; no platform fee on inference, including BYOK requests.
- AI Gateway features a model leaderboard tracking popular models by token volume.
📖 Reader Mode
~1 min readGemini 3.6 Flash and Gemini 3.5 Flash-Lite are now available on AI Gateway.
Gemini 3.6 Flash improves quality across coding, agentic tasks, and web development while consuming fewer tokens and making fewer model calls. It produces cleaner web and app development output.
Gemini 3.5 Flash Lite upgrades the agentic capabilities of the Flash-Lite tier, making it a good fit for subagents that handle scoped parts of a larger task.
To use them, set model to google/gemini-3.6-flash or google/gemini-3.5-flash-lite in the AI SDK:
import { streamText } from 'ai';
const result = streamText({
model: 'google/gemini-3.6-flash',
prompt: 'Build a settings page with profile and notification sections.',
});
AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in custom reporting, Zero Data Retention support, budgets for API keys, routing rules, and more.
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests.
Try Gemini 3.6 Flash in the model playground.
AI Gateway: Track top AI models by usage
The AI Gateway model leaderboard tracks the most popular models over time, ranking them by the total volume of tokens processed across all Gateway traffic.
View the leaderboard
— Originally published at vercel.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from Vercel AI
See more →
The Agent Stack
The Agent Stack by Vercel AI provides essential building blocks for creating production-grade agents, enabling seamless integration across multiple AI models and secure operations. It features components like AI Gateway for model routing, Workflow SDK for durable execution, and Vercel Connect for scoped access, streamlining agent development and deployment across various platforms.

