Gemini 3.8 Flash now available on AI Gateway
Gemini 3.8 Flash from Google is now available on AI Gateway. The model is 50% off through December 31st.
Google's Gemini 3.8 Flash is now available on Vercel's AI Gateway, priced 50% off through December 31st. It has a 1M token context window, accepts text, image, PDF, and video input, and returns text with support for tool calling and web search. Maximum output is 65,536 tokens, and thinking is on by default.
Vercel says the model improves on prior Flash models at software engineering, agent work, and multi-step reasoning, at the same speed and cost as the previous release. To use it, set the model to google/gemini-3.8-flash, or run vercel ai-gateway coding-agents setup to select it inside coding agents like Claude Code, OpenCode, Cursor, and Pi. It can also be tried in the model playground. As with other AI Gateway models, pricing reflects provider cost with no markup and no platform fee on inference, including on Bring Your Own Key requests.
