AI Signal 651 2 feeds carried it
Gemini 3.8 Flash now available on AI Gateway
Google's Gemini 3.8 Flash model is now routable through Vercel AI Gateway under the id google/gemini-3.8-flash, with a 50% promotional price running through December 31.
Engineers already on Vercel AI Gateway can swap in a model the vendor says improves on prior Flash releases for software engineering, agent work, and multi-step reasoning, at the same speed and cost as before, with thinking enabled by default. The 1M-token context, multimodal input, and existing streamText integration mean pipelines can adopt the new id with minimal reconfiguration. Because Vercel adds no markup on inference, the only price signal to track is Google's year-end discount window.
Written by elseif from the cluster below · every claim links back to a sourceThe three things worth knowing
AI Gateway exposes google/gemini-3.8-flash with provider pricing and no platform markup, including on BYOK requests.
The model ships with thinking on by default, a 1M-token context, text/image/PDF/video input, and a 65,536-token max output.
Pricing is 50% off the prior Flash rate through December 31, after which the discount expires.
THE CLUSTER
↗