Gemini 3.8 Flash now available on AI Gateway
Gemini 3.8 Flash is now available on AI Gateway with a 1M token context window, multi-modal input support, tool calling, and web search capabilities. The model is 50% off through December 31st and features improved performance on software engineering and reasoning tasks.
is now available on AI Gateway.Gemini 3.8 Flash from Google
The model is 50% off through December 31st. It has a 1M token context window, accepts text, image, PDF, and video input, returns text, and supports tool calling and web search. Maximum output is 65,536 tokens.
Gemini 3.8 Flash improves on prior Flash models at software engineering, agent work, and multi-step reasoning, at the same speed and cost as the previous release. Thinking is on by default.
To use Gemini 3.8 Flash, set to :modelgoogle/gemini-3.8-flash
To use it in a coding agent, see the , then run to connect agents like Claude Code, OpenCode, Cursor, Pi, and more and select inside the agent.coding agents guidevercel ai-gateway coding-agents setupgoogle/gemini-3.8-flash
Try Gemini 3.8 Flash in the .model playground
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on (BYOK) requests.Bring Your Own Key
You can view available on AI Gateway.all language models
Source: original entry ↗