megachangelog
Feature

DeepSeek V4.1 Flash now available on AI Gateway

DeepSeek V4.1 Flash is now available on Vercel's AI Gateway with native image understanding, supporting text and images in the same request, a 1 million token context window, and responses up to 384,000 tokens. The model includes reasoning, tool use, and prompt caching capabilities with optimized architecture for reduced computation.

is now available on AI Gateway with native image understanding. It accepts text and images in the same request, so you can ask questions about screenshots, read charts, and extract information from visual content.DeepSeek V4.1 Flash

The model has a 1 million token context window and supports responses up to 384,000 tokens, along with reasoning, tool use, and prompt caching. Its new architecture processes input and generates output with separate components, reducing the active computation needed for each stage.

Use as the model name:deepseek/deepseek-v4.1-flash

To use it in Claude Code, Codex, Cursor, and more, install the latest Vercel CLI and run setup:

Then select in the agent. See the for details.deepseek/deepseek-v4.1-flashcoding agents guide

Try DeepSeek V4.1 Flash in the .model playground

AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in , , , , and more.custom reportingZero Data Retention supportbudgets for API keysrouting rules

AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on (BYOK) requests.Bring Your Own Key

You can view available on AI Gateway.all language models

Read more

ai-gatewaydeepseekmodelsapivision

Source: original entry ↗