Qwen3-Next models now supported in Vercel AI Gateway
Vercel AI Gateway now supports Qwen3-Next models, allowing access to ultra-efficient 3B parameter models through a unified API with integrated observability, BYOK support, and intelligent provider routing without requiring separate provider accounts.
You can now access , two ultra-efficient models from , designed to activate only 3B parameters, using Vercel's with no other provider accounts required.Qwen3 NextQwenLMAI Gateway
AI Gateway lets you call the model with a consistent unified API and just a single string update, track usage and cost, and configure performance optimizations, retries, and failover for higher than provider-average uptime.
To use it with the , start by installing the package:AI SDK v5
Then set the model to :alibaba/qwen3-next-80b-a3b-thinking
Includes built-in , , and intelligent with automatic retries.observabilityBring Your Own Key supportprovider routing
Learn more about and access the model . AI Gatewayhere
Source: original entry ↗
More from Vercel
Follow Vercel to get its new changes in your feed and email digest.
OpenAI Decisions API now available on AI Gateway
OpenAI's Decisions API is now accessible through Vercel's AI Gateway with an OpenAI-compatible endpoint, enabling decision models to answer typed questions and return probabilities, choices, and scores for routing, triage, and guardrails use cases. Support is available across the OpenAI SDK, AI SDK, HTTP API, and CLI with the latest versions.
Timestamp attributes now supported in Vercel Flags
Vercel Flags now supports timestamp attributes for entities, allowing you to create time-based targeting rules. Use this feature to run limited-time campaigns, show content between specific dates, or target users based on registration date.
Glyph Cluster now available in stealth on AI Gateway
Glyph Cluster, a reasoning model for coding and long-context analysis, is now available as a stealth model on Vercel's AI Gateway for Pro and Enterprise plan teams with purchased AI Gateway credits at no cost during the stealth period. The model supports function calling, streams responses, and can be accessed via AI SDK, OpenAI-compatible APIs, and coding agents.