megachangelog
Feature

Video Generation with AI Gateway

AI Gateway now supports video generation with beta access for Pro, Enterprise, and paid users. You can generate cinematic videos through text prompts, animate images, define frame transitions, extract character identity, or apply style transfers using four initial video models from xAI, Alibaba, Kling, and Google through a unified API.

AI Gateway now supports video generation, so you can create cinematic videos with photorealistic quality, synchronized audio, generate personalized content with consistent identity, all through AI SDK 6.

Video generation is in beta and currently available for Pro and Enterprise plans and paid AI Gateway users.

Video models require more than just describing what you want. Unlike image generation, video prompts can include motion cues (camera movement, object actions, timing) and optionally audio direction. Each provider exposes different capabilities through that unlock fundamentally different generation modes. See the for model-specific options.providerOptionsdocumentation

AI Gateway initially supports 4 types of video generation:

Across the model creators, their current capabilities across the models on AI Gateway are listed below:

Describe what you want, get a video. The model handles visuals, motion, and optionally audio. Great for hyperrealistic, production-quality footage with just a simple text prompt.

Generate videos on demand for your app, platform, or content pipeline. No licencing fees or production required, just prompts and outputs.Example: Programmatic video at scale.

This example uses to generate video from a text prompt with a specified aspect ratio and duration.klingai/kling-v2.6-t2v

Turn a simple prompt into polished video clips for social media, ads, or storytelling with natural motion and cinematic quality.Example: Creative content generation.

By setting a very specific and descriptive prompt, generates video with immense detail and the exact desired motion.google/veo-3.1-generate-001

Provide a starting image and animate it. Control the initial composition, then let the model generate motion.

Turn existing product photos into interactive videos.Example: Animate product images.

The model animates a product image after you pass an image URL and motion description in the prompt.klingai/kling-v2.6-i2v

Bring static artwork to life with subtle motion. Perfect for thematic content or marketing at scale.Example: Animated illustrations.

Add subtle motion to food, beverage, or lifestyle shots for social content.Example: Lifestyle and product photography.

Here, a picture of coffee is rendered for a more interactive video, with lighting direction and minute details.

Define the start and end states, and the model generates a seamless transition between them.

Outfit swaps, product comparisons, changes over time. Upload two images, get a seamless transition.Example: Before/after reveals.

The start and end states are defined here with two images that used in the prompt and provider options.

In this example, lets you define the start frame in and the end frame in . The model generates the transition between them.klingai/kling-v3.0-i2vimagelastFrameImage

Provide reference videos or images of a person/character, and the model extracts their appearance and voice to generate new scenes starring them with consistent identity.

In this example, 2 reference images of dogs are used to generate the final video.

Using here, you can instruct the model to utilize the people/characters within the prompt. Wan suggests using , , etc. in the prompt for multi-reference to video to get the best results.alibaba/wan-v2.6-r2v-flashcharacter1character2

Transform existing videos with style transfer. Provide a video URL and describe the transformation you want. The model applies the new style while preserving the original motion.

Here, utilizes a source video from a previous generation to edit into a watercolor style.xai/grok-imagine-video

For more examples and detailed configuration options for video models, check out the . You can also find simple getting started scripts with the .Video Generation DocumentationVideo Generation Quick Start

Check out the changelogs for these video models for more detailed examples and prompts.

Read more

Two ways to get started

Four initial video models; 17 variations

Understanding video requests

Generation types

Text-to-video

Image-to-video

First and last frame

Reference-to-video

Video Editing

Get started

  • : Generate videos programmatically with the same interface you use for text and images. One API, one authentication flow, one observability dashboard across your entire AI pipeline.AI SDK 6

  • : Experiment with video models with no code in the configurable that's embedded in each model page. Compare providers, tweak prompts, and download results without writing code. To access, click any video gen model in the .AI Gateway PlaygroundAI Gateway playgroundmodel list

  • from xAI is fast and great at instruction following. Create and edit videos with style transfer, all in seconds.Grok Imagine

  • from Alibaba specializes in reference-based generation and multi-shot storytelling, with the ability to preserve identity across scenes.Wan

  • excels at image to video and native audio. The new 3.0 models support multishot video with automatic scene transitions.Kling

  • from Google delivers high visual fidelity and physics realism. Native audio generation with cinematic lighting and physics.Veo

Type

Inputs

Description

Example use cases

Text-to-video

Text prompt

Describe a scene, get a video

Ad creative, explainer videos, social content

Image-to-video

Image, text prompt optional

Animate a still image with motion

Product showcases, logo reveals, photo animation

First and last frame

2 images, text prompt optional

Define start and end states, model fills in between

Before/after reveals, time-lapse, transitions

Reference-to-video

Images or videos

Extract a character from reference images or videos and place them in new scenes

Spokesperson content, consistent brand characters

Model Creator

Capabilities

xAI

Text-to-video, image-to-video, video editing, audio

Wan

Text-to-video, image-to-video, reference-to-video, audio

Kling

Text-to-video, image-to-video, first and last frame, audio

Veo

Text-to-video, image-to-video, audio

video-generationai-gatewayai-sdkfeature

Source: original entry ↗

More from Vercel

Follow Vercel to get its new changes in your feed and email digest.

Feature

OpenAI Decisions API now available on AI Gateway

OpenAI's Decisions API is now accessible through Vercel's AI Gateway with an OpenAI-compatible endpoint, enabling decision models to answer typed questions and return probabilities, choices, and scores for routing, triage, and guardrails use cases. Support is available across the OpenAI SDK, AI SDK, HTTP API, and CLI with the latest versions.

ai-gatewayopenaiapidecisionssdks
Feature

Timestamp attributes now supported in Vercel Flags

Vercel Flags now supports timestamp attributes for entities, allowing you to create time-based targeting rules. Use this feature to run limited-time campaigns, show content between specific dates, or target users based on registration date.

flagstargetingfeaturetimestampscampaigns
Feature

Glyph Cluster now available in stealth on AI Gateway

Glyph Cluster, a reasoning model for coding and long-context analysis, is now available as a stealth model on Vercel's AI Gateway for Pro and Enterprise plan teams with purchased AI Gateway credits at no cost during the stealth period. The model supports function calling, streams responses, and can be accessed via AI SDK, OpenAI-compatible APIs, and coding agents.

ai-gatewaymodelscodingstealth
See all Vercel changes →