Vercel agent orchestration platform for production deployment
Vercel announces an agent orchestration platform designed to handle production deployment of AI agents at scale. The platform provides infrastructure primitives including Sandboxes for secure execution, Fluid compute for cost-efficient resource scaling, AI Gateway for unified model access, Workflows for reliable multi-step operations, and Observability tools—eliminating the operational complexity that previously required months of engineering effort.
Prototyping is democratized, but production deployment isn't.
AI models have commoditized code and agent generation, making it possible for anyone to build sophisticated software in minutes. Claude can scaffold a fully functional agent before your morning coffee gets cold. But that same AI will happily architect a $5,000/month DevOps setup when the system could run efficiently at $500/month.
In a world where anyone can build internal tools and agents, the build vs. buy equation has fundamentally changed. Competitive advantage no longer comes from whether you can build. It comes from rapid iteration on AI that solves real problems for your business and, more importantly, reliably operating those systems at scale.
To do that, companies need an internal AI stack as robust as their external product infrastructure. That's exactly what Vercel's agent orchestration platform provides.
For decades, the economics of custom internal tools only made sense at large-scale companies. The upfront engineering investment was high, but the real cost was long-term operation with high SLAs and measurable ROI. For everyone else, buying off-the-shelf software was the practical option.
AI has fundamentally changed this equation. Companies of any size can now create agents quickly, and customization delivers immediate ROI for specialized workflows:
Today the question isn’t build vs. buy. The answer is build . Instead of separating internal systems and vendors, companies need a single platform that can handle the unique demands of agent workloads.and run
The number of use cases for internal apps and agents is exploding, but here's the problem: production is still hard.
Vibe coding has created one of the largest shadow IT problems in history, and understanding production operations requires expertise in security, observability, reliability, and cost optimization. These skills remain rare even as building becomes easier.
The ultimate challenge for agents isn't building them, it's the platform they run on.
Like OpenAI, we built our own internal data agent named d0 (OSS template ). At its core, d0 is a text-to-SQL engine, which is not a new concept. What made it a successful product was the platform underneath.here
Using Vercel’s built-in primitives and deployment infrastructure, one person built d0 in a few weeks using 20% of their time.
This was only possible because Sandboxes, Fluid compute and AI Gateway automatically handled the operational complexity that would have normally taken months of engineering effort to scaffold and secure.
Today, d0 has completely democratized data access that was previously limited to professional analysts. Engineers, marketers, and executives can all ask questions in natural language and get immediate, accurate answers from our data warehouse.
Here’s how it works:
Vercel provides the infrastructure primitives purpose-built for agent workloads, both internal and customer-facing. You build the agent, Vercel runs it. And it just works.
Using our own agent orchestration platform has enabled us to build and manage an increasing number of custom agents.
Internally, we run:
On the product side:
Both products run on the same primitives as our internal tools.
give agents a secure, isolated environment for executing sensitive autonomous actions. This is critical for protecting your core systems. When agents generate and run untested code or face prompt injection attacks, sandboxes contain the damage within isolated Linux VMs. When agents need filesystem access for information discovery, sandboxes can dynamically mount VMs with secure access to the right resources. Sandboxes
automatically handles the unpredictable, long-running compute patterns that agents create. It’s easy to ignore compute when agents are processing text, but when usage scales and you add data-heavy workloads for files, images, and video, cost becomes an issue quickly. Fluid compute automatically scales up and down based on demand, and you're only charged for compute time, keeping costs low and predictable.Fluid compute
gives you unified access to hundreds of models with built-in budget control, usage monitoring, and load balancing across providers. This is important for avoiding vendor lock-in and getting instant access to the latest models. When your agent needs to handle different types of queries, AI Gateway can route simple requests to fast, inexpensive models while sending complex analysis to more capable ones. If your primary provider hits rate limits or goes down, traffic automatically fails over to backup providers.AI Gateway
give agents the ability to perform complex, multi-step operations reliably. When agents are used for critical business processes, failures are costly. Durable orchestration provides retry logic and error handling at every step so that interruptions don't require manual intervention or restart the entire operation.Workflows
reveals what agents are actually doing beyond basic system metrics. This data is essential for debugging unexpected behavior and optimizing agent performance. When your agent makes unexpected decisions, consumes more tokens than expected, or underperforms, observability shows you the exact prompts, model responses, and decision paths, letting you trace issues back to specific model calls or data sources.Observability
In the future, every enterprise will build their version of d0. And their internal code review agent. And their customer support routing agent. And hundreds of other specialized tools.
The success of these agents depends on the platform that runs them. Companies who invest in their internal AI stack now will not only move faster, they'll experience far higher ROI as their advantages compound over time.
Build vs. buy ROI has fundamentally changed
The platform is the product: how our data agent runs on Vercel
Vercel is the platform for agents
Build your agents, Vercel will run them
OpenAI deployed an to democratize analyticsinternal data agent
Vercel’s helps one SDR do the work of 10 (template )lead qualification agenthere
Stripe built a (on a flight!)customer-facing financial impact calculator
"What was our Enterprise ARR last quarter?" d0 receives the message, determines the right level of data access based on the permissions of the user, and starts the agent workflow.A user asks a question in Slack:
The semantic layer is a file system of 5 layers of YAML-based configs that describe our data warehouse, our metrics, our products, and our operations. The agent explores a semantic layer:
Streaming responses, tool use, and structured outputs all work out of the box. We didn't build custom LLM plumbing, we used the same abstractions any Vercel developer can use.AI SDK handles the model calls:
If a step fails (Snowflake timeout, model hiccup), Vercel Workflows handles retries and state recovery automatically.Agent steps are orchestrated durably:
: File exploration, SQL generation, and query execution all happen in a secure Vercel Sandbox. Runaway operations can't escape, and the agent can execute arbitrary Python for advanced analysis.Automated actions are executed in isolation
: AI Gateway routes simple requests to fast models and complex analysis to Claude Opus, all in one code base. Multiple models are used to balance cost and accuracy
formatted results, often with a chart or Google Sheet link, are delivered back to the Slack using the AI SDK Chatbot primitive. The answer arrives in Slack:
A lead qualification agent
d0, our analytics agent
A customer support agent (handles 87% percent of initial questions)
An abuse detection agent that flags risky content
A content agent that turns Slack threads into draft blog posts.
v0 is a code generation agent, and
Vercel Agent can review pull requests, analyze incidents, and recommend actions.
Every company needs an internal AI stack
Source: original entry ↗
More from Vercel
Follow Vercel to get its new changes in your feed and email digest.
OpenAI Decisions API now available on AI Gateway
OpenAI's Decisions API is now accessible through Vercel's AI Gateway with an OpenAI-compatible endpoint, enabling decision models to answer typed questions and return probabilities, choices, and scores for routing, triage, and guardrails use cases. Support is available across the OpenAI SDK, AI SDK, HTTP API, and CLI with the latest versions.
Timestamp attributes now supported in Vercel Flags
Vercel Flags now supports timestamp attributes for entities, allowing you to create time-based targeting rules. Use this feature to run limited-time campaigns, show content between specific dates, or target users based on registration date.
Glyph Cluster now available in stealth on AI Gateway
Glyph Cluster, a reasoning model for coding and long-context analysis, is now available as a stealth model on Vercel's AI Gateway for Pro and Enterprise plan teams with purchased AI Gateway credits at no cost during the stealth period. The model supports function calling, streams responses, and can be accessed via AI SDK, OpenAI-compatible APIs, and coding agents.