Vercel Sandbox now supports Devin Outposts
Vercel Sandbox now supports Devin Outposts, allowing each Devin session to execute in an isolated Sandbox microVM with session orchestration and command execution running in the user's Vercel project.…
Vercel Sandbox now supports Devin Outposts, allowing each Devin session to execute in an isolated Sandbox microVM with session orchestration and command execution running in the user's Vercel project.…
Vercel's AI Gateway is offering DeepSeek V4 Flash at a 90% discount through Novita for Vercel Pro customers until August 11, with the model accessible via the identifiers deepseek/deepseek-v4-flash or…
Vercel WAF for Blob is now generally available and supported for production use on all plans, with existing beta rules carrying over unchanged. The firewall protects Blob stores with custom rules that…
Factory, an AI-powered software development platform, runs its entire cloud backend on Vercel's Next.js, handling tens of millions of daily API requests with a p95 response time of 350ms or below, wit…
Factory, an AI-powered software development platform, scaled its cloud backend on Vercel to serve one billion daily API requests with a p95 response time of 350ms or below, without a dedicated infrast…
Vercel's AI Gateway now supports Qwen 3.8 Max, a 2.4-trillion-parameter model handling text and vision tasks with a 1-million-token context window. The model, accessed via model ID alibaba/qwen3.8-max…
Vercel's AI Gateway now supports spend budgets scoped to teams, projects, or API keys, allowing dollar limits that halt requests when exceeded. Budgets can be set via the dashboard or CLI, with option…
Vercel's AI Gateway now serves DeepSeek V4 Flash with updated weights by default, boosting its Terminal-Bench score to 82.7, up 25.8 points from 56.9 in the April preview. Requests to `deepseek/deepse…
Vercel has launched a dedicated Logs page for its AI Gateway, listing every request sent through the gateway with cost, token counts, duration, and routing details. The page, available at team and pro…
Poolside's Laguna S 2.1 model now has 10x more capacity on Vercel's AI Gateway, applying to both the paid version `poolside/laguna-s-2.1` and the free version `poolside/laguna-s-2.1-free`, enabling hi…
Vercel MCP now supports the 2026-07-28 MCP specification, enabling newer clients to use a stateless request model and updated authorization behavior without client-side changes. Both protocol versions…
Shopify and Vercel are rebuilding Hydrogen, Shopify's framework for headless storefronts, to be open source and runtime agnostic, enabling developers to use any JavaScript framework. The new version c…
Vercel's @vercel/sandbox SDK now supports multiple Linux users and groups, enabling developers to run isolated agents side by side in a single Sandbox with private home directories and optional shared…
Vercel's AI Gateway now supports MiniMax H3, a model that generates 2K video from text prompts, images, keyframes, or reference material including images, video, and audio. The model outputs mp4 at 2K…
Vercel released mcp-handler@2.0.0 with support for the 2026-07-28 Model Context Protocol specification and MCP TypeScript SDK v2, enabling stateless protocol support without Redis dependency. The upda…
AI Gateway has reduced prices for GPT-5.6 Luna and GPT-5.6 Terra, with Luna seeing an 80% price cut to $0.2 per 1M input tokens and $1.2 per 1M output tokens, and Terra dropping 20% to $2 and $12 resp…
Inkling Small from Thinking Machines is now available on Vercel's AI Gateway, offering performance comparable to the larger Inkling model at about a quarter of the size with lower compute per task. Th…
Vercel's AI Gateway has launched a unified fast mode abstraction in beta, allowing users to request lower-latency or higher-throughput serving for any supported model by setting speed to 'fast' under …
XAI's Grok Voice Think Fast 2.0, a speech-to-speech voice model with improved reasoning, transcription accuracy, and conversation, is now available on Vercel's AI Gateway. The model reasons in paralle…
Vercel's AI Gateway now supports regional inference, allowing users to pin requests to the US or EU via a single `inferenceRegion` field, with responses reporting the serving region. The feature repla…