Your agent can now use the Vercel CLI to buy a domain. The agent discovers candidates with vercel domains search, checks availability and registrar pricing with vercel domains check and vercel domains price, and initiates the purchase with vercel domains buy. The purchase decision itself stays with you by design. Run non-interactively, vercel domains buy returns a structured, machine-readable error and the suggested next commands, so the agent hands the confirmation back to you, prohibiting it from spending money without your approval. You can confirm the price, the auto-renew choice, and the…
New Pro teams now default to a 30-day deployment retention period. Coding agents have accelerated development, increasing deployment frequency and the amount of stored deployment history. Shorter retention keeps recent history available while limiting older history and its storage costs. You can choose a longer retention period in your team's settings when you need more history for inspection or rollback. Deployment Storage and Functions Storage each cost $0.10 per GB-month. Older deployments may be deleted under your retention policy. Deleted deployments cannot be restored or used for…
Deployment Storage and Functions Storage billing will begin for all Pro teams at $0.10 per GB-month. Check your email for details about when billing will start for your team. To keep storage costs low, the deployment retention window also changes to 30 days. Deployments older than 30 days will be deleted starting October 23, unless you opt out in your retention settings before then. Retention settings already at 30 days or shorter are unchanged. Control your Deployment Storage Review retention periods for Pre-Production, Production, Canceled, and Errored deployments with a Deployment…
Rillet on Vercel Tripled shipping rate in three months with eve agents deployed on Vercel Two-person team closed 800+ Linear tickets in its first 4 months Builder access via Enterprise Managed Users, SSO, and Directory Sync Ships customer-requested changes to production in as little as two hours Gives every builder access to models through AI Gateway Rillet is an AI-native enterprise resource planning (ERP) platform. Their AI agents do accounting work inside a real-time general ledger, with human approval and a full audit trail. Rillet serves more than 600 customers working toward a zero-day…
Decision-1 from Microsoft is now available on AI Gateway. It answers typed questions about text with probabilities, choices, and scores. It can be used to classify support requests, select the next step in a workflow, or assess an AI response against a rubric. The model can answer yes/no questions, choose among named categories, or score an input against an ordered set of criteria. Multiple questions can share the same input in one request. Choice and score answers include probabilities for each option, so your application can decide when to act or send an uncertain result for review. Call…
Liquid d1 is now available on AI Gateway. d1 is a decision model that evaluates shared state against typed questions for classification, routing, and scoring. It returns structured answers with probabilities without generating text tokens. d1 also has vision support: the model can answer typed questions about images for visual classification, inspection, and scoring. Use liquid/d1 through the AI SDK decision API, the OpenAI-compatible Decisions API, or the TypeSafe-compatible API. These examples ask whether a support agent issued a refund: Image input This example reads a local PNG named…
You can now skip sending client request bodies to Routing Middleware with skipMiddlewareRequestBody. This reduces incoming Fast Origin Transfer usage for middleware invocations and can improve time to first byte, especially for requests with large bodies. Enable this option only if your Routing Middleware doesn’t read request bodies. Vercel Functions and rewrite targets still receive the body, so you can read and process it there. Set skipMiddlewareRequestBody to true in your project’s JSON configuration file, vercel.json: If you use a TypeScript configuration file, add the same setting to…
A new Vercel Sandbox is now ready to run commands and read or write files when creation completes. Sandbox.create() now resolves only after snapshot restoration completes, so the first command or file operation no longer waits for it. Startup errors also reject Sandbox.create() instead of surfacing during that operation. Snapshot restoration time now counts toward creation instead of the first operation. Creation is about 50 ms longer at p50, with a larger increase possible for uncached snapshots. The total time through the first operation remains the same. This change applies automatically,…
Grok Imagine Video 1.5 Lite from xAI is now available on AI Gateway. This lightweight model turns text prompts and images into video with native audio. The model produces clips from 1 to 15 seconds long in 480p, 720p, and 1080p. Seven aspect ratios cover widescreen, square, and portrait formats, so you can create horizontal scenes or vertical videos from the same generation workflow. To start generating videos, set the model to spacexai/grok-imagine-video-1.5-lite in the AI SDK. Chain an image model with Grok Imagine Video 1.5 Lite to create a still and animate it in one flow: You can also…
Step 5 Preview, StepFun's flagship model, is now available on AI Gateway. It's built for agentic coding, professional knowledge work, and financial analysis, from building apps and debugging code to research and analytical reports. Step 5 Preview supports text and image input with a 1M-token context window, allowing you to work with large codebases, a stack of documents, or screenshots and charts in a single request. Use stepfun/step-5-preview as the model name: For other coding agents such as Claude Code, Codex, Cursor, and more, install the latest Vercel CLI and run setup: Then select…
OpenAI's Decisions API is now available on AI Gateway through an OpenAI-compatible /v1/decisions endpoint. Use it with GPT-6 Luna Decisions or any other decision model on AI Gateway. Decision models answer typed questions about a shared input and return probabilities, choices, and scores instead of generated text, which suits routing, triage, guardrails, and rubric scoring. You can ask predicate (yes/no), choice, and score questions against the same input in one request. openai/gpt-6-luna-decisions is a separate model ID from openai/gpt-6-luna, which stays a language model for text…
Claude Haiku 5.5 from Anthropic is now available on AI Gateway. It is built for high-volume, cost-sensitive tasks such as summaries, context compaction, database queries, and classification, and works as a subagent alongside larger Claude models on coding work. It is also the fastest Claude model at standard speed, which suits live customer support and browser use. Test it as a direct upgrade wherever you use Claude Haiku 4.5 today. Haiku 5.5 is the first Haiku model with effort levels. It uses adaptive thinking and supports low, medium, high, xhigh, and max effort, which set how much the…
Vercel Flags now supports timestamp attributes for entities, which represent the users, teams, devices, or requests evaluated by a flag. Use timestamp attributes to show a campaign page only between two dates, run a Black Friday promotion, or target users who registered before a cutoff. Configure a timestamp attribute You can configure timestamp attributes in the dashboard or via the CLI. In the Vercel dashboard, open Flags, select Entities, and create or select an entity. Add an attribute and set its data type to timestamp. Or define the attribute with the CLI: Whether you configure the…
Microfrontends Routing is now free for requests that the Web Application Firewall (WAF) denies, challenges, or rate limits. CDN Requests and Fast Data Transfer are already free for this traffic. The change applies automatically to every project using microfrontends. Learn how to set up WAF rules in the Firewall documentation. Read more
AI Gateway can now escalate a decision request to a fallback model when the primary model's answer trips a confidence condition you set. Conditions can be combined, so an escalation can depend on more than one signal. The plain model names you already list in models keep catching outright errors, so existing fallbacks are unaffected. Confidence conditions cover Choice and Score questions. Boolean questions escalate on a probability range instead. Set the fallback by adding one conditional object to providerOptions.gateway.models. Requests without it keep their existing behavior. Leave out…
Nano Banana 2.1 from Google is now available on AI Gateway. It improves image generation and editing over earlier Nano Banana models. This release targets product recontextualization, mask- and ink-based editing, and factuality. Product recontextualization places an existing product in a new scene. Mask- and ink-based edits change only the region marked by a mask or by strokes drawn on the image. Factuality is how accurately images depict real-world subjects. The model renders photorealistic skin tones, intricate materials, sharp lighting, and coherent backgrounds at the latency and cost of a…
Mistral Large 4 is now available on AI Gateway. The open-weight, natively multimodal model combines reasoning with text and image understanding. It is useful for cyber defense, manufacturing, and finance workflows. You can access the model through AI Gateway with one API key, without creating a separate Mistral account. Use Mistral Large 4 (mistral/mistral-large-4 ) with the AI SDK, OpenAI-compatible Chat Completions API, Responses API, or Anthropic Messages API. You can also select it in a coding agent connected to AI Gateway. For coding agents, run npx vercel ai-gateway setup and select…
FLUX 3 Image from Black Forest Labs is now available on AI Gateway. Generate images from text or edit existing images using the same model, with one API key and no Black Forest Labs account required. FLUX 3 Image supports up to ten reference images for editing, combining subjects, or changing the composition. It supports five resolution tiers from 768 × 768 to 4K, with square, portrait, and landscape aspect ratios. Set AI_GATEWAY_API_KEY, then use bfl/flux-3-image as the model ID in AI SDK 7 or later. To generate an image, call generateImage with a text prompt: To edit an image, pass it in…
AI Gateway now supports OpenAI's Ultrafast service tier for GPT-6 Astra, providing faster output for interactive applications and rapid coding iterations. To use Ultrafast, request it for openai/gpt-6-astra through AI SDK, the Chat Completions API, or Responses API: For workflows with frequent tool calls, OpenAI recommends the Responses API over a persistent WebSocket connection to reduce overhead between turns. See the Ultrafast service-tier examples for persistent connections, including AI SDK over WebSocket. Ultrafast supports US and global processing. Requests pinned to unsupported…
It's simply impossible to not hear about Jev. Seemingly everyone is tinkering with it in some way, from using it to make trading decisions (what could possibly go wrong?) to generating UIs with it (maybe we're onto something here!). Jev is a new kind of AI model. You feed it data and ask it a set of multiple-choice questions, and it responds with its answers and how confident it is in each. Just like with any other model, Jev can and will make mistakes, but it will make them fast. The bottom line is: Jev is built for making narrow decisions. Whether it's accurate enough for your use cases…
Read at the source
Your visit, your choice.
Optional Google Analytics helps us understand visits. Microsoft Clarity records masked interactions to improve the site. Optional tools stay off unless you choose them. Privacy details.