Vercel AI Gateway vs OpenRouter: which to pick in 2026
By Felix Wunderlich -
Updated August 15, 2026. We revisit this comparison as the market moves.
TL;DR: price will not decide this one, both gateways pass provider list prices through with no markup on tokens. Pick Vercel AI Gateway if your team already deploys on Vercel and you want gateway-level zero data retention as the default posture, pick OpenRouter if you want the widest independent catalog, 400+ models from 70+ providers, and an enterprise EU routing option. If you need European hosting without an enterprise contract, neither offers it, which is where Opper, our own gateway, enters as the honestly labeled third option below.
Vercel AI Gateway and OpenRouter answer the same question, how to reach many models through one key without paying a markup on tokens, and they answer it from opposite directions. Vercel folds a gateway into the deployment platform your frontend may already live on, OpenRouter runs the largest independent model marketplace. Because both pass provider list prices straight through, the real differences sit in catalogs, retention guarantees, plan gates and where your data is processed, so that is what this guide compares, with every claim linked to the vendor's own pages as verified on August 11, 2026. One disclosure before we start: we build Opper, an AI gateway in the same category, and it appears below as a clearly labeled third column rather than a hidden thumb on the scale.
Vercel AI Gateway vs OpenRouter at a glance
All figures are vendor-stated and current, sources linked throughout the article.
| Dimension | Vercel AI Gateway | OpenRouter | Opper |
|---|---|---|---|
| Models | "Hundreds of models", no published total; text, image and video plus embeddings | 400+ models from 70+ providers | 700+ models incl. image, voice and video |
| Inference pricing | Provider list price, no markup, no platform fee | Provider list price, no markup | Provider list price, no markup |
| Buying credits | Card processing fees passed through on top-ups, enterprise invoicing avoids them | 5.5% fee ($0.80 min) on card purchases, 5% crypto | Fee on credit purchases, numbers on /pricing |
| BYOK | No markup or fee, paid tier only | Free up to $25k/month list-price usage, then 5% | Free on the Gateway plan (details) |
| API surfaces | AI SDK, OpenAI Chat Completions and Responses, Anthropic Messages, OpenResponses | OpenAI-compatible API, official AI SDK provider | OpenAI, Anthropic and Gemini SDK compatible |
| Failover | Automatic provider retries, configurable model fallbacks | Price-prioritized load balancing, uptime filtering, fallbacks | Aliases with automatic fallback |
| Prompt retention | Gateway-level zero data retention by default, routing enforcement gated to Pro/Enterprise | Varies per provider, displayed but not routed on by default | Metadata only by default, no prompts stored |
| EU option | No published EU hosting option | Enterprise-only EU endpoint, on request | EU-hosted (AWS Stockholm), EU routes end to end |
| Free tier | $5 of credits every 30 days on a subset of models, ends on first purchase | Free model variants with daily caps | See /pricing |
| Account | Requires a Vercel team | Standalone | Standalone |
How is pricing different when neither adds a markup?
Both vendors say the same thing about tokens. Vercel's pricing page states that AI Gateway "charges no markup and no platform fee on tokens", you prepay credits and spend them at the provider's list price, and OpenRouter's FAQ says it passes through underlying provider pricing "without any markup". The difference is the on-ramp. OpenRouter charges 5.5% with a $0.80 minimum on card credit purchases, or 5% for crypto. Vercel charges no platform fee but notes you are "responsible for any payment processing fees that may apply" on top-ups, and only enterprise teams paying by invoice avoid those entirely.
Bring-your-own-keys is where the models diverge more. Vercel's BYOK terms carry no markup or fee at all, but BYOK requires the paid tier, and failed BYOK requests are retried on Vercel's own credentials and charged to your balance. OpenRouter prices BYOK with a monthly allowance: it is free up to $25,000 per month of list-price usage on pay-as-you-go ($200,000 on enterprise), with 5% of the equivalent OpenRouter cost deducted from credits above that.
Free tiers follow the same pattern of generosity with strings attached. Every Vercel team gets $5 of credits every 30 days, which the current pricing doc restricts to a subset of models with per-model rate limits below the paid tier, and the monthly free credit no longer applies once you purchase credits. OpenRouter's free usage runs through its free model variants instead, capped at 50 requests per day until you have bought $10 of credits, 1,000 per day after, at 20 requests per minute. Opper follows the same no-markup principle with a fee on credit purchases, and since fee schedules change faster than blog posts, the live numbers stay on the pricing page.
Who has more models, and who routes them better?
OpenRouter publishes concrete numbers on its homepage, 400+ models from 70+ providers, and that catalog plus its ecosystem remains the strongest reason to choose it. Vercel says "hundreds of models" across text, image and video plus embeddings, without publishing an exact total. For reference, Opper routes 700+ models including image, voice and video models, and you can put any two head to head on the compare pages.
Routing philosophies differ in ways that matter under load. Vercel's gateway automatically retries requests on other providers when one fails, and you can configure model fallbacks on top. OpenRouter goes further by default, load-balancing across providers with price as the priority, deprioritizing providers with recent outages, with fallbacks on by default and fine-grained controls on top, sort by price, throughput or latency, :nitro and :floor suffixes, or an explicit provider order. If routing control is your main criterion, OpenRouter's surface is the most configurable of the three.
Observability is the reverse trade. OpenRouter returns detailed usage information with every response, native-tokenizer token counts and cost in credits, with a /generation endpoint for async retrieval. Vercel includes usage, latency and spend monitoring in its dashboard, but deeper export is metered: Custom Reporting bills per write and per query, and Trace Drains, a Pro and Enterprise feature, bill on delivered traces and trace data volume.
What happens to your prompts?
Vercel has the strongest default retention story on paper. Its ZDR page states the gateway "does not retain prompts, outputs, or sensitive data" and that user data is "immediately and permanently deleted after requests are completed". Read the fine print though, because gateway retention and provider retention are different things. Enforcement that keeps your traffic on providers with zero-retention agreements is Pro and Enterprise only, free per request, $0.10 per 1,000 requests team-wide, and model-level exceptions exist: Vercel's own page notes that anthropic/claude-fable-5 supports ZDR on no provider, with prompts retained for 30 days and not used for training. BYOK keys are skipped under ZDR unless you mark them compliant yourself.
OpenRouter's posture is transparency rather than enforcement. Its privacy policy states it does not use your inputs or outputs for model training and that media files are not persisted beyond what routing requires, with carve-outs for abuse detection, security, billing and legal compliance. Downstream, provider data policies vary widely, and OpenRouter displays them but does not route on them by default, its docs say plainly that routing rules do not change based on provider retention policies, though you can disallow providers that train on prompts account-wide and filter by data policy per request. The same policy notes personal data may be transferred to servers in the US or other countries outside the European Economic Area.
On residency the positions are clear cut. OpenRouter now documents a concrete EU option: enterprise customers can request in-region routing via eu.openrouter.ai, where prompts and completions are processed within the European Union, enterprise status and an explicit request required. Vercel publishes no EU hosting option anywhere in its security and compliance section, its controls are provider-level rather than geographic. Keep one distinction in mind with any gateway: where the gateway processes a request and where the model actually runs are separate questions, so residency needs both the gateway leg and the inference leg pinned. That combination is Opper's home turf, the platform runs in AWS Stockholm with EU routes end to end and metadata-only retention by default, documented with a public sub-processor list on the compliance page, and the wider European field is covered in our European AI gateways guide.
When should you pick Vercel AI Gateway?
Pick Vercel when your team already deploys there. The gateway authenticates with API keys or Vercel OIDC tokens, credits and monitoring live in the dashboard you already use, and the default zero-retention posture is genuinely good. One myth to retire: it is not locked to Next.js or the AI SDK, the docs list OpenAI Chat Completions, OpenAI Responses and Anthropic Messages surfaces at ai-gateway.vercel.sh/v1, callable from any language. The real anchor is the account and the plan, you need a Vercel team, and ZDR enforcement, provider allowlists and trace export sit behind Pro and Enterprise, so the further you are from the Vercel platform, the less of the product you can reach.
When should you pick OpenRouter?
Pick OpenRouter for breadth and ecosystem. Its 400+ models from 70+ providers arrive with the most granular routing controls in this comparison, nearly every AI tool ships an OpenRouter integration, and it maintains an official provider for the Vercel AI SDK, so choosing OpenRouter does not mean giving up Vercel's tooling. Teams with enterprise contracts get the EU routing endpoint on request. One thing to watch during procurement: OpenRouter is reportedly in acquisition talks with Stripe, unannounced as of this writing. If you are weighing OpenRouter against Opper specifically, we keep a direct head-to-head on the OpenRouter alternative page.
Where does Opper fit as the third option?
Opper is our product, so weigh this section accordingly, but the gap it fills is visible in the table above: European hosting as the default rather than an enterprise negotiation. The gateway runs in AWS Stockholm with EU routes end to end, stores metadata only, no prompts, unless you opt into tracing on the Control Plane plan, and routes 700+ models including image, voice and video. Every call is metered with cost, latency and token accounting, uptime is published on a public status page, and aliases with automatic fallback handle the failover story. Switching is a base-URL change because the gateway speaks the OpenAI, Anthropic and Gemini SDK surfaces on one endpoint:
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.opper.ai/v3/compat",
api_key=os.environ["OPPER_API_KEY"],
)
For how the whole category shakes out beyond these three, our ten-gateway comparison covers the full field.
Vercel AI Gateway vs OpenRouter FAQ
Is Vercel AI Gateway cheaper than OpenRouter?+
Inference costs the same, both pass through provider list prices with no markup. The difference is the on-ramp: OpenRouter charges 5.5% with a $0.80 minimum on card credit purchases (5% for crypto), while Vercel adds no platform fee but notes you are responsible for any payment processing fees on top-ups, avoidable with enterprise invoicing. On BYOK, Vercel charges no fee but requires the paid tier, and OpenRouter is free up to $25,000 of monthly list-price usage on pay-as-you-go, then 5%.
Does Vercel AI Gateway only work with Next.js or the AI SDK?+
No. It exposes OpenAI Chat Completions, OpenAI Responses and Anthropic Messages surfaces at ai-gateway.vercel.sh/v1, callable from any language or framework. The tie is to the account rather than the framework: you need a Vercel team, and features like ZDR enforcement and provider allowlists are gated to Pro and Enterprise plans.
Does Vercel AI Gateway store my prompts?+
Vercel states a zero-data-retention policy at the gateway level, prompts and outputs are deleted immediately after the request completes. Gateway retention and provider retention are separate questions though: routing enforcement that keeps your traffic on providers with zero-retention agreements is Pro and Enterprise only, and model-level exceptions exist, Vercel's own ZDR page documents one model whose prompts are retained by the provider for 30 days on every host.
Does OpenRouter store my prompts?+
OpenRouter's privacy policy states it does not use your inputs or outputs for model training, and that media files are not persisted beyond what routing requires. Retention downstream depends on each provider's own policy, which OpenRouter displays but does not route on by default; you can disallow providers that train on prompts account-wide or filter by data policy per request. Its policy also notes personal data may be transferred to servers in the US.
Can I keep prompts and completions in the EU with either gateway?+
With OpenRouter yes, on an enterprise contract: its eu.openrouter.ai endpoint processes prompts and completions within the EU, enabled on explicit request. Vercel publishes no EU hosting option for AI Gateway. Opper is EU-hosted by default, running in AWS Stockholm with EU routes end to end, documented on the compliance page.
Can I use OpenRouter together with the Vercel AI SDK?+
Yes, OpenRouter maintains an official provider package for the AI SDK, @openrouter/ai-sdk-provider, so AI SDK versus OpenRouter is a false choice: you can build with the AI SDK and route through OpenRouter, or point the SDK at any OpenAI-compatible gateway.
Pick the gateway that matches where you run
If your app deploys on Vercel and Pro-tier controls cover your compliance needs, Vercel AI Gateway is a clean choice, and if catalog breadth with configurable routing matters most, OpenRouter has earned its ecosystem. If the deciding factor is European hosting with no markup on inference and observability built in, create a free Opper account and make your first call in minutes, browse the 700+ models it routes to, or check the numbers on pricing yourself.