AI Model Directory

Every model on the Opper gateway

Search by what matters. Zero data retention, residency, training and more.

by Anthropic

Claude Fable 5 by Anthropic, state-of-the-art generalist model with advanced vision and 1M token context for complex software engineering and reasoning.

Available on

Pricing from$10.00/ 1M input$50.00/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US, Multi
by OpenAI

OpenAI's frontier GPT-5.6 model for complex professional work. 1M context window with reasoning.

Available on

Pricing from$5.00/ 1M input$30.00/ 1M output1.1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by OpenAI

GPT-5.6 model that balances intelligence and cost. 1M context window with reasoning.

Available on

Pricing from$2.00/ 1M input$12.00/ 1M output1.1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by OpenAI

GPT-5.5 is OpenAI's newest frontier model, combining strong coding, reasoning, and web search and understanding multi-part tasks, with a 1M context window and vision.

Available on

Pricing from$5.00/ 1M input$30.00/ 1M output1.1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by xAI

xAI's Grok 4.5 model with toggleable reasoning, 500K context window, vision, and agentic tool use

Available on

Pricing from$2.00/ 1M input$6.00/ 1M output500Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by OpenAI

GPT-5.4 is OpenAI's frontier model for complex professional work, with a 1M context window, vision, tool use, structured output, and configurable reasoning.

Available on

Pricing from$2.50/ 1M input$15.00/ 1M output1.1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by OpenAI

GPT-5.6 model optimized for cost-sensitive, high-volume workloads. 1M context window with reasoning.

Available on

Pricing from$0.20/ 1M input$1.20/ 1M output1.1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Google

Gemini 3.5 Flash by Google pairs frontier-level intelligence with fast throughput and a 1M token context for agentic and coding tasks.

Available on

Pricing from$1.50/ 1M input$9.00/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Google

Google Gemini 3.6 Flash — improved token efficiency and agentic planning at a lower price point than 3.5 Flash

Available on

Pricing from$1.50/ 1M input$7.50/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Google

Gemini 3.1 Pro Preview by Google, a frontier reasoning LLM with 1M context and strong agentic coding, scoring 77% on ARC-AGI-2 and 81% on SWE-Bench Verified.

Available on

Pricing from$2.00/ 1M input$12.00/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Alibaba

Qwen3.7-Max by Alibaba, a 1M-context proprietary reasoning agent with extended thinking for multi-step code and autonomous workflows.

Available on

Pricing from$1.25/ 1M input$3.75/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
Multi, US
by MiniMax

MiniMax M3: natively multimodal 428B Mixture-of-Experts model with ~23B active parameters, MiniMax Sparse Attention for long context, and 59.0% on SWE-Bench Pro

Available on

Pricing from$0.30/ 1M input$1.20/ 1M output1Mcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by OpenAI

GPT-5.3 Codex is OpenAI's most capable agentic coding model, reaching state-of-the-art on SWE-Bench Pro and Terminal-Bench while using fewer tokens than prior models.

Available on

Pricing from$1.75/ 1M input$14.00/ 1M output272Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by OpenAI

GPT-5.2 by OpenAI, flagship coding and agentic model with a 272K context, vision, reasoning, and tool use.

Available on

Pricing from$1.75/ 1M input$14.00/ 1M output272Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Moonshot

Kimi K2.7 Code by Moonshot: open-weight, code-specialized agentic model on a 1 trillion parameter MoE, tuned for long-horizon software engineering and tool use.

Available on

Pricing from$0.95/ 1M input$4.00/ 1M output262Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by Z.ai

GLM-5.1 by Z.ai is a 754B open-weight (MIT) coding and agentic reasoning model with 202K context that tops SWE-Bench Pro at 58.4.

Available on

Pricing from$0.89/ 1M input$3.50/ 1M output205Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by OpenAI

GPT-5.4 Mini is OpenAI's strongest small model, improving over GPT-5 Mini in coding and reasoning while running more than 2x faster, with a 400k context window.

Available on

Pricing from$0.75/ 1M input$4.50/ 1M output400Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Z.ai

GLM 5 by Zhipu: 744B frontier MoE model with 40B active, reasoning-first design, and Huawei Ascend training.

Available on

Pricing from$0.60/ 1M input$1.60/ 1M output205Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by Alibaba

Qwen3.7-Plus by Alibaba, a 1M-context multimodal agent with vision, tool calling, and agentic iteration for GUI and CLI automation.

Available on

Pricing from$0.30/ 1M input$1.18/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
Multi
by OpenAI

GPT-5.4 Nano is the fastest, lightest GPT-5.4 variant, built for classification, data extraction, and lightweight coding subagents that need low latency.

Available on

Pricing from$0.20/ 1M input$1.25/ 1M output400Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by MiniMax

MiniMax M2.7-highspeed: throughput-optimized M2.7 variant delivering the self-evolving model's capabilities at higher output speed with Agent Teams

Available on

Pricing from$0.60/ 1M input$2.40/ 1M output205Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Z.ai

GLM-5-Turbo by Z.ai is a speed-optimized agentic model built for OpenClaw workflows, with thinking modes, reliable tool calling, and 200K context.

Available on

Pricing from$1.20/ 1M input$4.00/ 1M output203Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by xAI

Grok 4.3 by xAI combines always-on reasoning with a 1M context window, native vision, and agentic tool calling in one model.

Available on

Pricing from$1.25/ 1M input$2.50/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Alibaba

Qwen3.6-27B by Alibaba: 27B open-weight multimodal model with strong agentic coding (77.2 SWE-bench Verified) and reasoning across 262K context.

Available on

Pricing from$0.29/ 1M input$2.40/ 1M output262Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by OpenAI

GPT-5.1 by OpenAI, reasoning model with a 272K-token context, vision, and structured output for complex problem-solving.

Available on

Pricing from$1.25/ 1M input$10.00/ 1M output272Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Google

Google Gemini 3.5 Flash Lite — low-latency, highly cost-effective subagent model for high-volume automation

Available on

Pricing from$0.30/ 1M input$2.50/ 1M output1Mcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Moonshot

Kimi K2.5 by Moonshot: native multimodal 1T-parameter MoE with vision, 256K context, hybrid reasoning, and an agent swarm for parallel task execution.

Available on

Pricing from$0.40/ 1M input$2.25/ 1M output262Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU/EEA + US
by OpenAI

GPT-5 by OpenAI, multimodal reasoning model with a 272K-token context window for coding, analysis, and agentic work.

Available on

Pricing from$1.25/ 1M input$10.00/ 1M output272Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
EU + US
by Z.ai

GLM-5V-Turbo by Z.ai is a native multimodal model with vision, text, and video input that turns design mockups into front-end code.

Available on

Pricing from$1.20/ 1M input$4.00/ 1M output205Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US
by Alibaba

Qwen3.5-27B by Alibaba: 27B multimodal model with vision, tool calling, and thinking mode for reasoning and coding.

Available on

Pricing from$0.26/ 1M input$2.40/ 1M output262Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Alibaba

Qwen3.5 397B A17B by Alibaba: large sparse MoE model with 397B total (17B active) parameters for advanced reasoning and coding.

Available on

Pricing from$0.45/ 1M input$3.00/ 1M output262Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by Anthropic

Claude Opus 4.1 by Anthropic, improved agentic reasoning and multi-file code refactoring at 74.5% SWE-bench.

Available on

Pricing from$15.00/ 1M input$75.00/ 1M output200Kcontext

Privacy

Zero data retention
Enterprise
Training
No
Region
US
by MiniMax

MiniMax M2.5: 229B MoE model trained on hundreds of thousands of real environments, scoring 80.2% on SWE-Bench Verified and 76.3% on BrowseComp

Available on

Pricing from$0.30/ 1M input$1.10/ 1M output205Kcontext

Privacy

Zero data retention
Always-on
Training
No
Region
EU + US

Compare 300+ AI models on one gateway

The Opper model directory lists every large language model on the gateway, with image, video, and voice models alongside them. Filter by price, context window, capability, benchmark score, hosting region, and zero data retention, then run any of them through one OpenAI-compatible API and a single key. The default order leads with the models Artificial Analysis ranks highest on its intelligence index, which you can see in full on the LLM leaderboard. Put two models head to head on the comparison pages, or read how routing and fallbacks work on the LLM gateway.

Frequently asked questions

How do you compare AI models on Opper?

+
Every model on the gateway is listed with its price per million tokens, context window, capabilities, benchmark scores, hosting region, and zero-data-retention posture, so you can sort and filter to the right model, then switch between them through one OpenAI-compatible API and a single key. Compare models side by side

What is the cheapest LLM for my use case?

+
Sort the directory by input or output price to find the lowest cost per million tokens, from small budget models for high-volume work to frontier models for hard reasoning. Because every model runs through one gateway, you can route cheap models for easy calls and keep frontier models for the few that need them. See gateway pricing

How do I choose an LLM router?

+
An LLM router sends each request to the best model and falls back automatically when one is slow or unavailable. Opper routes across 300+ models behind a single API key, with built-in observability and European hosting, so you can change models without touching client code. How the LLM gateway works

Which AI models are hosted in the EU with zero data retention?

+
Filter the directory by European residency and zero data retention to see every model on a route where prompts and outputs are not stored or used for training. GDPR applies the same as it does across the EU, and a Data Processing Agreement is available on supported routes. Browse hosting providers

Does the directory include image, video, and voice models?

+
Yes. Alongside text LLMs, the directory lists image, video, text-to-speech, speech-to-text, and realtime models, priced per image, per second, or per character instead of per token, so you can compare media models on the same gateway and key as your text models. Compare every model

Start building with 300+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation