EU data residency

AI models hosted in the EU

Every model you can run in the EU through Opper, with where it runs, what the provider keeps and what it costs.

162 models270 routes21 providers10 countries

Model families in the EU

Every model that runs in the EU

Each card shows only the model's routes in Europe, so the hosts, zero data retention and prices are the ones you get there. Norway applies the GDPR in full through the EEA, so its routes are included. EU or EEA, and why it matters

by Anthropic

Claude Opus 5 is Anthropic's flagship for complex agentic coding and enterprise work, with a 1M token context window and adaptive thinking.

Available on

Pricing from$5.50/ 1M input$27.50/ 1M output$0.55/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI GPT-6 Sol for complex professional work, coding, and agent workflows.

Available on

Pricing from$2.40/ 1M input$12.00/ 1M output$0.24/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.6 Sol tops OpenAI's GPT-5.6 family, the frontier tier for complex coding, deep research, and long-running agents with a 1M token context.

Available on

Pricing from$5.00/ 1M input$30.00/ 1M output$0.50/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-5.3 by Z.ai is a 744B open-weight coding and agentic model, post-trained on the GLM-5.2 base with a 1M-token context and low, high, and max reasoning effort levels.

Available on

Pricing from$0.79/ 1M input$2.65/ 1M output$0.14/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K3 by Moonshot: open-weight 2.8 trillion parameter MoE with 104 billion active, native vision, and a 1M token context for long-horizon coding and agentic work.

Available on

Pricing from$3.00/ 1M input$15.00/ 1M output$0.45/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.6 Terra is the balanced middle tier of OpenAI's GPT-5.6 family, pairing near-flagship quality with everyday speed and a 1M token context.

Available on

Pricing from$2.00/ 1M input$12.00/ 1M output$0.20/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Opus 4.8 by Anthropic, highly capable model for complex reasoning, coding, and agentic work with 1M context and improved efficiency.

Available on

Pricing from$5.00/ 1M input$25.00/ 1M output$0.50/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-5.3-Flash by Z.ai is a natively multimodal 320B MoE with 18B active parameters, hybrid sparse and linear attention, a 1M-token context, and MIT open weights for coding and agent work.

Available on

Pricing from$0.13/ 1M input$0.50/ 1M output$0.03/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.8 Flash is Google's workhorse multimodal model for software engineering and agentic tasks, pairing 1M context with image, audio, video, and PDF input.

Available on

Pricing from$0.82/ 1M input$4.13/ 1M output$0.08/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Opus 4.7 by Anthropic, advanced vision and coding model with 1M context for complex long-running agentic workflows.

Available on

Pricing from$5.00/ 1M input$25.00/ 1M output$0.50/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.8-2.4T-A95B is the open-weight release of Alibaba's Qwen3.8-Max flagship, a 2.4 trillion parameter Mixture-of-Experts activating 95B per token.

Available on

Pricing from$2.50/ 1M input$6.00/ 1M output$0.63/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.8 Flash-Next by Alibaba is a multimodal MoE with a 125B backbone and 6B active parameters, an early look at the Qwen4 architecture with 262K native context for coding and agents.

Available on

Pricing from$0.20/ 1M input$0.50/ 1M output$0.05/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.7 Flash is Google's August 2026 Flash release, a large step up in coding and agentic automation over 3.6 while keeping the 1M-token multimodal context.

Available on

Pricing from$0.75/ 1M input$3.75/ 1M output$0.07/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4.1 Flash is a 552B multimodal MoE with a causal encoder-decoder design, 8B to 16B active parameters, a 1M-token context, and MIT weights for fast agentic coding.

Available on

Pricing from$0.23/ 1M input$1.14/ 1M output$0.01/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.4 is OpenAI's frontier model for complex professional work, with a 1M context window, vision, tool use, structured output, and configurable reasoning.

Available on

Pricing from$2.50/ 1M input$15.00/ 1M output$0.25/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.5 is OpenAI's newest frontier model, combining strong coding, reasoning, and web search and understanding multi-part tasks, with a 1M context window and vision.

Available on

Pricing from$5.00/ 1M input$30.00/ 1M output$0.50/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Sonnet 5 is Anthropic's balanced production model, bringing near-Opus agentic coding and tool use to a 1M token context window.

Available on

Pricing from$2.20/ 1M input$11.00/ 1M output$0.22/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.6 Luna, the fast high-volume tier of OpenAI's GPT-5.6 family, handles classification, summarization, and bulk pipelines with 1M context.

Available on

Pricing from$0.20/ 1M input$1.20/ 1M output$0.02/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI GPT-6 Luna for cost-sensitive, high-volume work and coding agents.

Available on

Pricing from$0.12/ 1M input$0.60/ 1M output$0.01/ 1M cache read1.1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4 Pro, flagship 1.6T MoE with 1M context and hybrid thinking for frontier reasoning and agents.

Available on

Pricing from$1.72/ 1M input$3.45/ 1M output$0.14/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4 Pro 0813 is the general availability build of DeepSeek's 1M-context flagship, tuned for agentic tool use and terminal-driven work.

Available on

Pricing from$1.14/ 1M input$3.41/ 1M output$0.04/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4 Flash, lightweight 284B MoE with 1M context and hybrid thinking for cost-efficient agents.

Available on

Pricing from$0.14/ 1M input$0.28/ 1M output$0.03/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.6 Flash, Google's token-efficient multimodal Flash model, cuts output tokens by 17% while improving coding and agentic planning.

Available on

Pricing from$0.75/ 1M input$3.75/ 1M output$0.07/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.5 Flash by Google pairs frontier-level intelligence with fast throughput and a 1M token context for agentic and coding tasks.

Available on

Pricing from$1.50/ 1M input$9.00/ 1M output$0.15/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.3 Codex is OpenAI's most capable agentic coding model, reaching state-of-the-art on SWE-Bench Pro and Terminal-Bench while using fewer tokens than prior models.

Available on

Pricing from$1.75/ 1M input$14.00/ 1M output$0.17/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Opus 4.6 by Anthropic, professional knowledge-work model with 1M context, extended thinking, and 128K token outputs.

Available on

Pricing from$5.00/ 1M input$25.00/ 1M output$0.50/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by MiniMax

MiniMax M3: natively multimodal 428B Mixture-of-Experts model with ~23B active parameters, MiniMax Sparse Attention for long context, and 59.0% on SWE-Bench Pro

Available on

Pricing from$0.40/ 1M input$2.00/ 1M output$0.10/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Opus 4.5 by Anthropic, highly capable model for coding and agents with extended thinking and a tunable effort parameter.

Available on

Pricing from$5.00/ 1M input$25.00/ 1M output$0.50/ 1M cache read200Kcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM 5 by Zhipu: 744B frontier MoE model with 40B active, reasoning-first design, and Huawei Ascend training.

Available on

Pricing from$1.00/ 1M input$3.20/ 1M output$0.25/ 1M cache read203Kcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K2.6 by Moonshot: open-weight agentic model with long-horizon coding, 300-agent swarms, video understanding, and 256K context.

Available on

Pricing from$0.47/ 1M input$2.58/ 1M output$0.13/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-5-Turbo by Z.ai is a speed-optimized agentic model built for OpenClaw workflows, with thinking modes, reliable tool calling, and 200K context.

Available on

Pricing from$1.20/ 1M input$4.00/ 1M output$0.30/ 1M cache read203Kcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-5.1 by Z.ai is a 754B open-weight (MIT) coding and agentic reasoning model with 202K context that tops SWE-Bench Pro at 58.4.

Available on

Pricing from$1.40/ 1M input$4.40/ 1M output$0.35/ 1M cache read203Kcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K2.7 Code by Moonshot: open-weight, code-specialized agentic model on a 1 trillion parameter MoE, tuned for long-horizon software engineering and tool use.

Available on

Pricing from$0.71/ 1M input$3.30/ 1M output$0.18/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by xAI

Grok 4.3 by xAI combines always-on reasoning with a 1M context window, native vision, and agentic tool calling in one model.

Available on

Pricing from$1.25/ 1M input$2.50/ 1M output$0.20/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.1 by OpenAI, reasoning model with a 272K-token context, vision, and structured output for complex problem-solving.

Available on

Pricing from$1.25/ 1M input$10.00/ 1M output$0.13/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.4 Mini is OpenAI's strongest small model, improving over GPT-5 Mini in coding and reasoning while running more than 2x faster, with a 400k context window.

Available on

Pricing from$0.75/ 1M input$4.50/ 1M output$0.07/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K2.5 by Moonshot: native multimodal 1T-parameter MoE with vision, 256K context, hybrid reasoning, and an agent swarm for parallel task execution.

Available on

Pricing from$0.40/ 1M input$2.60/ 1M output$0.13/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU/EEA
by Z.ai

GLM-5V-Turbo by Z.ai is a native multimodal model with vision, text, and video input that turns design mockups into front-end code.

Available on

Pricing from$1.20/ 1M input$4.00/ 1M output$0.30/ 1M cache read203Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5 by OpenAI, multimodal reasoning model with a 272K-token context window for coding, analysis, and agentic work.

Available on

Pricing from$1.25/ 1M input$10.00/ 1M output$0.13/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by MiniMax

MiniMax M2.5: 229B MoE model trained on hundreds of thousands of real environments, scoring 80.2% on SWE-Bench Verified and 76.3% on BrowseComp

Available on

Pricing from$0.30/ 1M input$1.20/ 1M output$0.08/ 1M cache read197Kcontext

Privacy

ZDR
Training
No
Region
EU
by MiniMax

MiniMax M2.7: self-evolving 229B model with 56.22% SWE-Pro, native multi-agent teams, and 1495 ELO on GDPval-AA (highest open-weight)

Available on

Pricing from$0.68/ 1M input$2.73/ 1M output197Kcontext

Privacy

ZDR
Training
No
Region
EU
by xAI

Grok 4 by xAI is a reasoning model with vision, native tool use, and real-time web and X search across a 256k context window.

Available on

Pricing from$3.00/ 1M input$15.00/ 1M output$0.75/ 1M cache read256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.5 Flash Lite is Google's low-latency workhorse for high-volume automation, pairing 1M context with audio, video, and PDF understanding.

Available on

Pricing from$0.30/ 1M input$2.50/ 1M output$0.03/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM 4.7 by Zhipu: 358B reasoning model with interleaved thinking, 200K context, and frontier coding performance.

Available on

Pricing from$0.80/ 1M input$2.84/ 1M output205Kcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-4.7-FP8 by Zhipu: FP8-quantized GLM-4.7 Flash, shrinking the BF16 footprint to roughly 30GB for efficient inference.

Available on

Pricing from$0.80/ 1M input$2.84/ 1M output

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V3.2, GPT-5-class open MoE that integrates reasoning directly into tool-use.

Available on

Pricing from$0.30/ 1M input$0.50/ 1M output$0.08/ 1M cache read164Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5 397B A17B by Alibaba: large sparse MoE model with 397B total (17B active) parameters for advanced reasoning and coding.

Available on

Pricing from$0.68/ 1M input$4.10/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6-27B by Alibaba: 27B open-weight multimodal model with strong agentic coding (77.2 SWE-bench Verified) and reasoning across 262K context.

Available on

Pricing from$0.15/ 1M input$1.00/ 1M output$0.07/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Sonnet 4.5 by Anthropic, strong coding and agentic model scoring 61.4% on the OSWorld computer-use benchmark, 30+ hour focus.

Available on

Pricing from$3.00/ 1M input$15.00/ 1M output$0.30/ 1M cache read200Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.4 Nano is the fastest, lightest GPT-5.4 variant, built for classification, data extraction, and lightweight coding subagents that need low latency.

Available on

Pricing from$0.20/ 1M input$1.25/ 1M output$0.02/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5 Mini is a fast, cost-efficient GPT-5 model for well-defined tasks and precise prompts, delivering near-frontier intelligence for high-volume workloads.

Available on

Pricing from$0.25/ 1M input$2.00/ 1M output$0.03/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.1 Codex Mini by OpenAI, fast, cost-efficient code model with a 272K context for everyday development tasks.

Available on

Pricing from$0.25/ 1M input$2.00/ 1M output$0.03/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemma 4 31B is Google DeepMind's 30.7B parameter dense open model, the largest of the Gemma 4 family, with 256K context, thinking mode, and native function calling.

Available on

Pricing from$0.11/ 1M input$0.34/ 1M output$0.02/ 1M cache read256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6 35B-A3B by Alibaba: sparse MoE (35B total, 3B active) with vision, reasoning, and agentic coding for repository-level tasks.

Available on

Pricing from$0.28/ 1M input$1.71/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6 35B A3B FP8 by Alibaba: fine-grained FP8 quantized variant of the sparse MoE model with near-identical performance.

Available on

Pricing from$0.34/ 1M input$1.37/ 1M output

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5-122B-A10B by Alibaba: sparse MoE multimodal model with 122B total (10B active) parameters and 262K context for reasoning, vision, and agentic tasks.

Available on

Pricing from$0.50/ 1M input$3.50/ 1M output$0.13/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Muse Glimmer is Meta's 30B open-weight agentic model, distilled from Muse Spark and built to plan, call tools, and recover from failures on a single consumer GPU.

Available on

Pricing from$0.23/ 1M input$1.14/ 1M output$0.06/ 1M cache read128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude Haiku 4.5 by Anthropic, fast model matching Claude Sonnet 4 coding performance at a fraction of the cost, more than twice the speed.

Available on

Pricing from$1.00/ 1M input$5.00/ 1M output$0.10/ 1M cache read200Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 2.5 Pro by Google, an advanced reasoning LLM with 1M context and thinking mode, strong at complex coding, math, STEM, and deep document analysis.

Available on

Pricing from$1.25/ 1M input$10.00/ 1M output$0.13/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM 4.7 Flash by Zhipu: 30B MoE (3B active) with 200K context, tuned for local deployment and fast agentic coding.

Available on

Pricing from$0.03/ 1M input$0.20/ 1M output$0.01/ 1M cache read203Kcontext

Privacy

ZDR
Training
No
Region
EU
by xAI

Grok 4.20 Non-Reasoning by xAI delivers fast, direct responses across a 2M context window with real-time X data.

Available on

Pricing from$2.00/ 1M input$6.00/ 1M output$0.20/ 1M cache read2Mcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5 9B by Alibaba: 9B multimodal model with vision, tool calling, and reasoning for balanced performance and efficiency.

Available on

Pricing from$0.04/ 1M input$0.07/ 1M output$0.02/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek R1 0528: updated reasoning model with deeper chain-of-thought, system prompts, and structured output

Available on

Pricing from$0.66/ 1M input$2.60/ 1M output$0.17/ 1M cache read164Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 2.5 Flash (Google). Fast multimodal model for text, images, video, and audio, with reasoning, tools, and structured output.

Available on

Pricing from$0.30/ 1M input$2.50/ 1M output$0.03/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5 Nano is the fastest, lightest GPT-5 variant, built for summarization, classification, and simple tasks that need minimal latency.

Available on

Pricing from$0.05/ 1M input$0.40/ 1M output$0.0050/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Small 4 (March 2026) unifies reasoning, coding, and multimodal in one MoE model with 256K context and configurable reasoning effort.

Available on

Pricing from$0.15/ 1M input$0.60/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-OSS 20B by OpenAI: open-weight Apache 2.0 model, 21B params (3.6B active), near o3-mini, for efficient and on-device reasoning.

Available on

Pricing from$0.0090/ 1M input$0.04/ 1M output$0.0045/ 1M cache read131Kcontext

Privacy

ZDR
Training
No
Region
EU/EEA
by Alibaba

Qwen3 Coder Next by Alibaba, an 80B ultra-sparse MoE model with 3B active parameters and 256K context for efficient local coding agents.

Available on

Pricing from$0.57/ 1M input$2.28/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 3.1 Flash Lite by Google, a cost-effective, low-latency multimodal LLM with 1M context, thinking mode, and tool use for efficient task automation.

Available on

Pricing from$0.25/ 1M input$1.50/ 1M output$0.03/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 2.5 Flash Lite by Google, an efficient multimodal LLM with 1M context, optimized for cost and speed across text, vision, and structured tasks.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output$0.01/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-4o by OpenAI is a capable general-purpose omnimodal model with a 128K context window, vision input, structured output, and tool use.

Available on

Pricing from$2.75/ 1M input$11.00/ 1M output$1.38/ 1M cache read128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 VL 30B A3B Instruct by Alibaba, an efficient open-weight MoE vision-language model with 3B active parameters and 131K context.

Available on

Pricing from$0.23/ 1M input$0.91/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.3 70B Instruct: refined 70B model with improved multilingual performance, structured output, and tool calling for reasoning workloads.

Available on

Pricing from$0.68/ 1M input$1.02/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.3 70B Instruct in an FP8 build, Meta's multilingual open-weight workhorse with 128K context, quantized for faster serving at near-identical quality.

Available on

Pricing from$1.14/ 1M input$1.14/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Amazon

Amazon Nova Pro, a balanced multimodal model combining accuracy, speed and cost for text, image and video understanding.

Available on

Pricing from$0.80/ 1M input$3.20/ 1M output300Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.1 8B Instruct: efficient open-weight small model with a 131K context, tool calling, and structured output for production deployments.

Available on

Pricing from$0.01/ 1M input$0.02/ 1M output$0.0050/ 1M cache read16Kcontext

Privacy

ZDR
Training
No
Region
EU
by Amazon

Amazon Nova Lite, a cost-effective multimodal model with 300K context for document analysis, visual Q&A, and video understanding.

Available on

Pricing from$0.06/ 1M input$0.24/ 1M output300Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-4 by OpenAI, an older high-intelligence text model with function calling, now superseded by newer options.

Available on

Pricing from$30.00/ 1M input$60.00/ 1M output8Kcontext

Privacy

ZDR
Training
No
Region
EU
by Amazon

Amazon Nova Micro, a fast text-only model for real-time summarization, translation, classification and simple reasoning.

Available on

Pricing from$0.04/ 1M input$0.14/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Anthropic

Claude 3 Haiku by Anthropic, fast and affordable vision model for customer support and batch processing of large documents.

Available on

Pricing from$0.25/ 1M input$1.25/ 1M output200Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-5.3 Chat by OpenAI, chat-tuned model with a 272K context, vision, and tools for conversational use.

Available on

Pricing from$1.75/ 1M input$14.00/ 1M output$0.17/ 1M cache read272Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-3.5 Turbo by OpenAI, a legacy text and code model with fast inference, now superseded by more capable alternatives.

Available on

Pricing from$0.50/ 1M input$1.50/ 1M output16Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI's 1.55B-parameter speech-to-text model supporting 99 languages with 10-20% error reduction over Whisper Large v2, released November 2023.

Available on

Pricing$0.0023 / min

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI's pruned 809M-parameter speech-to-text model with 4 decoder layers, delivering much faster multilingual transcription at minimal quality loss.

Available on

Pricing$0.0023 / min

Privacy

ZDR
Training
No
Region
EU
by Google

Gemini 2.5 Flash Image (Google). Fast multimodal model with native image generation and conversational editing.

Available on

Pricing from$0.30/ 1M input$2.50/ 1M output1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Google Gemma 4 31B IT, instruction-tuned 31B dense model with multimodal reasoning, function calling, and 256K context.

Available on

Pricing from$0.07/ 1M input$0.35/ 1M output$0.01/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Google Gemma 4 26B-A4B IT, mixture-of-experts open model with only 3.8B active parameters, multimodal reasoning, and function calling.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output$0.05/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Small 3.1 24B is an open-weight instruction-tuned model with 128K context, vision, and function calling at around 150 tokens per second.

Available on

Pricing from$0.37/ 1M input$0.37/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Devstral 2: code agents model for software engineering, 256k context, 123B params, 72.2% on SWE-bench Verified.

Available on

Pricing from$0.40/ 1M input$2.00/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 14B: largest Ministral 3 model, 256k context, Apache 2.0, tuned for local and edge deployment.

Available on

Pricing from$0.20/ 1M input$0.20/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 3B: smallest Ministral 3 model, 256k context, Apache 2.0, ultra-efficient edge deployment.

Available on

Pricing from$0.10/ 1M input$0.10/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 8B, edge-optimized multimodal model with 256k context, vision, and structured outputs.

Available on

Pricing from$0.15/ 1M input$0.15/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Large 3, open-weight mixture-of-experts model with 256k context and native multimodal vision.

Available on

Pricing from$0.50/ 1M input$1.50/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Codestral 25.08: code-optimized model for fill-in-the-middle completion and fast code generation, 128k context.

Available on

Pricing from$0.30/ 1M input$0.90/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Medium 3.1, frontier-class multimodal model balancing reasoning, vision, and agentic control.

Available on

Pricing from$0.40/ 1M input$2.00/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Pixtral Large is a 124B multimodal vision model with 128K context, built for document analysis, visual reasoning, and frontier image understanding.

Available on

Pricing from$2.00/ 1M input$6.00/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Latest routes to the current Voxtral Mini, a 32K-context speech model for voice agents with native audio understanding.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime 2602 delivers streaming live transcription with 13-language support, speaker diarization, and a 4B edge-deployable model.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime Latest routes to the current streaming transcription model, a 4B 13-language model for voice agents and live apps.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime by Mistral: streaming speech-to-text with configurable delay for live transcription, voice agents, and real-time subtitling.

Available on

Pricing$0.006 / min

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Small 2507 by Mistral: 24B audio-native LLM with function calling and structured outputs for speech understanding and agentic chat.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Small (latest) by Mistral: open-weight 24B audio-input instruct model that tracks the newest snapshot, with function calling and structured outputs.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by xAI

Grok 4.20 Reasoning by xAI generates visible chain-of-thought across a 2M context window with real-time data and tool use.

Available on

Pricing from$2.00/ 1M input$6.00/ 1M output$0.20/ 1M cache read2Mcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3-VL-Flash is the speed-focused tier of Alibaba's Qwen3-VL vision line, reading documents, charts, and video at low latency with 262K context.

Available on

Pricing from$0.05/ 1M input$0.40/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3-VL-Plus, Alibaba's flagship vision-language model for document understanding, video analysis, and visual agents, with 262K context and hybrid thinking.

Available on

Pricing from$0.20/ 1M input$1.60/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 30B A3B by Alibaba, an efficient MoE model with 3B active parameters and switchable thinking and non-thinking modes.

Available on

Pricing from$0.10/ 1M input$0.28/ 1M output$0.05/ 1M cache read41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 8B dense model on Fireworks

Available on

Pricing from$0.45/ 1M input$0.45/ 1M output$0.22/ 1M cache read33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 14B, Alibaba's April 2025 dense model with switchable thinking and direct modes for reasoning and tool use at 40K context.

Available on

Pricing from$0.10/ 1M input$0.22/ 1M output41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 32B by Alibaba, a dense 32B model with switchable thinking and non-thinking modes for reasoning and speed.

Available on

Pricing from$0.10/ 1M input$0.28/ 1M output$0.05/ 1M cache read16Kcontext

Privacy

ZDR
Training
No
Region
EU
by Amazon

Amazon Nova 2 Lite, a fast multimodal model with extended thinking, 1M context, web grounding and code execution for reasoning tasks.

Available on

Pricing from$0.30/ 1M input$2.50/ 1M output1Mcontext

Privacy

ZDR
Training
No
Region
EU
by NVIDIA

NVIDIA Nemotron 3 Nano 30B is a 30B Mixture-of-Experts model with about 3.5B active parameters, tuned for reasoning, tool-calling, and code.

Available on

Pricing from$0.05/ 1M input$0.20/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU/EEA
by Talkie LM

Talkie 1930 is a 13B open-weight model trained only on pre-1931 English text, giving a 1930 knowledge cutoff for historical reasoning and contamination-free research.

Available on

Pricing from—/ 1M input—/ 1M output8Kcontext

Privacy

ZDR
Training
No
Region
EU
by Swiss AI Initiative

Apertus 70B, the Swiss AI Initiative's fully open model from EPFL and ETH Zurich, trained on 15T tokens with over 1,800 natively supported languages.

Available on

Pricing from$0.46/ 1M input$2.39/ 1M output66Kcontext

Privacy

ZDR
Training
No
Region
EU
by Regolo

Brick Complexity Pro, Regolo's hosted prompt-complexity classifier that grades queries easy, medium, or hard so routing picks the right model tier.

Available on

Pricing from$0.11/ 1M input$0.46/ 1M output

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-l-raw GreenPT Backed by Mistral Small 3.2 24B Direct access to the same GreenPT-backed model as green-l without the built-in system prompt. Input €0.25 Output €0.80 Context 128k Max output 32k Released Jun 2025 Text Images Documents Multilingual No system prompt

Available on

Pricing from$0.28/ 1M input$0.91/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-r-raw GreenPT Backed by GPT-OSS Direct access to the same reasoning stack as green-r without the GreenPT system prompt. Input €0.35 Output €0.95 Context 128k Max output 32k Released Aug 2025 Text Images Documents Multilingual No system prompt

Available on

Pricing from$0.40/ 1M input$1.08/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-l GreenPT Backed by Mistral Small 3.2 24B GreenPT-branded chat model tuned for multilingual writing, image understanding, and Dutch grammar guardrails. Input €0.25 Output €0.80 Context 128k Max output 32k Released Jun 2025 Text Images Documents Multilingual Writing assistant Dutch grammar guardrails

Available on

Pricing from$0.28/ 1M input$0.91/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-r GreenPT Backed by GPT-OSS GreenPT-branded reasoning model for advanced analysis, writing, and content generation. Input €0.35 Output €0.95 Context 128k Max output 32k Released Aug 2025 Text Images Documents Multilingual Advanced reasoning Writing & content generation

Available on

Pricing from$0.40/ 1M input$1.08/ 1M output123Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Sao10K's L3.3 70B Euryale v2.3, a 70B Llama 3.3 creative roleplay model and the direct successor to v2.2.

Available on

Pricing from$0.65/ 1M input$0.75/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Barcelona Supercomputing Center

ALIA 40B Instruct 2601, Spain's publicly funded 40B multilingual model from the Barcelona Supercomputing Center, Apache 2.0 and pretrained across 35 European languages.

Available on

Pricing from$0.30/ 1M input$0.60/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

TheDrummer's UnslopNemo 12B v4.1, a Mistral Nemo fine-tune for adventure writing and roleplay with a 32K context.

Available on

Pricing from$0.40/ 1M input$0.40/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Kev-4B System One decision model, a fine-tune by Jared Palmer of Alibaba's Qwen3.5-4B-Base, for typed yes/no decisions (noul), classification (choice) and rubric scoring (score), returning a probability for every option instead of text. Text or structured text input only, up to 8,192 tokens for the state plus one question; trained on states of up to 384 tokens, so accuracy on long documents is lower. Use POST /v3/compat/v1/systemone with model opper/kev-4b. Weights by Jared Palmer: https://huggingface.co/jaredpalmer/kev-4b. Base model: Alibaba's Qwen3.5-4B-Base, https://huggingface.co/Qwen/Qwen3.5-4B-Base.

Available on

Pricing from$0.04/ 1M input$0/ 1M output8Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Undi95's ReMM SLERP L2 13B, a Llama 2 SLERP merge recreating the MythoMax recipe with updated component models.

Available on

Pricing from$0.45/ 1M input$0.65/ 1M output6Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Mythomax L2 13B by Gryphe is an open-weight Llama 2 model tuned for character-driven roleplay and creative storytelling with a consistent voice across longer exchanges.

Available on

Pricing from$0.06/ 1M input$0.06/ 1M output4Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Laya System One decision model by ConvAI Innovations (Nandha Kishor M), a non-autoregressive ModernBERT-large encoder (421M) for typed yes/no decisions (noul), classification (choice) and rubric scoring (score), returning a probability for every option instead of text. Text or structured text input only; the English checkpoint reads 512 tokens and truncates longer input. Use POST /v3/compat/v1/systemone with model berget/convaiinnovations/laya. Weights: https://huggingface.co/convaiinnovations/laya (Apache-2.0).

Available on

Pricing from$0.05/ 1M input$0/ 1M output512context

Privacy

ZDR
Training
No
Region
EU
by KBLab

KBLab's KB Whisper Large is a 2-billion-parameter Swedish speech-to-text model trained on 50,000 hours of audio, cutting word error rate by an average 47% versus Whisper-large-v3.

Available on

Pricing$0.0023 / min

Privacy

ZDR
Training
No
Region
EU
by Klang

Pianissimo from Klang — speech to text model on the Opper gateway.

Available on

Pricing$0.004 / min

Privacy

ZDR
Training
No
Region
EU
by Evroc

roc, Evroc's enterprise AI agent for writing, data analysis, coding support, and knowledge retrieval over company documents and systems.

Available on

Pricing from$2.84/ 1M input$11.38/ 1M output

Privacy

ZDR
Training
No
Region
EU

Keep every call in the EU

Every route has its own model id. Call that id on any plan and the request is served on that route, in that location. The examples below call Claude Opus 5.5 on AWS Bedrock in Sweden.

Set it up with your agent

Copy this into a coding agent like Claude Code, Cursor or Codex and it will wire up Opper and call this route.

Or call it directly

import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.OPPER_API_KEY,
baseURL: "https://api.opper.ai/v3/compat",
});
const completion = await client.chat.completions.create({
model: "aws/claude-opus-5-5",
messages: [{ role: "user", content: "Hello" }],
});
console.log(completion.choices[0].message.content);

Lock it in with a rule

A model access rule with Inference location: EU, set for your organization or a project, means no LLM or embedding call can reach a route outside the EU, even by mistake. Rules are part of the Control Plane. How model access rules work

How much EU do you need?

For most teams EU hosting is the right setting. Inference runs in European data centres operated by AWS, Azure, Google and European hosts, Opper itself runs in AWS Stockholm, and many routes carry zero data retention as well.

Need European ownership end to end?

Opper can run as a dedicated instance on evroc, the Swedish cloud, limited to European-owned providers.

Deployment options

Who you contract with

Company
Opper Technology AB, Sweden, reg. no. 559446-3720
Certification
Processor
Opper, for every modelSub-processor list
DPA
One DPA covers the whole catalogueRead the DPA

Frequently asked questions

Which AI models are hosted in the EU?

+
162 models on the Opper gateway run on European routes as of 2026-09-28, 125 of them language models, across 21 providers in Belgium, Finland, France, Germany, Ireland, Italy, the Netherlands, Norway, Spain and Sweden. By family: Claude 11, OpenAI 25, Gemini 11, Mistral 28, DeepSeek 7, Qwen 19, GLM 10, Kimi 4, Llama 3, Grok 4. This page lists every one of them with where it runs, what the provider keeps and what it costs.

Can I use Claude, GPT and Gemini with EU data residency?

+
Yes. Claude: 11 of 13 models, on AWS Bedrock, Google Cloud and Azure in Sweden, the EU multi-region, France and Belgium. OpenAI: 25 of 44 models, on Azure, Geodd, AWS Bedrock, evroc, GreenPT and 4 more in Sweden, Norway, Germany, Italy, France and the EU multi-region. Gemini: 11 of 22 models, on Google Cloud in the EU multi-region and the Netherlands. Call the EU route's model id and the request is served in the EU. Claude in the EU

Is an EU-hosted model GDPR (DSGVO) compliant?

+
Hosting in the EU settles where your data is processed, which is one of the questions the GDPR (DSGVO in German) asks. The rest depends on your own processing and on the terms of the service you use. On Opper, one Data Processing Agreement covers every model, Opper is your single sub-processor for all of them, and every route lists what the provider keeps and whether it trains on your data. Opper is certified against ISO/IEC 27001:2022; the controls, policies and sub-processors are on https://trust.opper.ai. Read the DPA

Do I need an EU-owned provider?

+
Most teams don't. EU hosting keeps processing in the EU, and every route states what the provider keeps. If a policy requires European ownership end to end, Opper can run as a dedicated instance on evroc, the Swedish cloud, with European-owned providers only, as part of Enterprise. Deployment options

How do I make sure a call never leaves the EU?

+
On any plan, call an EU route's model id and the request is served there. To lock it in, set a model access rule with Inference location: EU for your organization or a project, and no LLM or embedding call can reach a route outside the EU, even by mistake. Rules are part of the Control Plane. Model access rules

Do EU routes cost more?

+
Sometimes. Comparing input prices, of the 90 EU-hosted models that also run elsewhere, 41 cost the same, 14 cost less and 35 cost more on their cheapest European route. Where AWS, Azure or Google charge more for their EU region, the uplift is between 10 and 20%, typically 10%. Other hosts in Europe set their own prices, so for open-weight models the difference depends on the host. Each model page lists every route's own price.

Is zero data retention available on EU routes?

+
Yes, on pay-as-you-go. 121 models have a route in Europe where the provider keeps nothing and trains on nothing, Claude on AWS Bedrock among them. Routes marked abuse monitoring keep content for abuse monitoring only; the window is shown where the provider states one. For OpenAI models on Azure and Gemini on Google Vertex, the zero-retention routes without abuse monitoring (azure-zdr and vertexai-zdr) need a signed Enterprise agreement. How zero data retention works

Read this list from an agent

This list as data. The same routes as JSON at opper.ai/models/eu.json (model id, provider, country, zero data retention, training and price per route) and as markdown at opper.ai/models/eu.md.

The whole catalogue. The model catalogue is public JSON at https://api.opper.ai/v3/models, no API key needed. Filter it with query parameters, for example ?type=llm&limit=3, and page through it with limit and offset. Use it when you can fetch a URL but cannot connect to an MCP server. Coding agents that use MCP can connect to https://api.opper.ai/mcp. Two of the server's tools need no account and no sign-in: list_models searches the model catalogue by name, type, provider or capability, and get_guide returns short setup guides. Everything that touches your account needs you to approve the connection in the browser first. About the MCP server

Run your AI in the EU

One API key for every EU route and one DPA for every model, with each call served where you choose.

Get startedRead the DPA