Sovereign AI

Sovereign AI models hosted in Europe

Every model you can run with a European-owned provider in Europe, with where it runs, what the provider keeps and what it costs.

Sovereign AI on Opper
90 models154 routes15 providers9 home countries

Every model on a European-owned provider

Each card shows only the model's routes on European-owned providers, so the hosts, zero data retention and prices are the ones you get there. European-owned means the provider is headquartered in the EU, the EEA, Switzerland or the UK.

by Z.ai

GLM-5.3 by Z.ai is a 744B open-weight coding and agentic model, post-trained on the GLM-5.2 base with a 1M-token context and low, high, and max reasoning effort levels.

Available on

Pricing from$0.55/ 1M input$3.36/ 1M output$0.14/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K3 by Moonshot: open-weight 2.8 trillion parameter MoE with 104 billion active, native vision, and a 1M token context for long-horizon coding and agentic work.

Available on

Pricing from$3.00/ 1M input$15.00/ 1M output$0.45/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.8-2.4T-A95B is the open-weight release of Alibaba's Qwen3.8-Max flagship, a 2.4 trillion parameter Mixture-of-Experts activating 95B per token.

Available on

Pricing from$2.50/ 1M input$6.00/ 1M output$0.63/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.8 Flash-Next by Alibaba is a multimodal MoE with a 125B backbone and 6B active parameters, an early look at the Qwen4 architecture with 262K native context for coding and agents.

Available on

Pricing from$0.20/ 1M input$0.50/ 1M output$0.05/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4.1 Flash is a 552B multimodal MoE with a causal encoder-decoder design, 8B to 16B active parameters, a 1M-token context, and MIT weights for fast agentic coding.

Available on

Pricing from$0.22/ 1M input$1.12/ 1M output$0.01/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4 Pro, flagship 1.6T MoE with 1M context and hybrid thinking for frontier reasoning and agents.

Available on

Pricing from$2.00/ 1M input$4.00/ 1M output$0.50/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by DeepSeek

DeepSeek V4 Flash, lightweight 284B MoE with 1M context and hybrid thinking for cost-efficient agents.

Available on

Pricing from$0.25/ 1M input$0.30/ 1M output$0.06/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM-5.2 by Z.ai is a 744B open-weight (MIT) mixture-of-experts coding model with a 1M-token context, IndexShare sparse attention, and selectable thinking-effort levels.

Available on

Pricing from$1.23/ 1M input$4.39/ 1M output$0.14/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by MiniMax

MiniMax M3: natively multimodal 428B Mixture-of-Experts model with ~23B active parameters, MiniMax Sparse Attention for long context, and 59.0% on SWE-Bench Pro

Available on

Pricing from$0.40/ 1M input$2.00/ 1M output$0.07/ 1M cache read1Mcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K2.6 by Moonshot: open-weight agentic model with long-horizon coding, 300-agent swarms, video understanding, and 256K context.

Available on

Pricing from$0.47/ 1M input$2.45/ 1M output$0.10/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Moonshot

Kimi K2.7 Code by Moonshot: open-weight, code-specialized agentic model on a 1 trillion parameter MoE, tuned for long-horizon software engineering and tool use.

Available on

Pricing from$0.67/ 1M input$3.35/ 1M output$0.18/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by MiniMax

MiniMax M2.5: 229B MoE model trained on hundreds of thousands of real environments, scoring 80.2% on SWE-Bench Verified and 76.3% on BrowseComp

Available on

Pricing from$0.37/ 1M input$1.48/ 1M output197Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5 397B A17B by Alibaba: large sparse MoE model with 397B total (17B active) parameters for advanced reasoning and coding.

Available on

Pricing from$0.67/ 1M input$4.03/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6-27B by Alibaba: 27B open-weight multimodal model with strong agentic coding (77.2 SWE-bench Verified) and reasoning across 262K context.

Available on

Pricing from$0.15/ 1M input$1.00/ 1M output$0.07/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Gemma 4 31B is Google DeepMind's 30.7B parameter dense open model, the largest of the Gemma 4 family, with 256K context, thinking mode, and native function calling.

Available on

Pricing from$0.11/ 1M input$0.34/ 1M output$0.02/ 1M cache read256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6 35B-A3B by Alibaba: sparse MoE (35B total, 3B active) with vision, reasoning, and agentic coding for repository-level tasks.

Available on

Pricing from$0.28/ 1M input$1.68/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.6 35B A3B FP8 by Alibaba: fine-grained FP8 quantized variant of the sparse MoE model with near-identical performance.

Available on

Pricing from$0.34/ 1M input$1.34/ 1M output

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5-122B-A10B by Alibaba: sparse MoE multimodal model with 122B total (10B active) parameters and 262K context for reasoning, vision, and agentic tasks.

Available on

Pricing from$1.12/ 1M input$4.71/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Z.ai

GLM 4.7 Flash by Zhipu: 30B MoE (3B active) with 200K context, tuned for local deployment and fast agentic coding.

Available on

Pricing from$0.03/ 1M input$0.20/ 1M output$0.01/ 1M cache read131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3.5 9B by Alibaba: 9B multimodal model with vision, tool calling, and reasoning for balanced performance and efficiency.

Available on

Pricing from$0.04/ 1M input$0.07/ 1M output$0.02/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-OSS 120B by OpenAI: open-weight Apache 2.0 reasoning model, 117B params (5.1B active), near o4-mini on reasoning, tools and structured output.

Available on

Pricing from$0.17/ 1M input$0.66/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Small 4 (March 2026) unifies reasoning, coding, and multimodal in one MoE model with 256K context and configurable reasoning effort.

Available on

Pricing from$0.15/ 1M input$0.60/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

GPT-OSS 20B by OpenAI: open-weight Apache 2.0 model, 21B params (3.6B active), near o3-mini, for efficient and on-device reasoning.

Available on

Pricing from$0.0090/ 1M input$0.04/ 1M output$0.0045/ 1M cache read131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 Coder Next by Alibaba, an 80B ultra-sparse MoE model with 3B active parameters and 256K context for efficient local coding agents.

Available on

Pricing from$0.56/ 1M input$2.24/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 VL 30B A3B Instruct by Alibaba, an efficient open-weight MoE vision-language model with 3B active parameters and 131K context.

Available on

Pricing from$0.22/ 1M input$0.90/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.3 70B Instruct: refined 70B model with improved multilingual performance, structured output, and tool calling for reasoning workloads.

Available on

Pricing from$0.67/ 1M input$1.01/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.3 70B Instruct in an FP8 build, Meta's multilingual open-weight workhorse with 128K context, quantized for faster serving at near-identical quality.

Available on

Pricing from$1.12/ 1M input$1.12/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Meta

Llama 3.1 8B Instruct: efficient open-weight small model with a 131K context, tool calling, and structured output for production deployments.

Available on

Pricing from$0.01/ 1M input$0.02/ 1M output$0.0050/ 1M cache read16Kcontext

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI's 1.55B-parameter speech-to-text model supporting 99 languages with 10-20% error reduction over Whisper Large v2, released November 2023.

Available on

Pricing$0.0022 / min

Privacy

ZDR
Training
No
Region
EU
by OpenAI

OpenAI's pruned 809M-parameter speech-to-text model with 4 decoder layers, delivering much faster multilingual transcription at minimal quality loss.

Available on

Pricing$0.0022 / min

Privacy

ZDR
Training
No
Region
EU
by Google

Google Gemma 4 26B-A4B IT, mixture-of-experts open model with only 3.8B active parameters, multimodal reasoning, and function calling.

Available on

Pricing from$0.09/ 1M input$0.30/ 1M output$0.05/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Google

Google Gemma 4 31B IT, instruction-tuned 31B dense model with multimodal reasoning, function calling, and 256K context.

Available on

Pricing from$0.03/ 1M input$0.17/ 1M output$0.01/ 1M cache read262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Medium 3.5 from Mistral — text model on the Opper gateway.

Available on

Pricing from$1.50/ 1M input$7.50/ 1M output262Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Small 3.1 24B is an open-weight instruction-tuned model with 128K context, vision, and function calling at around 150 tokens per second.

Available on

Pricing from$0.36/ 1M input$0.36/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 14B: largest Ministral 3 model, 256k context, Apache 2.0, tuned for local and edge deployment.

Available on

Pricing from$0.20/ 1M input$0.20/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 3B: smallest Ministral 3 model, 256k context, Apache 2.0, ultra-efficient edge deployment.

Available on

Pricing from$0.10/ 1M input$0.10/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Ministral 3 8B, edge-optimized multimodal model with 256k context, vision, and structured outputs.

Available on

Pricing from$0.15/ 1M input$0.15/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Large 3, open-weight mixture-of-experts model with 256k context and native multimodal vision.

Available on

Pricing from$0.50/ 1M input$1.50/ 1M output256Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Mistral Codestral 25.08: code-optimized model for fill-in-the-middle completion and fast code generation, 128k context.

Available on

Pricing from$0.30/ 1M input$0.90/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Latest routes to the current Voxtral Mini, a 32K-context speech model for voice agents with native audio understanding.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime 2602 delivers streaming live transcription with 13-language support, speaker diarization, and a 4B edge-deployable model.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime Latest routes to the current streaming transcription model, a 4B 13-language model for voice agents and live apps.

Available on

Pricing from$0.04/ 1M input$0.04/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Mini Transcribe Realtime by Mistral: streaming speech-to-text with configurable delay for live transcription, voice agents, and real-time subtitling.

Available on

Pricing$0.006 / min

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Small 2507 by Mistral: 24B audio-native LLM with function calling and structured outputs for speech understanding and agentic chat.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Mistral

Voxtral Small (latest) by Mistral: open-weight 24B audio-input instruct model that tracks the newest snapshot, with function calling and structured outputs.

Available on

Pricing from$0.10/ 1M input$0.40/ 1M output33Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 14B, Alibaba's April 2025 dense model with switchable thinking and direct modes for reasoning and tool use at 40K context.

Available on

Pricing from$0.10/ 1M input$0.22/ 1M output41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 30B A3B by Alibaba, an efficient MoE model with 3B active parameters and switchable thinking and non-thinking modes.

Available on

Pricing from$0.10/ 1M input$0.28/ 1M output$0.05/ 1M cache read41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 32B by Alibaba, a dense 32B model with switchable thinking and non-thinking modes for reasoning and speed.

Available on

Pricing from$0.10/ 1M input$0.28/ 1M output$0.05/ 1M cache read41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Alibaba

Qwen3 8B dense model on Fireworks

Available on

Pricing from$0.45/ 1M input$0.45/ 1M output$0.22/ 1M cache read41Kcontext

Privacy

ZDR
Training
No
Region
EU
by Talkie LM

Talkie 1930 is a 13B open-weight model trained only on pre-1931 English text, giving a 1930 knowledge cutoff for historical reasoning and contamination-free research.

Available on

Pricing from—/ 1M input—/ 1M output8Kcontext

Privacy

ZDR
Training
No
Region
EU
by Swiss AI Initiative

Apertus 70B, the Swiss AI Initiative's fully open model from EPFL and ETH Zurich, trained on 15T tokens with over 1,800 natively supported languages.

Available on

Pricing from$0.45/ 1M input$2.35/ 1M output66Kcontext

Privacy

ZDR
Training
No
Region
EU
by Regolo

Brick Complexity Pro, Regolo's hosted prompt-complexity classifier that grades queries easy, medium, or hard so routing picks the right model tier.

Available on

Pricing from$0.11/ 1M input$0.45/ 1M output

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-l-raw GreenPT Backed by Mistral Small 3.2 24B Direct access to the same GreenPT-backed model as green-l without the built-in system prompt. Input €0.25 Output €0.80 Context 128k Max output 32k Released Jun 2025 Text Images Documents Multilingual No system prompt

Available on

Pricing from$0.28/ 1M input$0.90/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-r-raw GreenPT Backed by GPT-OSS Direct access to the same reasoning stack as green-r without the GreenPT system prompt. Input €0.35 Output €0.95 Context 128k Max output 32k Released Aug 2025 Text Images Documents Multilingual No system prompt

Available on

Pricing from$0.39/ 1M input$1.06/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-l GreenPT Backed by Mistral Small 3.2 24B GreenPT-branded chat model tuned for multilingual writing, image understanding, and Dutch grammar guardrails. Input €0.25 Output €0.80 Context 128k Max output 32k Released Jun 2025 Text Images Documents Multilingual Writing assistant Dutch grammar guardrails

Available on

Pricing from$0.28/ 1M input$0.90/ 1M output128Kcontext

Privacy

ZDR
Training
No
Region
EU
by GreenPT

green-r GreenPT Backed by GPT-OSS GreenPT-branded reasoning model for advanced analysis, writing, and content generation. Input €0.35 Output €0.95 Context 128k Max output 32k Released Aug 2025 Text Images Documents Multilingual Advanced reasoning Writing & content generation

Available on

Pricing from$0.39/ 1M input$1.06/ 1M output123Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Sao10K's L3.3 70B Euryale v2.3, a 70B Llama 3.3 creative roleplay model and the direct successor to v2.2.

Available on

Pricing from$0.65/ 1M input$0.75/ 1M output131Kcontext

Privacy

ZDR
Training
No
Region
EU
by Barcelona Supercomputing Center

ALIA 40B Instruct 2601, Spain's publicly funded 40B multilingual model from the Barcelona Supercomputing Center, Apache 2.0 and pretrained across 35 European languages.

Available on

Pricing from$0.30/ 1M input$0.60/ 1M output66Kcontext

Privacy

ZDR
Training
No
Region
EU
by Cloudflare

Clef decision model by Cloudflare (27B), hosted by Opper in the EU. Answers noul (yes/no), choice and score questions with a probability for every option instead of text. Text or JSON state up to 16,384 tokens including the questions. Use POST /v3/compat/v1/systemone with model opper/clef. Weights: https://huggingface.co/Cloudflare/clef.

Available on

Pricing from$0.24/ 1M input$0/ 1M output16Kcontext

Privacy

ZDR
Training
No
Region
EU
by Cloudflare

Clef-flash decision model by Cloudflare (9B), hosted by Opper in the EU. Answers noul (yes/no), choice and score questions with a probability for every option instead of text. Text or JSON state up to 16,384 tokens including the questions. Use POST /v3/compat/v1/systemone with model opper/clef-flash. Weights: https://huggingface.co/Cloudflare/clef-flash.

Available on

Pricing from$0.04/ 1M input$0/ 1M output16Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Kev-4B System One decision model, a fine-tune by Jared Palmer of Alibaba's Qwen3.5-4B-Base, for typed yes/no decisions (noul), classification (choice) and rubric scoring (score), returning a probability for every option instead of text. Text or structured text input only, up to 8,192 tokens for the state plus one question; trained on states of up to 384 tokens, so accuracy on long documents is lower. Use POST /v3/compat/v1/systemone with model opper/kev-4b. Weights by Jared Palmer: https://huggingface.co/jaredpalmer/kev-4b. Base model: Alibaba's Qwen3.5-4B-Base, https://huggingface.co/Qwen/Qwen3.5-4B-Base.

Available on

Pricing from$0.04/ 1M input$0/ 1M output8Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Undi95's ReMM SLERP L2 13B, a Llama 2 SLERP merge recreating the MythoMax recipe with updated component models.

Available on

Pricing from$0.45/ 1M input$0.65/ 1M output6Kcontext

Privacy

ZDR
Training
No
Region
EU
by Community

Laya System One decision model by ConvAI Innovations (Nandha Kishor M), a non-autoregressive ModernBERT-large encoder (421M) for typed yes/no decisions (noul), classification (choice) and rubric scoring (score), returning a probability for every option instead of text. Text or structured text input only; the English checkpoint reads 512 tokens and truncates longer input. Use POST /v3/compat/v1/systemone with model berget/convaiinnovations/laya. Weights: https://huggingface.co/convaiinnovations/laya (Apache-2.0).

Available on

Pricing from$0.05/ 1M input$0/ 1M output512context

Privacy

ZDR
Training
No
Region
EU
by KBLab

KBLab's KB Whisper Large is a 2-billion-parameter Swedish speech-to-text model trained on 50,000 hours of audio, cutting word error rate by an average 47% versus Whisper-large-v3.

Available on

Pricing$0.0022 / min

Privacy

ZDR
Training
No
Region
EU
by Klang

Pianissimo from Klang — speech to text model on the Opper gateway.

Available on

Pricing$0.0039 / min

Privacy

ZDR
Training
No
Region
EU
by Evroc

roc, Evroc's enterprise AI agent for writing, data analysis, coding support, and knowledge retrieval over company documents and systems.

Available on

Pricing from$2.80/ 1M input$11.20/ 1M output

Privacy

ZDR
Training
No
Region
EU

Sovereign end to end

The gateway on a European cloud too

The providers on this list are European companies. For the gateway as well, Opper runs as its own instance on evroc, the Swedish cloud, set up with your team.

Sovereign AI on Opper

Frequently asked questions

Which AI models are sovereign?

+
90 models on the Opper gateway run on routes operated by European-owned providers and hosted in Europe as of 2026-10-06, on Inceptron, Melious, Sference, GreenPT, Mistral, TensorX, Berget, Greenference, Scaleway, Regolo, evroc, Infercom, Nextbit, Opper and SLNG. Families include OpenAI's open weights, Mistral, DeepSeek, Qwen, GLM, Kimi, Llama and others. This page lists every one of them with where it runs, what the provider keeps and what it costs.

What makes an AI model sovereign?

+
The route, not the model. A route counts as sovereign here when the company that operates it is headquartered in Europe and inference runs in Europe, so only European companies under European law handle your prompt. Open-weight models from any lab qualify when a European-owned provider runs them, and the lab never receives your data.

Can I run Claude, Gemini and Grok on a European-owned provider?

+
Not on Opper today. Claude, Gemini and Grok run in the EU on AWS Bedrock, Google Cloud, Azure and xAI, which keeps processing in the EU with operators headquartered outside Europe. For EU residency, which is what most teams need, they are on the EU list. Every EU-hosted model

How do I make sure calls only reach European-owned providers?

+
Call a route from this list by its model id. On Control Plane, set one Model access rule with provider country and inference location in Europe and it becomes the default for your whole organization: a call to anything else fails with a clear error, and Opper never switches provider silently. Model access rules

Is the gateway European-owned too?

+
On Opper on evroc, yes. Opper runs as its own instance on evroc, the Swedish cloud, so the gateway and the model provider are both European companies. The main Opper gateway runs in AWS Stockholm. Access to Opper on evroc is by onboarding. Sovereign AI on Opper

Read this list from an agent

This list as data. The same routes as JSON at opper.ai/models/eu.json (model id, provider, country, zero data retention, training and price per route) and as markdown at opper.ai/models/eu/sovereign.md.

The whole catalogue. The model catalogue is public JSON at https://api.opper.ai/v3/models, no API key needed. Filter it with query parameters, for example ?type=llm&limit=3, and page through it with limit and offset. Use it when you can fetch a URL but cannot connect to an MCP server. Coding agents that use MCP can connect to https://api.opper.ai/mcp. Two of the server's tools need no account and no sign-in: list_models searches the model catalogue by name, type, provider or capability, and get_guide returns short setup guides. Everything that touches your account needs you to approve the connection in the browser first. About the MCP server

Run sovereign AI on Opper

Opper on evroc is set up with your team. Get in touch and we onboard your organization.

Sovereign AI on Opper