Anthropic's Opus model for complex agentic coding and enterprise work. 1M context, adaptive thinking.
Available on
Privacy
- ZDR
- Training
- No
- Region
- US, EU, Multi
AI Model Directory
Search by what matters. Zero data retention, residency, training and more.
Anthropic's Opus model for complex agentic coding and enterprise work. 1M context, adaptive thinking.
Claude Fable 5.1 is Anthropic's frontier model for agentic coding, knowledge work, and long-running problem solving, with a 1M token context and 128K output.
GPT-6 Astra is OpenAI's flagship model for computer use, software engineering, and scientific work, with a 1.05M token context and 128K output.
Claude Opus 5 is Anthropic's flagship for complex agentic coding and enterprise work, with a 1M token context window and adaptive thinking.
Claude Fable 5 by Anthropic, state-of-the-art generalist model with advanced vision and 1M token context for complex software engineering and reasoning.
Muse Spark 1.3 is the latest of Meta's Muse Spark models, cutting tool calls and tokens versus Muse Spark 1.2 across a 1M token context with image and video input.
OpenAI GPT-6.1 Sol for coding, computer use, and complex professional work.
OpenAI GPT-6 Sol for complex professional work, coding, and agent workflows.
GPT-5.6 Sol tops OpenAI's GPT-5.6 family, the frontier tier for complex coding, deep research, and long-running agents with a 1M token context.
xAI's Grok 4.7 frontier model for coding, agentic tasks, and knowledge work, with reasoning, 500K context window, vision, and agentic tool use
MiMo V2.6 Pro from Xiaomi — text model on the Opper gateway.
Qwen3.8-Max, Alibaba's 2.4T-parameter MoE flagship with hybrid thinking, vision input, and a context window approaching one million tokens.
GLM-5.3 by Z.ai is a 744B open-weight coding and agentic model, post-trained on the GLM-5.2 base with a 1M-token context and low, high, and max reasoning effort levels.
Kimi K3 by Moonshot: open-weight 2.8 trillion parameter MoE with 104 billion active, native vision, and a 1M token context for long-horizon coding and agentic work.
GPT-5.6 Terra is the balanced middle tier of OpenAI's GPT-5.6 family, pairing near-flagship quality with everyday speed and a 1M token context.
Claude Opus 4.8 by Anthropic, highly capable model for complex reasoning, coding, and agentic work with 1M context and improved efficiency.
GLM-5.3-Flash by Z.ai is a natively multimodal 320B MoE with 18B active parameters, hybrid sparse and linear attention, a 1M-token context, and MIT open weights for coding and agent work.
Gemini 3.8 Flash is Google's workhorse multimodal model for software engineering and agentic tasks, pairing 1M context with image, audio, video, and PDF input.
Claude Opus 4.7 by Anthropic, advanced vision and coding model with 1M context for complex long-running agentic workflows.
Qwen3.8-2.4T-A95B is the open-weight release of Alibaba's Qwen3.8-Max flagship, a 2.4 trillion parameter Mixture-of-Experts activating 95B per token.
Qwen3.8 Flash-Next by Alibaba is a multimodal MoE with a 125B backbone and 6B active parameters, an early look at the Qwen4 architecture with 262K native context for coding and agents.
Gemini 3.7 Flash is Google's August 2026 Flash release, a large step up in coding and agentic automation over 3.6 while keeping the 1M-token multimodal context.
Muse Spark 1.2, Meta's flagship coding and multimodal model with a million-token context, built for large codebases and agentic tool use.
DeepSeek V4.1 Flash is a 552B multimodal MoE with a causal encoder-decoder design, 8B to 16B active parameters, a 1M-token context, and MIT weights for fast agentic coding.
Claude Sonnet 5.5 via the Anthropic API
Claude Sonnet 5 is Anthropic's balanced production model, bringing near-Opus agentic coding and tool use to a 1M token context window.
GPT-5.6 Luna, the fast high-volume tier of OpenAI's GPT-5.6 family, handles classification, summarization, and bulk pipelines with 1M context.
OpenAI GPT-6 Luna for cost-sensitive, high-volume work and coding agents.
DeepSeek V4 Pro, flagship 1.6T MoE with 1M context and hybrid thinking for frontier reasoning and agents.
DeepSeek V4 Pro 0813 is the general availability build of DeepSeek's 1M-context flagship, tuned for agentic tool use and terminal-driven work.
DeepSeek V4 Flash 0731, the July 2026 retrained checkpoint with major agentic and coding gains over the original release.
Also available as deepseek-v4-flash-latest.
DeepSeek V4 Flash, lightweight 284B MoE with 1M context and hybrid thinking for cost-efficient agents.
Gemini 3.6 Flash, Google's token-efficient multimodal Flash model, cuts output tokens by 17% while improving coding and agentic planning.
Qwen3.8 27B by Alibaba is a dense 27B open-weight vision-language model with image and video input, a 262K native context, thinking on by default, and Apache 2.0 weights.
GLM-5.2 by Z.ai is a 744B open-weight (MIT) mixture-of-experts coding model with a 1M-token context, IndexShare sparse attention, and selectable thinking-effort levels.
Gemini 3.5 Flash by Google pairs frontier-level intelligence with fast throughput and a 1M token context for agentic and coding tasks.
GPT-5.3 Codex is OpenAI's most capable agentic coding model, reaching state-of-the-art on SWE-Bench Pro and Terminal-Bench while using fewer tokens than prior models.
Claude Opus 4.6 by Anthropic, professional knowledge-work model with 1M context, extended thinking, and 128K token outputs.
Claude Sonnet 4.6 by Anthropic, full upgrade across coding, computer use, and long-context reasoning.
Gemini 3.1 Pro Preview by Google, a frontier reasoning LLM with 1M context and strong agentic coding, scoring 77% on ARC-AGI-2 and 81% on SWE-Bench Verified.
Qwen3.7-Max by Alibaba, a 1M-context proprietary reasoning agent with extended thinking for multi-step code and autonomous workflows.
MiniMax M3: natively multimodal 428B Mixture-of-Experts model with ~23B active parameters, MiniMax Sparse Attention for long context, and 59.0% on SWE-Bench Pro
Claude Opus 4.5 by Anthropic, highly capable model for coding and agents with extended thinking and a tunable effort parameter.
The Opper model directory lists every large language model on the gateway, with image, video, and voice models alongside them. Filter by price, context window, capability, benchmark score, hosting region, and zero data retention, then run any of them through one OpenAI-compatible API and a single key. The default order leads with the models Artificial Analysis ranks highest on its intelligence index, which you can see in full on the LLM leaderboard. Put two models head to head on the comparison pages, or read how routing and fallbacks work on the LLM gateway.