Qwen3.8 Flash-Next

by Alibaba

Released 12 days after Qwen3.8 27B · All Alibaba releases

Qwen3.8 Flash-Next is a natively multimodal MoE model from Alibaba with 125B total / 6B active parameters and a 256K context window. It targets frontier-grade coding and agentic work at very low active-parameter cost (Artificial Analysis intelligence index 55.8, coding index 73.1).

Key info

Input
Output
Features
Context window
262K
Max output
Input price
$0.20 /1M
Output price
$0.50 /1M
Released
  • EU residency available
  • Zero data retention on pay-as-you-go
  • No training by default
  • GDPR DPA available

Available routes

Qwen3.8 Flash-Next runs on 1 route through the Opper gateway. Compare residency, ZDR, and training posture at a glance — full data-handling detail per route below.

ProviderRegionZero data retentionTrainingInputOutput
EUZero data retentionNo$0.20$0.50

Data handling per route

Each route hosting Qwen3.8 Flash-Next has its own privacy posture, residency, and GDPR terms. Postures are maintained by Opper with a last-verification timestamp.

TensorX European Union🇪🇺

Zero data retention is on by default on Pay-as-you-go — no action required. No training on customer data. EU; DPA available.

Zero data retention
On by default on Pay-as-you-go.
Training
No training on customer data.
Logging
None
Third-party access
None disclosed
GDPR DPA
DPA available
Transfer mechanism
Not applicable — data stays in EU

Benchmarks

Independent benchmark scores — composite indices for reasoning, coding, and math, plus individual eval scores where available.

Global rank#22 of 611 LLMs
TierFrontier
Output speed70 tok/s
First token1.70s
Intelligence Index55.8
Coding Index73.1
Reasoning & knowledge
GPQA Diamond
92%
Humanity's Last Exam
38%
Long-context reasoning
77%
Coding
SciCode
47%

Get started

Call Qwen3.8 Flash-Next through the Opper gateway with one API key. Let your coding agent set it up, or call it directly — Opper is drop-in compatible with the OpenAI, Anthropic, and Google AI SDKs.

Set it up with your agent

Copy this and paste it into a coding agent like Claude Code, Cursor or Codex and it'll wire up Opper for you.

Or call it directly

import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.OPPER_API_KEY,
baseURL: "https://api.opper.ai/v3/compat",
});
const completion = await client.chat.completions.create({
model: "tensorx/qwen/qwen3.8-flash-next",
messages: [{ role: "user", content: "Hello" }],
});
console.log(completion.choices[0].message.content);

Compare Qwen3.8 Flash-Next with…

Side-by-side on privacy, EU hosting, pricing, and benchmarks.

Other models from Alibaba

Start building with 700+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation