Qwen 3 Next 80B A3B Thinking

by Alibaba

Qwen 3 Next 80B A3B Thinking is an open-weight, sparse mixture-of-experts model that pairs the efficiency of roughly 3B active parameters with extended reasoning. Unlike the Instruct variant, it generates reasoning traces before its final answer, exposing its problem-solving steps across math, logic, coding, and planning tasks. With a 128K context window and a sparse MoE design (512 experts, 10 activated per token), it outperforms Gemini 2.5 Flash Thinking on several benchmarks, including AIME 2025 and Arena-Hard, while staying computationally efficient. Its strengths show clearly on competition math and demanding coding challenges. It is best suited to educational AI, math tutoring, code debugging systems, and applications where interpretable reasoning is important. The open-weight release allows fine-tuning for specialized reasoning domains and local deployment.

Key info

Input
Output
Features
Context window
128K
Max output
66K

Available routes

No routes currently available — Qwen 3 Next 80B A3B Thinking isn't routed through the Opper gateway right now. It may return.

Contact us about this model →

Available models from Alibaba

Start building with 700+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation