China AI Hub China AI Hub

Models / Qwen3.8-Max

Qwen3.8-Max

Qwen3.8 family · Provider: alibaba-cloud · Status: active · Released: 2026-08

Qwen3.8-Max is a Qwen3.8 model developed by alibaba-cloud, a Chinese AI company, released 2026-08 with a 1,048,576-token context window and closed weights.

Qwen3.8-Max
Image: AI-generated illustration (Seedream)
Last verified: · Data status: Current · Next review:
Key facts for Qwen3.8-Max
Key factValue
Model IDqwen3.8-max
Version0902
Architecture2.4T-parameter MoE, 95B activated, 512 experts (10 routed + 1 shared per token), 92 layers, Gated DeltaNet + Gated Attention hybrid
Context window 1,048,576 tokens
Max output131,072 tokens
Open weightsNo
Licenseproprietary
Self-hostingNo
API availableYes
API pricing $2 input / $6 output per 1M tokens (USD) · provider pricing page
Regionschina-beijing, singapore, hong-kong, germany-frankfurt, us-virginia, japan-tokyo
Cloud providersAlibaba Cloud

Capabilities

Capabilities of Qwen3.8-Max
CapabilitySupported
ReasoningYes
CodingYes
MathUnknown
ChineseUnknown
EnglishUnknown
MultilingualUnknown
VisionYes
AudioUnknown
VideoYes
Tool callingYes
Function callingUnknown
Structured outputYes
Agent capabilityUnknown
RAGUnknown
Computer useUnknown

Benchmark results

Benchmark results for Qwen3.8-Max
BenchmarkVersionScoreMetricDateSource typeSource
Terminal-Bench 2.1 86.6 accuracy 2026-08 vendor_reported link
SWE-bench Pro 67.7 accuracy 2026-08 vendor_reported link
GPQA Diamond 92.6 accuracy 2026-08 vendor_reported link
HLE 43.6 (56.2 with tools) accuracy 2026-08 vendor_reported link
MRCR v2 256K 92.9 accuracy 2026-08 vendor_reported link

Benchmark scores are single data points, not universal rankings. Vendor-reported scores are labeled as such.

Known limitations

Qwen3.8-Max is Alibaba’s flagship API model: a 2.4T-parameter MoE (95B activated) with a 1M-token context window (991,808 max input), 131,072 max output and 262,144 max chain-of-thought. Input modalities are image, text and video; output is text. It supports thinking and non-thinking modes, function calling, structured outputs, web search, prefix completion, context caching and batch inference (Beijing only).

The 0902 snapshot (2026-09-02) upgraded coding and vision. API pricing: Singapore $2 input / $6 output per 1M tokens; Beijing and other Global regions $1.65 / $4.951 (as of 2026-09-20). Fine-tuning is not supported.

Provider

Pricing

Benchmarks with results for this model

Agents built on this model

Related models (same family)

API

Model family timeline

Release timeline for the Qwen3.8 family
ModelReleasedStatus
Qwen3.8-Max (this page) 2026-08 active
Qwen3.8-Flash 2026-08 active
Qwen3.8-2.4T-A95B 2026-08-12 active

Data interpretation

Evidence types for Qwen3.8-Max
FieldValueEvidence type
Context window1,048,576 tokensOfficial
Architecture2.4T-parameter MoE, 95B activated, 512 experts (10 routed + 1 shared per token), 92 layers, Gated DeltaNet + Gated Attention hybridOfficial
Open weightsNoOfficial
LicenseproprietaryOfficial
API pricing $2 / $6 per 1M tokens (USD) Official
Terminal-Bench 2.1 86.6 accuracy Vendor-reported
SWE-bench Pro 67.7 accuracy Vendor-reported
GPQA Diamond 92.6 accuracy Vendor-reported
HLE 43.6 (56.2 with tools) accuracy Vendor-reported
MRCR v2 256K 92.9 accuracy Vendor-reported
Family position3 models in the Qwen3.8 familyChina AI Hub analysis

Evidence types: Official = vendor documentation, pricing pages or model cards. Vendor-reported = benchmark scores published by the vendor. China AI Hub analysis = derived from the database itself. See the sourcing policy.

Sources

Confidence and source hierarchy per the sourcing policy. Facts change; check the source before relying on this page.

What is the context window of Qwen3.8-Max?

Qwen3.8-Max has a 1,048,576-token context window and a maximum output of 131,072 tokens.

Is Qwen3.8-Max open weight?

No — Qwen3.8-Max is not open weight.

How much does Qwen3.8-Max cost through the API?

$2 per 1M input tokens and $6 per 1M output tokens (USD).

Where does China AI Hub get its Qwen3.8-Max data?

From 3 sources (official pages first), last verified 2026-09-20. Benchmark scores are labeled by source type; see the sourcing policy for details.