China AI Hub China AI Hub

Models / Qwen3.8-Flash

Qwen3.8-Flash

Qwen3.8 family · Provider: alibaba-cloud · Status: active · Released: 2026-08

Qwen3.8-Flash is a Qwen3.8 model developed by alibaba-cloud, a Chinese AI company, released 2026-08 with a 1,048,576-token context window and closed weights.

Qwen3.8-Flash
Image: AI-generated illustration (Seedream)
Last verified: · Data status: Current · Next review:
Key facts for Qwen3.8-Flash
Key factValue
Model IDqwen3.8-flash
ArchitectureNot publicly disclosed
Context window 1,048,576 tokens
Max output131,072 tokens
Open weightsNo
Licenseproprietary
Self-hostingNo
API availableYes
API pricing $0.15 input / $0.47 output per 1M tokens (USD) · provider pricing page
Regionschina-beijing, singapore, hong-kong, germany-frankfurt, us-virginia, japan-tokyo
Cloud providersAlibaba Cloud

Capabilities

Capabilities of Qwen3.8-Flash
CapabilitySupported
ReasoningYes
CodingUnknown
MathUnknown
ChineseUnknown
EnglishUnknown
MultilingualUnknown
VisionYes
AudioUnknown
VideoYes
Tool callingUnknown
Function callingUnknown
Structured outputUnknown
Agent capabilityUnknown
RAGUnknown
Computer useUnknown

Known limitations

Qwen3.8-Flash is the lightweight, low-cost model of the Qwen3.8 family: 1M-token context (991,808 max input, 131,072 max output), multimodal input (image, text, video) with text output, and context caching. Alibaba states it is fully compatible with both OpenAI and Anthropic API protocols.

Pricing: Singapore $0.15 input / $0.47 output per 1M tokens; Beijing and other Global regions $0.113 / $0.382 (as of 2026-09-20). Batch inference is not supported.

Provider

Pricing

Agents built on this model

Related models (same family)

API

Model family timeline

Release timeline for the Qwen3.8 family
ModelReleasedStatus
Qwen3.8-Max 2026-08 active
Qwen3.8-Flash (this page) 2026-08 active
Qwen3.8-2.4T-A95B 2026-08-12 active

Data interpretation

Evidence types for Qwen3.8-Flash
FieldValueEvidence type
Context window1,048,576 tokensOfficial
Open weightsNoOfficial
LicenseproprietaryOfficial
API pricing $0.15 / $0.47 per 1M tokens (USD) Official
Family position3 models in the Qwen3.8 familyChina AI Hub analysis

Evidence types: Official = vendor documentation, pricing pages or model cards. Vendor-reported = benchmark scores published by the vendor. China AI Hub analysis = derived from the database itself. See the sourcing policy.

Sources

Confidence and source hierarchy per the sourcing policy. Facts change; check the source before relying on this page.

What is the context window of Qwen3.8-Flash?

Qwen3.8-Flash has a 1,048,576-token context window and a maximum output of 131,072 tokens.

Is Qwen3.8-Flash open weight?

No — Qwen3.8-Flash is not open weight.

How much does Qwen3.8-Flash cost through the API?

$0.15 per 1M input tokens and $0.47 per 1M output tokens (USD).

Where does China AI Hub get its Qwen3.8-Flash data?

From 2 sources (official pages first), last verified 2026-09-20. Benchmark scores are labeled by source type; see the sourcing policy for details.