Models / Qwen3.8-Flash
Qwen3.8-Flash
Qwen3.8 family · Provider: alibaba-cloud · Status: active · Released: 2026-08
Qwen3.8-Flash is a Qwen3.8 model developed by alibaba-cloud, a Chinese AI company, released 2026-08 with a 1,048,576-token context window and closed weights.
| Key fact | Value |
|---|---|
| Model ID | qwen3.8-flash |
| Architecture | Not publicly disclosed |
| Context window | 1,048,576 tokens |
| Max output | 131,072 tokens |
| Open weights | No |
| License | proprietary |
| Self-hosting | No |
| API available | Yes |
| API pricing | $0.15 input / $0.47 output per 1M tokens (USD) · provider pricing page |
| Regions | china-beijing, singapore, hong-kong, germany-frankfurt, us-virginia, japan-tokyo |
| Cloud providers | Alibaba Cloud |
Capabilities
| Capability | Supported |
|---|---|
| Reasoning | Yes |
| Coding | Unknown |
| Math | Unknown |
| Chinese | Unknown |
| English | Unknown |
| Multilingual | Unknown |
| Vision | Yes |
| Audio | Unknown |
| Video | Yes |
| Tool calling | Unknown |
| Function calling | Unknown |
| Structured output | Unknown |
| Agent capability | Unknown |
| RAG | Unknown |
| Computer use | Unknown |
Known limitations
- Batch inference not supported
- Prices differ by region: Singapore $0.15/$0.47; Beijing and Global regions $0.113/$0.382 per 1M tokens
- Architecture details not published on the fetched official pages
Qwen3.8-Flash is the lightweight, low-cost model of the Qwen3.8 family: 1M-token context (991,808 max input, 131,072 max output), multimodal input (image, text, video) with text output, and context caching. Alibaba states it is fully compatible with both OpenAI and Anthropic API protocols.
Pricing: Singapore $0.15 input / $0.47 output per 1M tokens; Beijing and other Global regions $0.113 / $0.382 (as of 2026-09-20). Batch inference is not supported.
Provider
Pricing
Agents built on this model
Related models (same family)
API
Model family timeline
| Model | Released | Status |
|---|---|---|
| Qwen3.8-Max | 2026-08 | active |
| Qwen3.8-Flash (this page) | 2026-08 | active |
| Qwen3.8-2.4T-A95B | 2026-08-12 | active |
Data interpretation
| Field | Value | Evidence type |
|---|---|---|
| Context window | 1,048,576 tokens | Official |
| Open weights | No | Official |
| License | proprietary | Official |
| API pricing | $0.15 / $0.47 per 1M tokens (USD) | Official |
| Family position | 3 models in the Qwen3.8 family | China AI Hub analysis |
Evidence types: Official = vendor documentation, pricing pages or model cards. Vendor-reported = benchmark scores published by the vendor. China AI Hub analysis = derived from the database itself. See the sourcing policy.
Sources
Confidence and source hierarchy per the sourcing policy. Facts change; check the source before relying on this page.
What is the context window of Qwen3.8-Flash?
Qwen3.8-Flash has a 1,048,576-token context window and a maximum output of 131,072 tokens.
Is Qwen3.8-Flash open weight?
No — Qwen3.8-Flash is not open weight.
How much does Qwen3.8-Flash cost through the API?
$0.15 per 1M input tokens and $0.47 per 1M output tokens (USD).
Where does China AI Hub get its Qwen3.8-Flash data?
From 2 sources (official pages first), last verified 2026-09-20. Benchmark scores are labeled by source type; see the sourcing policy for details.