Qwen3.8-Flash
Released 2026-08-26; listed at $0.16 / $0.47 on QwenCloud, callable on QCode
Alibaba's Qwen team released Qwen3.8-Flash on 2026-08-26 (the architecture-preview weights are often called Qwen3.8-Flash-Next). Hosted list price is $0.16 input / $0.47 output per 1M tokens, with 1M context advertised. This tracker follows the official API, weights and Bailian RMB pricing, and states where the id actually sits in QCode's catalog and usage table.
Updated 2026-09-19
Claude, GPT and Chinese models on one key, billed per token; check /models for callable models.
- Billed per token — live rates on /models
- Pay by card (Visa / Mastercard / AMEX), Apple Pay, Google Pay or crypto
- Self-service: top up and activate instantly
Four facts you can quote
Public QwenCloud USD list
Launch write-ups and pricing pages put the production SKU at $0.16 in / $0.47 out per million. That is Alibaba’s list, not a reseller invoice.
Announce day
@Alibaba_Qwen that day; Reuters / Bloomberg same day. Weights on Hugging Face / ModelScope. Whether the production API is live for everyone is inconsistent across write-ups — this page treats rollout as unfinished.
Pitched context
Reports: native about 262,144 tokens, YaRN to 1M. Trust the model card. Do not treat a recap window as a number you already measured on an aggregator.
Keep it apart from Qwen3.7
qwen3.8-flash and qwen3.8-max are both listed on QCode, with 30-day traffic of 4,776 and 1,970 calls respectively (as of 2026-09-19). qwen3.7-max was delisted on 2026-09-09; the Qwen line today is qwen3.8-max, qwen3.8-flash and qwen3.7-plus.
Where Flash sits in the 3.8 family
Qwen3.8-Max shipped around 2026-08-03 as the flagship (public $2 / $6). Flash is the cheap tier of the same generation; vendor copy frames it as near-flagship at roughly one-twelfth the price. Preview weights (Flash-Next) and the hosted production SKU (Qwen3.8-Flash) share a name family — do not treat them as one checkpoint. Downloadable weights ≠ your account is billed on a hosted id.
Do not merge RMB and USD into one row
Alibaba Cloud Bailian notice 2026-08-26 23:18: from 2026-08-27 12:00 Beijing time, Qwen3.8-Flash input ¥1.00 → ¥0.80, output ¥3.00 → ¥2.70 per million tokens. International pages quote $0.16 / $0.47. Keep unit and channel in the same sentence. Bailian RMB is not QwenCloud USD, and neither is a QCode price.
Timeline
Around 2026-08-03 Qwen3.8-Max shipped as the flagship at roughly $2 / $6 per 1M tokens with 1M context. That page also said the id was not yet in our catalog — both ids are in it now.
2026-08-26 Qwen3.8-Flash / Flash-Next announced. USD list $0.16 / $0.47. Weights and tech report on GitHub / HF / ModelScope.
Bailian cuts RMB unit prices to ¥0.80 / ¥2.70. As of 2026-08-30 this page still splits QwenCloud USD and Bailian RMB.
Confirmed vs rumor
Confirmed
Announce day 2026-08-26, QwenCloud USD $0.16 / $0.47, Bailian RMB cut on 08-27, Max vs Flash price story, public weights — primary or official notices. No qwen3.8* id in this site’s usage table on 2026-08-30, also checked.
Rumor / misread
"Flash must fit a laptop" — parameter and active-count claims vary by repost; trust the model card, and don't read 'Flash' as 'small'. "Already on every aggregator" — reports contradict each other, some say API coming soon, some say Bailian already bills; this page does not pick a side. "QCode's catalog still lacks it" — outdated: re-checked 2026-09-18, the catalog and usage table both carry qwen3.8-flash.
Tongyi you can call today vs 3.8-Flash still in tracking
Need it today: both 3.8 ids are in the catalog
qwen3.8-max / qwen3.8-flash / qwen3.7-plus all show up in the usage table and are callable today (qwen3.7-max was retired on 2026-09-09). Price and window follow the live /models list.
3.8-Flash: follow official channels; do not hard-code yet
Official API, Bailian and the HF weights are three different landing paths. Before hard-coding qwen3.8-flash, confirm your own console returns that id. Our catalog listed qwen3.8-flash and qwen3.8-max as of 2026-09-18.
How to track (official channels)
Read @Alibaba_Qwen, the qwen.ai blog, Bailian notices and the model card rather than aggregator reposts. Keep $ for USD and ¥ for RMB in the same sentence. If you need Qwen today, qwen3.8-flash, qwen3.8-max and qwen3.7-plus are all in our catalog.
This site does not describe Qwen3.8-Flash as callable here
qwen3.8-flash is in QCode's /models list and usage table (4,776 calls over 30 days as of 2026-09-19), and qwen3.8-max is on sale too (1,970). qwen3.7-max was delisted on 2026-09-09; the Qwen line is now qwen3.8-max, qwen3.8-flash and qwen3.7-plus. Defer to what /models lists.
FAQ
What is the public list price for Qwen3.8-Flash?
International pages use $0.16 in / $0.47 out per million (QwenCloud). Bailian RMB after 2026-08-27 12:00 Beijing time is ¥0.80 / ¥2.70. Trust the console you actually pay.
How far from Qwen3.8-Max?
Max public list is about $2 / $6. Several write-ups frame Flash at roughly one-twelfth. That is vendor tier copy, not a score we measured.
Are Qwen3.8-Flash-Next and Flash the same id?
Next is usually the architecture-preview / weights name; Flash is the hosted production SKU. Do not paste an HF repo name into a random vendor’s model field.
Can I call qwen3.8-flash on QCode today?
Yes. Both qwen3.8-flash and qwen3.8-max are in the catalog (4,776 / 1,970 calls over 30 days as of 2026-09-19). qwen3.7-max was delisted on 2026-09-09.
Is 1M context on by default?
The pitch is 1M; native length is reported near 262K plus YaRN. Default-on and extra billing follow the official model note. This page does not invent a hidden flag.
Can I run it locally?
Weights are on HF / ModelScope. VRAM and active parameters follow the card. The word Flash does not guarantee a consumer GPU will run the full hosted spec comfortably.
Sources
@Alibaba_Qwen announcement of 2026-08-26; QwenCloud $0.16 / $0.47 as reported at launch; Alibaba Cloud Bailian notice 2026-08-26 23:18 (¥0.80 / ¥2.70 from 2026-08-27 12:00 Beijing). Max background is on our Qwen3.8-Max guide. Catalog and usage conclusions come from /models and the 30-day table on 2026-09-18.
Tongyi 3.8-Flash is still landing on official channels; we track it here
qwen3.8-flash and qwen3.8-max are both on QCode's list — same key, change the model. Billing is official price × service fee.
Related
Qwen3.8-Max tracker
Same-generation flagship at a public $2/$6, in our catalog alongside Flash.
Qwen3.8-27B
A downloadable 3.8 small tier — not the same id as the hosted Flash SKU.
Qwen3.7-Plus
One of the Tongyi tiers that actually matches today’s catalog.
Compiled from public Qwen / Alibaba Cloud notices and press retellings; not Alibaba's official position. Prices and availability follow the console you use. QCode only calls an id available when it appears in both the catalog and the usage table.