Status tracker · Qwen3.8-Flash

Qwen3.8-Flash
Released 2026-08-26; listed at $0.16 / $0.47 on QwenCloud, callable on QCode

Alibaba's Qwen team released Qwen3.8-Flash on 2026-08-26 (the architecture-preview weights are often called Qwen3.8-Flash-Next). Hosted list price is $0.16 input / $0.47 output per 1M tokens, with 1M context advertised. This tracker follows the official API, weights and Bailian RMB pricing, and states where the id actually sits in QCode's catalog and usage table.

Updated 2026-09-19

Claude, GPT and Chinese models on one key, billed per token; check /models for callable models.

#Qwen3.8-Flash#$0.16 / $0.47#2026-08-26#Callable on QCode

Four facts you can quote

$0.16 / $0.47

Public QwenCloud USD list

Launch write-ups and pricing pages put the production SKU at $0.16 in / $0.47 out per million. That is Alibaba’s list, not a reseller invoice.

2026-08-26

Announce day

@Alibaba_Qwen that day; Reuters / Bloomberg same day. Weights on Hugging Face / ModelScope. Whether the production API is live for everyone is inconsistent across write-ups — this page treats rollout as unfinished.

1M

Pitched context

Reports: native about 262,144 tokens, YaRN to 1M. Trust the model card. Do not treat a recap window as a number you already measured on an aggregator.

In our catalog

Keep it apart from Qwen3.7

qwen3.8-flash and qwen3.8-max are both listed on QCode, with 30-day traffic of 4,776 and 1,970 calls respectively (as of 2026-09-19). qwen3.7-max was delisted on 2026-09-09; the Qwen line today is qwen3.8-max, qwen3.8-flash and qwen3.7-plus.

Where Flash sits in the 3.8 family

Qwen3.8-Max shipped around 2026-08-03 as the flagship (public $2 / $6). Flash is the cheap tier of the same generation; vendor copy frames it as near-flagship at roughly one-twelfth the price. Preview weights (Flash-Next) and the hosted production SKU (Qwen3.8-Flash) share a name family — do not treat them as one checkpoint. Downloadable weights ≠ your account is billed on a hosted id.

Do not merge RMB and USD into one row

Alibaba Cloud Bailian notice 2026-08-26 23:18: from 2026-08-27 12:00 Beijing time, Qwen3.8-Flash input ¥1.00 → ¥0.80, output ¥3.00 → ¥2.70 per million tokens. International pages quote $0.16 / $0.47. Keep unit and channel in the same sentence. Bailian RMB is not QwenCloud USD, and neither is a QCode price.

Timeline

Around 2026-08-03

Around 2026-08-03 Qwen3.8-Max shipped as the flagship at roughly $2 / $6 per 1M tokens with 1M context. That page also said the id was not yet in our catalog — both ids are in it now.

2026-08-26

2026-08-26 Qwen3.8-Flash / Flash-Next announced. USD list $0.16 / $0.47. Weights and tech report on GitHub / HF / ModelScope.

2026-08-27 12:00 Beijing time

Bailian cuts RMB unit prices to ¥0.80 / ¥2.70. As of 2026-08-30 this page still splits QwenCloud USD and Bailian RMB.

Confirmed vs rumor

Confirmed

Announce day 2026-08-26, QwenCloud USD $0.16 / $0.47, Bailian RMB cut on 08-27, Max vs Flash price story, public weights — primary or official notices. No qwen3.8* id in this site’s usage table on 2026-08-30, also checked.

Rumor / misread

"Flash must fit a laptop" — parameter and active-count claims vary by repost; trust the model card, and don't read 'Flash' as 'small'. "Already on every aggregator" — reports contradict each other, some say API coming soon, some say Bailian already bills; this page does not pick a side. "QCode's catalog still lacks it" — outdated: re-checked 2026-09-18, the catalog and usage table both carry qwen3.8-flash.

Tongyi you can call today vs 3.8-Flash still in tracking

Need it today: both 3.8 ids are in the catalog

qwen3.8-max / qwen3.8-flash / qwen3.7-plus all show up in the usage table and are callable today (qwen3.7-max was retired on 2026-09-09). Price and window follow the live /models list.

3.8-Flash: follow official channels; do not hard-code yet

Official API, Bailian and the HF weights are three different landing paths. Before hard-coding qwen3.8-flash, confirm your own console returns that id. Our catalog listed qwen3.8-flash and qwen3.8-max as of 2026-09-18.

How to track (official channels)

Read @Alibaba_Qwen, the qwen.ai blog, Bailian notices and the model card rather than aggregator reposts. Keep $ for USD and ¥ for RMB in the same sentence. If you need Qwen today, qwen3.8-flash, qwen3.8-max and qwen3.7-plus are all in our catalog.

This site does not describe Qwen3.8-Flash as callable here

qwen3.8-flash is in QCode's /models list and usage table (4,776 calls over 30 days as of 2026-09-19), and qwen3.8-max is on sale too (1,970). qwen3.7-max was delisted on 2026-09-09; the Qwen line is now qwen3.8-max, qwen3.8-flash and qwen3.7-plus. Defer to what /models lists.

FAQ

What is the public list price for Qwen3.8-Flash?

International pages use $0.16 in / $0.47 out per million (QwenCloud). Bailian RMB after 2026-08-27 12:00 Beijing time is ¥0.80 / ¥2.70. Trust the console you actually pay.

How far from Qwen3.8-Max?

Max public list is about $2 / $6. Several write-ups frame Flash at roughly one-twelfth. That is vendor tier copy, not a score we measured.

Are Qwen3.8-Flash-Next and Flash the same id?

Next is usually the architecture-preview / weights name; Flash is the hosted production SKU. Do not paste an HF repo name into a random vendor’s model field.

Can I call qwen3.8-flash on QCode today?

Yes. Both qwen3.8-flash and qwen3.8-max are in the catalog (4,776 / 1,970 calls over 30 days as of 2026-09-19). qwen3.7-max was delisted on 2026-09-09.

Is 1M context on by default?

The pitch is 1M; native length is reported near 262K plus YaRN. Default-on and extra billing follow the official model note. This page does not invent a hidden flag.

Can I run it locally?

Weights are on HF / ModelScope. VRAM and active parameters follow the card. The word Flash does not guarantee a consumer GPU will run the full hosted spec comfortably.

Sources

@Alibaba_Qwen announcement of 2026-08-26; QwenCloud $0.16 / $0.47 as reported at launch; Alibaba Cloud Bailian notice 2026-08-26 23:18 (¥0.80 / ¥2.70 from 2026-08-27 12:00 Beijing). Max background is on our Qwen3.8-Max guide. Catalog and usage conclusions come from /models and the 30-day table on 2026-09-18.

Tongyi 3.8-Flash is still landing on official channels; we track it here

qwen3.8-flash and qwen3.8-max are both on QCode's list — same key, change the model. Billing is official price × service fee.

Related

Compiled from public Qwen / Alibaba Cloud notices and press retellings; not Alibaba's official position. Prices and availability follow the console you use. QCode only calls an id available when it appears in both the catalog and the usage table.

Try first, then decide

Not sure which tier? Start with Starter ($8.57/mo) and upgrade when you're happy — the unused value of the old plan goes back to your balance.