Qwen: Qwen3.8 2.4T A95B · Alibaba
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Alibaba with 95 billion active parameters out of 2.4 trillion total. It delivers advanced natural language processing capabilities with extended 1 million token context window.
About Qwen: Qwen3.8 2.4T A95B
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Alibaba with 95 billion active parameters out of 2.4 trillion total. It delivers advanced natural language processing capabilities with extended 1 million token context window.
Where it shines: 95B active parameters · 1M token context · Sparse MoE architecture.
How to use Qwen: Qwen3.8 2.4T A95B
-
1
Type or upload
Type what you want in the box above — or upload the file if the tool asks for one.
-
2
Generate
Click the main button. Wait 2-30 seconds depending on the model and input size.
-
3
Download or share
Download the result or share the direct link. No watermark, ready to use.
Frequently asked questions
How much does it cost to use Qwen: Qwen3.8 2.4T A95B?
Qwen: Qwen3.8 2.4T A95B is a Pro model: it costs 34 tokens per use (~$0.17 real cost for us). You need a Pro plan ($9/month → 15,000 tokens) or a one-shot pack. If you already have tokens in the free account, you can also spend them directly.
How many uses of Qwen: Qwen3.8 2.4T A95B are included in the Pro plan?
Pro ($9/month) gives you 15,000 recurring tokens. At 34 tokens per Qwen use: Qwen3.8 2.4T A95B, that's ~441 full uses per cycle. If you run out, one-shot packs (5,000 / 25,000 / 80,000 tokens) add to the balance without expiring before one year.
What makes Qwen: Qwen3.8 2.4T A95B special?
Alibaba leads Asian open models with efficient MoE architectures and huge context windows (262K+), with specific strengths of 95b active parameters, context 1m tokens and dispersed moe architecture.
How fast does Qwen: Qwen3.8 2.4T A95B respond?
Qwen: Qwen3.8 2.4T A95B has a balanced speed: 5-15 seconds per response — neither the fastest nor the slowest in the catalog.The actual time also depends on the length of the prompt and the load of the datacenter — models with huge context take longer when you enter very long texts.
How to use Qwen: Qwen3.8 2.4T A95B in ia.gratis?
You can use Qwen: Qwen3.8 2.4T A95B from /chat/ by selecting Qwen: Qwen3.8 2.4T A95B in the picker, or via the REST API with `model=qwen3-8-2-4t-a95b` in the POST body. Quick summary: qwen3.8 2.4T A95B: expert model of dispersed mix with 95B parameters active. The internal model identifier is `qwen3-8-2-4t-a95b` — useful when integrating by API.