Home / Alibaba / Qwen / Qwen 3.8 2.4T A95B
Qwen 3.8 2.4T A95B API price
qwen/qwen3.8-2.4t-a95b
The cheapest Qwen 3.8 2.4T A95B API is $2.00/M input at DeepInfra - all 8 providers compared below, per 1M tokens.
FlagshipAlibaba / Qwen1.0M context8 sellers
Input
$2.00
per 1M tokens · DeepInfra
up to $2.50
Output
$6.00
per 1M tokens · DeepInfra
up to $7.50
Where to buy it
Cheapest Qwen 3.8 2.4T A95B API providers compared - 8 sellers, from $2.00/M.
| Provider | Input $/M | Output $/M | Cached in | Uptime |
|---|---|---|---|---|
| DeepInfrafp4 | 2.00cheapest | 6.00 | 0.200 | 100% |
| Alibaba (via OpenRouter) | 2.00 | 6.00 | 0.250 | - |
| Modal (via OpenRouter)nvfp4 | 2.00 | 6.00 | 0.250 | 100% |
| Novita | 2.00 | 6.00 | 0.250 | 100% |
| SiliconFlow (via OpenRouter)fp8 | 2.00 | 6.00 | 0.250 | 100% |
| Together (via OpenRouter) | 2.50 | 6.25 | 0.500 | 75.6% |
| Together AI (inference host) | 2.50 | 6.25 | 0.500 | - |
| Venice | 2.50 | 7.50 | 0.312 | - |
1.2× between the cheapest and the dearest of 8 sellers.
Against its tier
- This model$8.00
- Cheapest · GLM 5.3 Flash$0.29
- Tier median · Grok 4.6$8.00
- Dearest · GPT 5.5 Pro (04-23 snapshot)$330.00
#15 of 31 flagship · cheaper than 53%
Details
- Context window
- 1.0M1,048,576 tokens
- Developer
- Alibaba / Qwen
- Intelligence
- 57.7
- Sellers
- 81.2× spread
- Cached input
- $0.20where the source publishes one
- Cheapest at
- DeepInfra
- Combined
- $8.00in + out per 1M
- Last checked
- 2026-08-28
- Availability
- 100.00%median of 5 OpenRouter endpoints, last 30 min