Qwen3 235B A22B: price on every cloud
Open weights · 235 billion parameters · Mixture of experts · 40,960-token context · licence apache-2.0
Per token, by cloud
| Cloud | Input, per 1M tokens | Output, per 1M tokens | Cheapest region |
|---|---|---|---|
| AWS | $0.11 | $0.44 | eu-north-1 · 13 regions |
| Alibaba Cloud | $0.287 | $1.15 | eu-central-1 · 3 regions |
| Azure | — | — | not sold per token |
| OCI | — | — | not sold per token |
| Core42 | — | — | not sold per token |
| Google Cloud | — | — | not sold per token |
Run it yourself
It needs 587.7 GB of GPU memory at bf16, the precision it ships in.
| Cloud | Cheapest instance that fits | GPUs | An hour | A month, around the clock |
|---|---|---|---|---|
| AWS | p4de.24xlarge | 8 × A100 80 GB | $27.45 | $20,036 |
| Azure | Standard_ND96amsr_A100_v4 | 8 × A100 80 GB | $32.77 | $23,922 |
| Google Cloud | a2-ultragpu-8g | 8 × A100 80 GB | $40.55 | $29,602 |
| Alibaba Cloud | ecs.gn8v-8x.16xlarge | 8 × GPU H 96 GB | $58.49 | $42,698 |
More from Qwen
Similar size, open weights
- Llama 4 Maverick 17B
- Llama 3.1 405B
- GLM-5.3-Flash
- DeepSeek-V4-Flash-0731
- DeepSeek-V4-Flash
- NVIDIA-Nemotron-3-Super-120B-A12B-BF16
Inside the UAE: no cloud sells it in-country.
List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.