LLM Opex Every way to run an LLM, every cloud

Models

Qwen3 235B A22B: price on every cloud

Open weights · 235 billion parameters · Mixture of experts · 40,960-token context · licence apache-2.0

Per token, by cloud

US dollars per million tokens, each cloud's cheapest region.
CloudInput, per 1M tokensOutput, per 1M tokensCheapest region
AWS$0.11$0.44eu-north-1 · 13 regions
Alibaba Cloud$0.287$1.15eu-central-1 · 3 regions
Azure——not sold per token
OCI——not sold per token
Core42——not sold per token
Google Cloud——not sold per token

Run it yourself

It needs 587.7 GB of GPU memory at bf16, the precision it ships in.

CloudCheapest instance that fitsGPUsAn hourA month, around the clock
AWSp4de.24xlarge8 × A100 80 GB$27.45$20,036
AzureStandard_ND96amsr_A100_v48 × A100 80 GB$32.77$23,922
Google Clouda2-ultragpu-8g8 × A100 80 GB$40.55$29,602
Alibaba Cloudecs.gn8v-8x.16xlarge8 × GPU H 96 GB$58.49$42,698

More from Qwen

Similar size, open weights

Inside the UAE: no cloud sells it in-country.

List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.