LLM Opex Every way to run an LLM, every cloud

Models

OpenAI gpt-oss 120B: price on every cloud

Open weights · 117 billion parameters · Mixture of experts · 131,072-token context · licence apache-2.0

Per token, by cloud

US dollars per million tokens, each cloud's cheapest region.
CloudInput, per 1M tokensOutput, per 1M tokensCheapest region
AWS$0.075$0.30eu-north-1 · 15 regions
Google Cloud$0.09$0.36global
Core42$0.15$0.37uae
Azure$0.15$0.60australiaeast · 41 regions
OCI$0.15$0.60global · 3 regions
Alibaba Cloud——not sold per token

Run it yourself

It needs 73 GB of GPU memory at int4, the precision it ships in.

CloudCheapest instance that fitsGPUsAn hourA month, around the clock
AzureStandard_NC24ads_A100_v41 × A100 80 GB$3.67$2,681
Google Cloudg4-standard-481 × RTX PRO 6000 96 GB$4.50$3,285
AWSp5.4xlarge1 × H100 80 GB$6.88$5,022
Alibaba Cloudecs.gn8v.4xlarge1 × GPU H 96 GB$7.31$5,337

At the cheapest per-token price, running it on Azure Standard_NC24ads_A100_v4 pays off above roughly 20.4 billion tokens a month.

More from OpenAI

Similar size, open weights

Inside the UAE: sold in-country by Core42, Azure.

List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.