LLM Opex Every way to run an LLM, every cloud

Models

Intern-S1: GPU memory and hosting cost

Open weights · 240.71 billion parameters · Mixture of experts · 65,536-token context · licence apache-2.0

No cloud sells this model per token: to use it you run it yourself, on GPU instances you rent.

Run it yourself

It needs 601.8 GB of GPU memory at bf16, the precision it ships in.

CloudCheapest instance that fitsGPUsAn hourA month, around the clock
AWSp4de.24xlarge8 × A100 80 GB$27.45$20,036
AzureStandard_ND96amsr_A100_v48 × A100 80 GB$32.77$23,922
Google Clouda2-ultragpu-8g8 × A100 80 GB$40.55$29,602
Alibaba Cloudecs.gn8v-8x.16xlarge8 × GPU H 96 GB$58.49$42,698

More from InternLM

Similar size, open weights

List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.