Ornith-1.5-35B-A3B-FP8: GPU memory and hosting cost
Open weights · 35.95 billion parameters · Mixture of experts · 262,144-token context · licence mit
No cloud sells this model per token: to use it you run it yourself, on GPU instances you rent.
Run it yourself
It needs 44.9 GB of GPU memory at fp8, the precision it ships in.
| Cloud | Cheapest instance that fits | GPUs | An hour | A month, around the clock |
|---|---|---|---|---|
| AWS | g6e.xlarge | 1 × L40S 48 GB | $1.86 | $1,359 |
| Alibaba Cloud | ecs.gn8is.2xlarge | 1 × L20 48 GB | $2.26 | $1,648 |
| Azure | Standard_NC72ds_xl_RTXPRO6000BSE_v6 | 1 × RTX PRO 6000 48 GB | $2.83 | $2,066 |
| Google Cloud | g4-standard-48 | 1 × RTX PRO 6000 96 GB | $4.50 | $3,285 |
More from Other
- dolphin-2.9.1-yi-1.5-34b
- OTel-2.0-LLM-31B-IT
- JiRackUltra_1b
- moondream2
- TinyLlama-1.1B-Chat-v1.0
- gpt2-large
Similar size, open weights
List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.