Nemotron-Labs-Diffusion-14B: GPU memory and hosting cost
Open weights · 13.51 billion parameters · 262,144-token context · licence other
No cloud sells this model per token: to use it you run it yourself, on GPU instances you rent.
Run it yourself
It needs 33.8 GB of GPU memory at bf16, the precision it ships in.
| Cloud | Cheapest instance that fits | GPUs | An hour | A month, around the clock |
|---|---|---|---|---|
| AWS | g6e.xlarge | 1 × L40S 48 GB | $1.86 | $1,359 |
| Alibaba Cloud | ecs.gn8is.2xlarge | 1 × L20 48 GB | $2.26 | $1,648 |
| Azure | Standard_NC72ds_xl_RTXPRO6000BSE_v6 | 1 × RTX PRO 6000 48 GB | $2.83 | $2,066 |
| Google Cloud | a2-highgpu-1g | 1 × A100 40 GB | $3.67 | $2,682 |
More from NVIDIA
- Qwen3.6-35B-A3B-NVFP4
- NVIDIA-Nemotron-3-Nano-4B-BF16
- Gemma-4-31B-IT-NVFP4
- Qwen2.5-VL-7B-Instruct-NVFP4
- NVIDIA-Nemotron-3-Super-120B-A12B-BF16
- Qwen3.5-122B-A10B-NVFP4
Similar size, open weights
List prices collected 5 October 2026, refreshed every three hours. No negotiated rates or taxes.