Crusoe Cloud
pricing

Flexible pricing for every workload — GPU instances, Managed Inference, and Serverless Fine-Tuning.

Two stylized white cloud shapes with gray sides and lime green square pixel patterns.

Built for value, scalability & speed

Cost-effective performance

Our AI-optimized hardware and lightweight virtualization cut waste and unlock performance, so you get more done with less.

Growth-aligned commitments

Avoid GPU lock-ins with tailored agreements that scale with your needs and budget. 

Flexible consumption models

Access our full portfolio of LLMs and generative models with pay-as-you-go pricing, or utilize our GPU/CPU offerings, with spot, on-demand, or reserved pricing options.

Pricing for 
compute, inference, and fine-tuning

Pricing for compute instances and managed AI services

GPU instances pricing

Access the latest high-performance GPUs including NVIDIA GB200 NVL72 and AMD MI355X. Pay by the hour for maximum agility and unthrottled compute, or contact us to lock in guaranteed resources at our lowest rates.

CPU instances pricing

Ideal for data processing, model checkpointing and orchestrating your GPU clusters. Choose from a variety of vCPU and RAM configurations.

CPU Type
On-demand

General-purpose

$0.04/vCPU-hr

Storage-optimized

$0.09/vCPU-hr

Storage pricing

Reliable, low-latency storage designed to handle the massive datasets and high-throughput demands of modern AI workloads.

Storage

Persistent disks

$0.08
per GiB/month

Shared disks

$0.07
per GiB/month

Container registry usage

$0.10

per GiB/month

Object Storage

$0.06

per GiB/month

Managed Kubernetes pricing

A fully managed cluster that simplifies deployment and scaling of your AI applications across GPU and CPU resources.

Managed Kubernetes

Cluster pricing

$0.10
per cluster hour

Severless Fine-Tuning pricing

Customize top performing models with your proprietary data.

Base model size
Price per 1M tokens
Models

< 16B

parameters

$0.40

Qwen3.5 2B
Qwen3.5 9B
Meta Llama 3.1 8B Instruct
Qwen3 8B

16B to 70B

parameters

$2.50

OpenAI gpt-oss 20b
Qwen3.6 35B A3B
Google gemma 4 31B it

70B to 300B

parameters

$6.00

Meta Llama 3.3 70B Instruct
OpenAI gpt-oss 120b
Qwen3 235B A22B Instruct 2507
DeepSeek V4 Flash

Serverless Inference pricing

Seamlessly integrate the industry's leading Large Language Models (LLMs) and generative models into your applications with flexible pay-as-you-go pricing.

Model
(Price per 1 million tokens)
Input tokens
Output tokens
Cached tokens

DeepSeek

V3 0324

$0.50

$1.50

$0.25

DeepSeek

V4 Pro

$1.74

$3.48

$0.15

DeepSeek

V4 Flash

$0.14

$0.28

$0.03

Gemma 4

31B-it

$0.14

$0.40

$0.14

GLM 5.1

$1.20

$4.40

$0.25

GLM 5.2

$1.40

$4.40

$0.26

GPT-OSS

120B

$0.05

$0.20

$0.05

Kimi K2.6

$0.70

$3.50

$0.35

Llama

3.3 70B Instruct

$0.25

$0.75

$0.13

Nemotron 3 Nano

30B-A3B-FP8

$0.05

$0.20

$0.03

Nemotron 3 Nano Omni 30B A3B Reasoning

(Text, Image, Video)

$0.30

$1.83

$0.30

Nemotron 3 Nano Omni 30B A3B Reasoning

(Audio)

$0.50

$1.83

$0.50

Nemotron 3 Super

120B-A12B-FP8

$0.30

$2.40

$0.15

Qwen3

235B A22B Instruct 2507

$0.22

$0.80

$0.11

Yutori n1.5

$1.50

$5.00

$1.50

Self-Serve Deployments pricing

Spin up dedicated endpoints for open and fine-tuned models in minutes — no sales engagement required. Contact sales for monthly and volume rates.

GPU type
Price per hour

Tailored Deployments pricing

Work directly with our team for the highest level of optimization and a dedicated, benchmarked endpoint. You bring the model and define the requirements, we handle the rest. Request benchmarking here.

Provisioned Throughput pricing

Ensure guaranteed throughput for your generative AI applications. Provisioned throughput is transacted via AI Model Units (AMUs). The longer your commitment, the lower your cost. Contact sales to learn more.

Frequently asked 
questions

Close-up of two server racks with illuminated green status indicator lights and bright green levers on hard drive bays.Close-up of a server rack with two rows of black hard drive bays featuring bright green release latches and green indicator lights.

Are you ready to build something amazing?

A rural landscape showing hybrid generation, a large array of solar panels alongside a canal and a line of wind turbines.