Crusoe
Renewable-powered cloud infrastructure and managed inference service for running large AI models.
| What is it | Renewable-powered cloud infrastructure and managed inference service for running large AI models. |
|---|---|
| Pricing | Paid |
| Free tier | No |
| Platform | Web Application |
| API | Yes |
| Best for | Deploying and scaling large language models in production, Running AI training workloads on high-performance GPUs |
| Domain registered | 2019 |
Data updated Aug. 1, 2026
What does Crusoe do?
Crusoe provides specialized cloud infrastructure designed specifically for artificial intelligence workloads. The company offers two main services: Crusoe Cloud, an AI-optimized infrastructure platform with high-performance NVIDIA and AMD GPUs, and Crusoe Managed Inference, a service for deploying and scaling large language models. What sets Crusoe apart is its focus on environmentally aligned energy sources—the company powers its data centers with renewable energy including wind, solar, and hydropower.
The platform includes the Crusoe Intelligence Foundry, which lets users select from top open-source models like Qwen, DeepSeek, Llama, and Nemotron, or bring their own fine-tuned models. The company claims its proprietary inference engine maintains ultra-low latency even for large-context AI workloads, with up to 9.9x faster time-to-first-token compared to alternatives. The infrastructure features optimized RDMA networking and accelerated storage designed to deploy models up to 20x faster while reducing costs.
This service is ideal for AI research teams and enterprise developers who need to run large-scale AI models in production without managing complex infrastructure. Companies building AI applications can use Crusoe to host their models with reliable, high-performance compute while benefiting from the environmental advantage of renewable-powered data centers. The platform handles the operational overhead through managed Kubernetes and Slurm clusters, allowing teams to focus on AI development rather than infrastructure management.
Key features
What makes it stand outWho is Crusoe for?
Who benefits most from this toolPricing
GPU instances
- per GPU hour pricing unit
- NVIDIA H200: $4.29/GPU-hr
- NVIDIA H100: $3.90/GPU-hr
- AMD MI300X: $3.45/GPU-hr
- NVIDIA A100 80GB SXM: $1.95/GPU-hr
- NVIDIA A100 80GB PCIe: $1.65/GPU-hr
- NVIDIA A100 40GB PCIe: $1.45/GPU-hr
CPU instances
- per vCPU hour pricing unit
- General-purpose: $0.04/vCPU-hr
- Storage-optimized: $0.09/vCPU-hr
Storage
- per GiB per month pricing unit
- Persistent disks: $0.08 per GiB/month
- Shared disks: $0.07 per GiB/month
- Container registry usage: $0.10 per GiB/month
- Object Storage: $0.06 per GiB/month
Managed Kubernetes
- per cluster hour pricing unit
- Cluster pricing: $0.10 per cluster hour
Managed inference (pay as you go)
- per 1 million tokens (input/output/cached) pricing unit
- DeepSeek V3 0324: $0.50/$1.50/$0.25 per 1M tokens
- DeepSeek V4 Pro: $1.74/$3.48/$0.15 per 1M tokens
- DeepSeek V4 Flash: $0.14/$0.28/$0.03 per 1M tokens
- Gemma-4-31B-it: $0.14/$0.40/$0.14 per 1M tokens
- GLM 5.1: $1.20/$4.40/$0.25 per 1M tokens
- GPT-OSS 120B: $0.05/$0.20/$0.05 per 1M tokens
- Llama 3.3 70B Instruct: $0.25/$0.75/$0.13 per 1M tokens
- Nemotron-3-Nano-30B-A3B-FP8: $0.05/$0.20/$0.03 per 1M tokens
- Nemotron-3-Nano-Omni 30B-A3B Reasoning (Text, Image, Video): $0.30/$1.83/$0.30 per 1M tokens
- Nemotron-3-Nano-Omni 30B-A3B Reasoning (Audio): $0.50/$1.83/$0.50 per 1M tokens
- Nemotron-3-Super-120B-A12B-FP8: $0.30/$2.40/$0.15 per 1M tokens
- Qwen3 235B A22B Instruct 2507: $0.22/$0.80/$0.11 per 1M tokens
Trust & presence
Gallery
Click any image to enlargeAlternatives in AI inference
On-demand access to NVIDIA H100, H200, and AMD MI300 GPU cloud clusters for AI and deep learning workloads.
Cloud platform providing on-demand access to multiple AI accelerators for development, training, and inference workloads.
AI cloud platform providing scalable NVIDIA GPU clusters for training and inference, with managed Kubernetes and Slurm orchestration.
Cloud infrastructure platform for deploying low-latency apps with GPUs, Kubernetes, and flat pricing
Managed platform for GPU infrastructure — turn your GPU fleet into self-service AI workspaces with Kubernetes, Slurm, and inference endpoints
A global GPU network for running AI models at scale — serverless, dedicated, or batch inference with pay-per-token pricing.
AI infrastructure platform — deploy, serve, and scale machine learning models in production with optimized inference
Cloud GPU platform for AI developers — deploy, train, and scale AI models with on-demand infrastructure