Skip to content
Dashboard

Crusoe models

Run model inference with fast time-to-first-token, low latency, limitless throughput, and resilient scaling.
Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Regional Inference
Free Tier
Released
1M
0.3s
508tps
$0.80/MFast $2.10/M
$2.52/MFast $6.60/M
Read:$0.14/M
Write:
alibaba logo
baseten logo
crusoe logo
+12
US
06/16/2026