Skip to content
Dashboard

Crusoe models

Run model inference with fast time-to-first-token, low latency, limitless throughput, and resilient scaling.
Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Regional Inference
Free Tier
Released
1M
0.4s
594tps
$0.90/M
$2.84/M
Read:$0.17/M
Write:
alibaba logo
baseten logo
crusoe logo
+11
US
06/16/2026