Replicate
Replicate
Model marketplace · Pay per second · One-line deploys
- Pay as you go$0/mo, then $5.49 per GPU hour
$1,098/mo
at 200 GPU hours
Replicate pricing detailsCompute & Storage · plans & pricing
4 providers · USD monthly pricing · ordered by popularity
GPU clouds rent accelerators by the second or hour for model inference, fine-tuning and training. Prices vary widely by GPU type, availability and serverless vs reserved capacity.
Replicate
Model marketplace · Pay per second · One-line deploys
$1,098/mo
at 200 GPU hours
Replicate pricing detailsRunPod
On-demand GPUs · Serverless endpoints · Community cloud
$398/mo
at 200 GPU hours
RunPod pricing detailsModal
Python-native · Scale to zero · Fast cold starts
$760/mo
at 200 GPU hours
Modal pricing detailsRelative to each provider's typical workload
1×typical usage
Providers bill in different units, so usage is never converted between them. Each estimate scales that provider's own typical workload, shown under its price. Estimates exclude taxes and optional add-ons.
Compare the full directory
| ReplicateReplicate | Model marketplace · Pay per second · One-line deploys | Trial | $1,098/moat 200 GPU hoursPricing details |
| RunPodRunPod | On-demand GPUs · Serverless endpoints · Community cloud | None | $398/moat 200 GPU hoursPricing details |
| ModalModal | Python-native · Scale to zero · Fast cold starts | $30/mo credits | $760/moat 200 GPU hoursPricing details |
| CoreWeaveCoreWeave | H100 clusters · Kubernetes-native · Enterprise AI | Custom | $1,232/moat 200 GPU hoursPricing details |
The buying guide
Replicate, RunPod, Modal are the most widely used providers in this list. The directory is ordered by popularity by default; sort by price or name to change it.
At each provider's typical usage, RunPod has the lowest estimate in this list at $398/mo. Estimates use tracked plans and exclude taxes and add-ons.
4 providers are tracked on this page, each with a dedicated pricing page.