Vast.ai

GPU cloud marketplace for renting AI compute.

Category: Inference
Pricing: PAID
Source: Proprietary
Hosting: Cloud
Platforms: WebCLIAPI
Models: Model-agnostic
Verified: Jun 14, 2026

Vast.ai is a marketplace for renting GPUs, connecting people who need AI compute with hosts who list spare hardware — from solo owners to Tier-4 data centers. Prices are set by supply and demand, billed per second, and queryable through code. It supports on-demand, interruptible (spot), and reserved instances across tens of thousands of GPUs and dozens of GPU types.

Capabilities 2

What it actually does — grouped by capability family.

GPU compute (primary capability)
Model inference / serving (secondary capability)

Pros & cons

Often the cheapest GPUs via marketplace
Per-second billing, $5 minimum
On-demand, spot, and reserved options
Large catalog of GPU types

Host quality and reliability vary
Not a managed inference platform
Interruptible instances can be reclaimed

View Runpod details
InferencePAID
Runpod
Runpod
GPU cloud for AI — on-demand instances and serverless inference.
Runpod is an AI developer cloud for renting GPUs on demand or running auto-scaling serverless inference endpoints. Serverless workers bill by the millisecond, scale to zero when idle, and advertise sub-200ms cold starts; on-demand Pods and multi-node Clusters cover training and long-running jobs. A Community Cloud tier offers cheaper, peer-sourced GPUs alongside the vendor-operated Secure Cloud.
Serverless auto-scaling inference
Community Cloud less reliable/secure
- gpu-cloud
- serverless
- inference
- deployment
- +1
Open
View CoreWeave details
InferencePAID
CoreWeave
CoreWeave
The AI hyperscaler — GPU cloud built for large-scale training and inference.
CoreWeave is a purpose-built AI cloud renting large-scale NVIDIA GPU capacity for training and inference, layered with managed Kubernetes, AI object storage, and Mission Control observability. Public on Nasdaq since March 2025, it counts most leading AI labs — including OpenAI, Meta, and Anthropic — among its customers, with a contracted revenue backlog reported near $100B in 2026.
Frontier-scale GPU capacity
Enterprise-oriented; no free tier
- gpu-cloud
- ai-hyperscaler
- training
- inference
- +1
Open
View Lambda details
InferencePAID
Lambda
Lambda
GPU cloud for AI training — on-demand GPUs, 1-Click Clusters, and superclusters.
Lambda is a GPU cloud for AI training and inference, spanning on-demand HGX B200 and H100 instances, self-serve 1-Click Clusters, and single-tenant superclusters built on NVIDIA's latest generations. A GPU specialist since 2012, it sells compute by the hour without long-term hyperscaler contracts and co-engineers large deployments with NVIDIA.
Single GPUs up to superclusters
No free tier
- gpu-cloud
- training
- clusters
- nvidia
- +1
Open
View Modal details
InferenceFREEMIUM
Modal
Modal Labs
Serverless GPUs. Run training, inference, batch jobs from Python.
Define cloud workloads in Python, deploy with one command — GPU access on demand, fast cold starts, fair-share pricing. The default 'I need to fine-tune a model from a Jupyter cell' platform.
Python-decorator infra, no YAML/Dockerfiles
SDK lock-in; migrating means rewriting
- gpu
- serverless
- python
- training
Open

Open Vast.ai

Vast.ai

Capabilities 2

Pros & cons

Tags

Further reading

Runpod

CoreWeave

Lambda

Modal