Vast.ai
GPU marketplace — rent consumer cards at half the hyperscaler price.
On-demand GPU pods for ComfyUI, vLLM, training.

Self-hosted server · Datacenter GPU (80 GB+)
Card mix ranges from RTX 3090 to H200 / B200.
Comfortable on a mid-range consumer card — no need to remortgage for an A100.
Power-user score 87/100 — consistently rated highly by people who use this every day, not just benchmark chasers.
Output is clean — you can ship it without scrubbing logos out.
Supports FP16, BF16, FP8 and 1 more — you can dial VRAM use up or down to match your card.
This page summarizes upstream documentation, release information, and editorially reviewed catalogue fields. It is not presented as a hands-on benchmark. Verify changing requirements at the official project; report stale data through our corrections channel.
A standard local install — download, install dependencies, point at your GPU.
Needs 8 GB minimum — RTX 3060 12GB or 4070 territory.
Exposes a stable API — you can build on top of it programmatically.
Catalogue entry last updated 61 days ago — re-verification due soon.
Three picks across different tradeoffs — so you don't end up with three near-clones of RunPod.
Rent A100 / H100 / 4090 / L40S pods by the minute. Templates for ComfyUI, A1111, vLLM, Ollama. Useful when your workflow outgrows your local box.
Pay-as-you-go. No free tier. Community-cloud pricing starts around $0.20/hr for older cards.