Icon for QuantaCloud

QuantaCloud

On-demand NVIDIA GPU VMs in US Midwest regions, prepaid and billed by the second, plus quoted reserved clusters

QuantaCloud rents on-demand NVIDIA GPU virtual machines from a self-serve console, reached over SSH or in the browser, with templates for PyTorch with Jupyter, Open WebUI with Ollama and ComfyUI. Instances run on sourced capacity in three US Midwest regions. A public offers API, no key needed, lists live configurations and prices.

Billing is prepaid: the first hour is charged at launch and unused seconds come back on stop. On 8 October 2026 an RTX A6000 started at $0.52/GPU-hr. Stopping an instance deletes its disk, with no volumes or snapshots. Reserved clusters are quoted separately.

Pricing: Hourly

Hosting Cloud
Pricing $0.52/GPU-hr (RTX A6000)
HQ 🇺🇸 United States
Screenshot of QuantaCloud webpage

Work on QuantaCloud? Feature it at the top of Inference APIs.

Is your product missing?

Add it here →