Packet.ai
On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access
Packet.ai is an on-demand GPU cloud for AI and ML workloads, built by hosted.ai and headquartered in San Jose, California. It offers NVIDIA B200, A100, RTX 6000 Pro, RTX 4090, and L40S GPUs on Dedicated (single-tenant) or Dynamic (shared, scheduler-isolated) plans, with full root SSH, a CLI, and API access.
Rates start at USD 0.39/hr for an RTX 4090 (Dedicated) up to USD 5.90/hr for a B200 (Dedicated), billed per second with no long-term contracts. Token Factory, an OpenAI-compatible per-token inference API, is announced but not yet live.
Pricing: Pay-as-you-go
Packet.ai Alternatives
Explore 100 products in the Inference APIs category. View all Packet.ai alternatives.
AKI.IO
European AI API for open-source models on EU infrastructure
CheapestInference
Flat-rate unlimited inference on open-weight models, sold in daily 8-hour windows
Work on Packet.ai? Feature it at the top of Inference APIs.
Is your product missing?