Icon for Packet.ai

Packet.ai

On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access

Packet.ai is an on-demand GPU cloud for AI and ML workloads, built by hosted.ai and headquartered in San Jose, California. It offers NVIDIA B200, A100, RTX 6000 Pro, RTX 4090, and L40S GPUs on Dedicated (single-tenant) or Dynamic (shared, scheduler-isolated) plans, with full root SSH, a CLI, and API access.

Rates start at USD 0.39/hr for an RTX 4090 (Dedicated) up to USD 5.90/hr for a B200 (Dedicated), billed per second with no long-term contracts. Token Factory, an OpenAI-compatible per-token inference API, is announced but not yet live.

Pricing: Pay-as-you-go

Hosting Cloud
Pricing Usage Based, $0.39/hr RTX 4090 (Dedicated)
HQ 🇺🇸 United States
Founded 2024
Compliance GDPR
Screenshot of Packet.ai webpage

Work on Packet.ai? Feature it at the top of Inference APIs.

Is your product missing?

Add it here →