Packet.ai
On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access
Packet.ai is an on-demand GPU cloud for AI and ML workloads, built by hosted.ai and headquartered in San Jose, California. It offers NVIDIA B200, A100, RTX 6000 Pro, RTX 4090, and L40S GPUs on Dedicated (single-tenant) or Dynamic (shared, scheduler-isolated) plans, with full root SSH, a CLI, and API access.
Rates start at USD 0.39/hr for an RTX 4090 (Dedicated) up to USD 5.90/hr for a B200 (Dedicated), billed per second with no long-term contracts. Token Factory, an OpenAI-compatible per-token inference API, is announced but not yet live.
Pricing: Pay-as-you-go
Packet.ai Alternatives
Explore 104 products in the Inference APIs category. View all Packet.ai alternatives.
AiQu
Swedish GPU infrastructure and LLM hosting platform with API-first deployment, no Kubernetes required
CheapestInference
Flat-rate unlimited inference on open-weight models, sold in daily 8-hour windows
SiliconFlow
OpenAI-compatible API serving 200+ open-source LLM and multimodal models
Work on Packet.ai? Feature it at the top of Inference APIs.
Is your product missing?