Anyscale
Fast, cost-efficient, serverless APIs for LLM Serving and Fine Tuning
Anyscale Endpoints offers fast, cost-efficient, serverless APIs for serving and fine-tuning Large Language Models (LLMs) with a focus on production-readiness. Users can start with common LLMs, including the Llama-2 family and Mistral 7B, and fine-tune them for specific applications.
Pricing: Pay-as-you-go
HQ
🇺🇸 United States
Anyscale Alternatives
Explore 91 products in the Inference APIs category. View all Anyscale alternatives.
Packet.ai
On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access
From $0.39/hr RTX 4090 (Dedicated)
vLLM
High-throughput LLM inference engine with PagedAttention for efficient GPU memory usage
Open Source
Free Trial
From Free (open-source)
Also listed in
Work on Anyscale? Feature it at the top of Inference APIs.
Is your product missing?