Replicate
Run and fine-tune open-source models. Deploy custom models at scale. All with one line of code.
Replicate allows users to run and fine-tune open-source models with a single line of code, offering deployment of custom models at scale. It provides serverless API endpoints for various AI tasks, including image generation, text generation, video generation, and more. The platform supports thousands of models contributed by the community and is designed for production-ready use.
Pricing: Per token usage
Replicate Alternatives
Explore 116 products in the Inference APIs category. View all Replicate alternatives.
Beam
Open-source serverless GPU cloud with sub-second cold starts and auto-scaling
BentoML
BentoML is the platform for software engineers to build AI products.
Cerebrium
Serverless GPU infrastructure for deploying AI models with sub-5 second cold starts
DeepInfra
Run the top AI models using a simple API, pay per use. Low cost, scalable and production ready infrastructure.
Fireworks AI
The production AI platform built for developers.
Groq
LPU-powered inference API for LLMs, speech, and vision models with usage-based pricing
Also listed in
Work on Replicate? Feature it at the top of Inference APIs.
Is your product missing?