Replicate
Run and fine-tune open-source models. Deploy custom models at scale. All with one line of code.
Replicate allows users to run and fine-tune open-source models with a single line of code, offering deployment of custom models at scale. It provides serverless API endpoints for various AI tasks, including image generation, text generation, video generation, and more. The platform supports thousands of models contributed by the community and is designed for production-ready use.
Pricing: Per token usage
Replicate Alternatives
Explore 100 products in the Inference APIs category. View all Replicate alternatives.
Modal
Run generative AI models, large-scale batch jobs, job queues, and much more.
Beam
Open-source serverless GPU cloud with sub-second cold starts and auto-scaling
BentoML
BentoML is the platform for software engineers to build AI products.
deepinfra
Run the top AI models using a simple API, pay per use. Low cost, scalable and production ready infrastructure.
fireworks.ai
The production AI platform built for developers.
Groq
LPU-powered inference API for LLMs, speech, and vision models with usage-based pricing
Also listed in
Work on Replicate? Feature it at the top of Inference APIs.
Is your product missing?