Simplismart
Deploy and serve open-source and custom models via API, on dedicated GPUs, or in your own cloud
Simplismart is an inference and deployment platform for LLMs, vision, diffusion and speech models. It offers a usage-based API for 150+ open-source models plus dedicated deployments for custom models. A bring-your-own-cloud or on-prem option suits teams that need to keep workloads in their own VPC on AWS, Azure or GCP.
The platform also runs training jobs for fine-tuning. API pricing is per 1M tokens and listed per model; dedicated and BYOC deployments are priced separately.
Pricing: Usage-based
Simplismart Alternatives
Explore 125 products in the Inference APIs category. View all Simplismart alternatives.
Replicate
Run and fine-tune open-source models. Deploy custom models at scale. All with one line of code.
Beam
Open-source serverless GPU cloud with sub-second cold starts and auto-scaling
BentoML
BentoML is the platform for software engineers to build AI products.
Work on Simplismart? Feature it at the top of Inference APIs.
Is your product missing?