SiliconFlow
OpenAI-compatible API serving 200+ open-source LLM and multimodal models
SiliconFlow is an inference platform that serves open-source LLMs alongside image, video, and audio models through a single OpenAI-compatible API. It hosts 200+ models, including the DeepSeek, Qwen, and Kimi families, with per-token usage pricing and serverless deployment.
It also offers reserved GPU options for predictable billing. Developers use it as a drop-in alternative to other hosted inference APIs, switching by changing the base URL and key.
Pricing: Per token usage
SiliconFlow Alternatives
Explore 92 products in the Inference APIs category. View all SiliconFlow alternatives.
Scaleway
European serverless AI inference APIs, 100% hosted in Europe
Varion
OpenAI-compatible proxy that trims LLM token usage before requests reach your provider
Alibaba Cloud Model Studio
Hosted API access to Qwen and third-party models across six global regions
Packet.ai
On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access
Work on SiliconFlow? Feature it at the top of Inference APIs.
Is your product missing?