IONOS AI Model Hub
OpenAI-compatible API for open-weight LLMs and image models, hosted in IONOS EU data centers
IONOS AI Model Hub is the inference API from IONOS, the German hosting and cloud provider. It serves open-weight models through an OpenAI-compatible REST API: Llama 3.1 from 8B to 405B, Mistral variants, gpt-oss-120b, Qwen3 Coder, FLUX.1 and FLUX.2 for image generation, and multilingual embedding models.
Pricing is per token, starting around $0.17/1M tokens for Llama 3.1 8B, with per-image pricing for the FLUX models. Everything runs in IONOS-operated EU data centers with ISO 27001 certification and GDPR compliance, which makes it one of the larger European alternatives to US-hosted inference APIs.
Pricing: Per token usage
IONOS AI Model Hub Alternatives
Explore 89 products in the Inference APIs category. View all IONOS AI Model Hub alternatives.
vLLM
High-throughput LLM inference engine with PagedAttention for efficient GPU memory usage
Genesis Cloud
European GPU cloud, website offline and company in liquidation as of August 2026
Infer by Flow7
Responses API gateway for coding agents, with a public model catalog and per-request spend ceilings
TensorX
EU-sovereign inference API with 42+ open-source models and zero data retention
Work on IONOS AI Model Hub? Feature it at the top of Inference APIs.
Is your product missing?