IONOS AI Model Hub
OpenAI-compatible API for open-weight LLMs and image models, hosted in IONOS EU data centers
IONOS AI Model Hub is the inference API from IONOS, the German hosting and cloud provider. It serves open-weight models through an OpenAI-compatible REST API: Llama 3.1 from 8B to 405B, Mistral variants, gpt-oss-120b, Qwen3 Coder, FLUX.1 and FLUX.2 for image generation, and multilingual embedding models.
Pricing is per token, starting around $0.17/1M tokens for Llama 3.1 8B, with per-image pricing for the FLUX models. Everything runs in IONOS-operated EU data centers with ISO 27001 certification and GDPR compliance, which makes it one of the larger European alternatives to US-hosted inference APIs.
Pricing: Per token usage
IONOS AI Model Hub prices by model
Per 1M tokens, read off IONOS AI Model Hub's own pricing page on the date shown.
| Model | Input / 1M | Cached / 1M | Output / 1M | Checked | Notes |
|---|---|---|---|---|---|
| Llama 3.3 70B Instruct | €0.65 | – | €0.65 | 25 Sep 2026 | Price on ionos.de; the US site lists $0.71 / $0.71 |
| gpt-oss-120b | €0.15 | – | €0.65 | 25 Sep 2026 | Price on ionos.de; the US site lists $0.17 / $0.71 |
IONOS AI Model Hub Alternatives
Explore 107 products in the Inference APIs category. View all IONOS AI Model Hub alternatives.
CheapestInference
Flat-rate unlimited inference on open-weight models, sold in daily 8-hour windows
LLM Tech
EU inference provider serving Qwen3.8-27B from dedicated Helsinki GPUs it rents and operates itself, with zero data r...
WAYSCloud
Norwegian cloud platform with an LLM inference API running open-weight models in Norway
Work on IONOS AI Model Hub? Feature it at the top of Inference APIs.
Is your product missing?