Builder Radar
July 2026
Five European inference providers landed in July, from Italy, Norway and three from Germany. That is more than any month since March, and they are no longer pitching sovereignty as a compliance checkbox. They publish per-token rates, ship OpenAI-compatible endpoints, and expect you to compare them to Together and Fireworks on price. Underneath that, the image and video wrapper wave grew rather than slowed, and a new tell emerged for spotting it.
249
Products spotted
European inference providers
Top trend
European inference stopped selling compliance and started selling price
regolo (Italy, run by the hosting company Seeweb), SimpleLLM (Germany), KugelAudio (Berlin, text-to-speech), IONOS AI Model Hub (Germany) and WAYSCloud (Norway) all arrived within a few weeks of each other. What changed is what they lead with. GDPR and data residency are still on the page, further down. At the top is a per-token rate and an OpenAI-compatible base URL, which means they expect to be compared against Together and Fireworks on cost first and jurisdiction second. GreenPT, which arrived at the start of August, publishes a per-model price table down to cached-input rates.
KugelAudio trained its own voice models instead of reselling
Most European inference providers rent GPU capacity and serve open-weight models. It is a real business with a thin moat, and it describes almost every provider in this group. KugelAudio is the exception: it trained its own text-to-speech models in Berlin, published one under MIT, and competes on latency against ElevenLabs and Cartesia. One company is not a trend. It is the harder path, and nobody else in July took it.
The image and video wrapper wave got bigger, and easier to identify
Image and video generators were 72 of the products we looked at, comfortably the largest group and larger than June. The tell has sharpened into four things that travel together: the domain carries a major vendor's model name, pricing is credit subscriptions in the $9 to $40 band, an API is advertised but never linked to documentation, and no company is named anywhere. Seedance, Hailuo, Wan, Nano Banana and Flux each appeared across multiple unrelated domains. One went further and presented ByteDance's Seedance 2.0 as its own proprietary model, without mentioning ByteDance anywhere on the page.
Three ways to spend less on tokens, none of them switching provider
KV Cache Store precomputes the KV cache for long prompts so you skip prefill on repeated context. Tokenade sits between a coding agent and the model and rewrites requests before they are sent, billing on tokens saved. Parity Layer shadow-tests cheaper models against your baseline and routes to them once your confidence threshold is met. Three products, three mechanisms, all arriving the same month. Three is not a category yet, but a year ago the answer to a large bill was to find a cheaper provider, and none of these do that.
Agent tooling moved down the stack
None of the agent products that landed were frameworks for writing agents. Wire stores documents and records that agents query over MCP, running hybrid retrieval internally. ego lite is a browser where agents work in isolated spaces alongside you, reusing your real logged-in sessions instead of re-authenticating. Redcells runs adversarial probes against models you own. Tura is a local coding agent under AGPL. The framework layer looks settled enough that the interesting work has moved to what agents run inside and against.
Worth a look
New AI products and tools that caught our eye in July 2026
regolo โ
OpenAI-compatible inference API run on Italian infrastructure with zero data retention
Operated by Seeweb, an Italian hosting company that has been around far longer than the AI boom. Serves open-weight models from EU data centres, and rents GPUs hourly if you want to deploy your own.
KugelAudio โ
Real-time text-to-speech in 26 languages, trained and hosted in Europe
Trained its own voice models and published one under MIT. Priced per audio minute and built for voice agents, where the constraint is time to first audio.
KV Cache Store โ
Build, share and reuse precomputed KV-cache artifacts to skip redundant prefill
Treats the KV cache as a distributable artifact with a registry, digests and model metadata. An unusual place to look for savings, and the open-source CLI works on its own without the hosted registry.
Wire โ
Hosted context containers that agents query over MCP, with hybrid retrieval built in
Delivers retrieval as MCP tools, so an agent gets scoped access without you building the layer. BM25, embeddings, HyDE and reranking all sit behind one search call.
ego lite โ
Chromium browser where humans and AI agents browse in parallel through isolated Spaces
A local-first answer to headless browser infrastructure. Agents reuse the sessions you are already logged into, driven by an MIT-licensed skill that works with several coding agents.
Runware โ
Unified API for image, video, audio and 3D generation running on custom inference hardware
Owns its hardware and scheduler, which is the line between this and the media-generation wrappers filling the same category. You can also upload custom LoRAs and checkpoints.
What people are building
Based on ~249 AI products spotted across directories and launch platforms in July 2026
Get Builder Radar in your inbox
Monthly brief, no spam. Next edition: September 2026.
Building something? Submit your product.
Is your product missing?