Regolo
OpenAI-compatible inference API run on Italian infrastructure with zero data retention
regolo is an inference API from Seeweb, an Italian hosting provider, serving open-weight models from European data centres. The endpoint is OpenAI-compatible, so existing clients work by changing the base URL.
The catalogue covers Mistral Small, Qwen3.5, GPT-OSS, Gemma and Llama models, plus Whisper for transcription. Beyond the shared API, you can deploy a custom model from Hugging Face onto rented GPUs, billed hourly.
Data residency is the pitch: EU-only processing, zero data retention, and infrastructure certifications including ISO 27001 and CISPE. Useful if GDPR or jurisdiction rules out US-hosted providers.
Pricing: Per token usage
Regolo prices by model
Per 1M tokens, read off Regolo's own pricing page on the date shown.
| Model | Input / 1M | Cached / 1M | Output / 1M | Checked | Notes |
|---|---|---|---|---|---|
| Qwen3.8 27B | €0.50 | – | €2.10 | 1 Oct 2026 | |
| gpt-oss-120b | €1.00 | – | €4.20 | 1 Oct 2026 | |
| gpt-oss-20b | €0.10 | – | €0.42 | 1 Oct 2026 |
Regolo Alternatives
Explore 113 products in the Inference APIs category. View all Regolo alternatives.
CheapestInference
Flat-rate unlimited inference on open-weight models, sold in daily 8-hour windows
DeepInfra
Run the top AI models using a simple API, pay per use. Low cost, scalable and production ready infrastructure.
Work on Regolo? Feature it at the top of Inference APIs.
Is your product missing?