The right provider for every scenario

RAGuardian supports three LLM deployment modes. Choose based on your risk level, compliance requirements and budget — and switch later without effort.

Regolo.ai RECOMMENDED

European AI inference provider with zero data retention. Inference runs on EU infrastructure, with a strong focus on Italy, aligning with European data sovereignty requirements.

Data residency EU infrastructure with a strong focus on Italy, reducing cross-border transfer concerns
Zero retention Prompts and outputs handled in-memory only, zero persistence
No training Your data is not used for training or fine-tuning
Compliance Simpler, auditable story for European data sovereignty requirements
Cloud embedding Cloud embedding with Qwen3-Embedding-8B model
Reranking Remote reranking after Source Diversity or MMR candidate selection
Available models:
gpt-oss-120b · gpt-oss-20b · Llama-3.3-70B · qwen3.5-122b · qwen3-coder-next · mistral-small-4-119b · more...

Enterprise controls: Customer prompts and outputs are handled in memory by the inference layer and are not persisted. Regolo does not use customer request content to train or fine-tune shared models. Logging and monitoring are designed to avoid storing full payloads, while operational access is limited to what is necessary for reliability and support under clear contractual terms.

Mistral AI

Leading French open-weight AI provider. EU data residency, transparent data usage policies.

Infrastructure Data centers in Europe
Data usage Data stored in the EU. Does not use customer data for training, except under specific conditions (feedback, moderation, free tier).
Open-weight Open-source models of European excellence
Compliance Clear, auditable policies
Available models:
mistral-medium · mistral-large · mistral-small · mistral-tiny · codestral

Local Model

Run LLMs on your server. No data ever leaves your network. Maximum privacy, zero external dependencies.

Total privacy Zero data transfer: data never leaves your datacenter
Air-gapped mode Works without internet connection
OpenAI-compatible Connect Ollama, LM Studio, vLLM or any compatible endpoint
Hardware Requires GPU for a quality model. Local embedding needs ~2GB RAM minimum.
Cost No API costs: once configured, costs are only infrastructure
Local embedding sentence-transformers: no cloud dependency for embeddings

Example config: Ollama with Qwen 3.5 122B on your server. Local embedding with sentence-transformers. No data ever leaves your network. Maximum compliance.

Choose your scenario

Feature Regolo.ai Mistral AI Local Model
Data residency European European Your datacenter
Zero data retention Yes, guaranteed Yes, with conditions* Zero transfer
No training on data Guaranteed Except specific cases Impossible
GDPR compliance By design Compliant Total
Operating costs Pay-per-use API Pay-per-use API GPU infrastructure
Hardware required None None GPU with 16GB+
Air-gapped mode Yes
Cloud embedding Qwen3-Embedding sentence-transformers

* See Mistral AI Data Usage Policy for full details.

Change provider in 30 seconds

1. Open admin dashboard

Go to /admin/config, select the provider and model for the instance. This acts as the central policy for all workspaces.

2. Set server credentials

For Regolo: REGOLO_API_KEY. For Mistral: MISTRAL_API_KEY. For local: OPENAI_BASE_URL=localhost:11434. Individual user API keys remain separate.

3. Done

Chats and connectors keep using each user's personal workspace. No data migration needed.

Need help choosing?

Our team can help evaluate the best scenario for your use case, compliance requirements and budget.