Greenference

French inference host for small open-weight models, served from GPU capacity in the EU.

Greenference SAS is a French company, registered in Lyon, that serves small open-weight models through an OpenAI-compatible API. It runs on GPU capacity rented inside the European Union, behind a request router in Amsterdam. On Opper it serves the Qwen3, Qwen3.5, Qwen3.6 and Qwen3.8 families alongside Gemma 4, GLM-4.7-Flash, gpt-oss-20b and Llama 3.1 8B. A good pick for inexpensive EU-hosted inference on 8B to 32B class models.

1 route11 modelsEU🇫🇷 HQ France
greenference.com

Models on Greenference

Every model we route through Greenference. Compare residency, zero data retention, training posture and price at a glance, with full data-handling detail per route below.

ModelRegionZero data retentionTrainingContextInputOutput
EUNo262K$0.05$0.90
EUNo16K$0.15$1.00
EUNo203K$0.03$0.20
EUNo16K$0.04$0.07
EUNo33K$0.0090$0.04
EUNo16K$0.01$0.02
EUNo131K$0.07$0.35
EUNo41K$0.10$0.28
EUNo33K$0.45$0.45
EUNo16K
EUNo16K$0.10$0.28

Data handling per route

Greenference hosts on 1 route. Each route has its own privacy posture, residency, and GDPR terms. Postures are maintained by Opper with a last-verification timestamp.

European Union🇪🇺

Zero data retention: nothing is logged, held for abuse monitoring or used for training. No training on customer data. EU; DPA available.

Zero data retention
Yes. Nothing is logged, held for abuse monitoring or used for training.
Training
No training on customer data.
Logging
None
Abuse monitoring
No classifier
Caching
Content cached for replay
Subprocessor access
Subprocessors may read content
GDPR DPA
DPA available
Transfer mechanism
Not applicable — data stays in EU

Start building with 700+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation