Sference

Managed inference for open models — EEA-resident, zero data retention by default.

Sference (operated by Neural Compute Ltd, registered in England and Wales) runs a managed inference platform for open-weight models, built as its own scheduler and runtime rather than a wrapper around a hyperscaler. Inference runs on infrastructure within the EEA, and every model is pinned to a specific version so a silent upstream swap can't change your outputs. On Opper it serves GLM-5.2, Kimi K3, DeepSeek V4 Flash, Qwen 3.6 35B-A3B, Qwen3-VL 30B, and BottleCap AI's reasoning-efficient ThinkingCap. A strong pick when you want European residency for frontier open weights.

1 route8 modelsEU🇬🇧 HQ United Kingdom
sference.com

Models on Sference

Every model we route through Sference. Compare residency, zero data retention, training posture and price at a glance, with full data-handling detail per route below.

ModelRegionZero data retentionTrainingContextInputOutput
EUNo1M$1.20$4.20
EUNo1M$3.00$15.00
EUNo1M$0.20$0.60
EUNo1M$0.50$1.50
EUNo1M$0.28$0.56
EUNo1M$0.28$0.56
EUNo1M$0.28$0.56
EUNo1M$1.20$4.20

Data handling per route

Sference hosts on 1 route. Each route has its own privacy posture, residency, and GDPR terms. Postures are maintained by Opper with a last-verification timestamp.

EU🇪🇺

Zero data retention: nothing is logged, held for abuse monitoring or used for training. No training on customer data. EEA; SCCs; DPA available.

Zero data retention
Yes. Nothing is logged, held for abuse monitoring or used for training.
Training
No training on customer data.
Logging
None
Abuse monitoring
No classifier
Caching
Content cached for replay
Subprocessor access
No subprocessor reads content
GDPR DPA
DPA available
Transfer mechanism
SCCs

Start building with 700+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation