Data residency for LLM APIs

Every model in the catalogue with the region its inference runs in and the retention its endpoint records, read from the endpoint rows the router itself filters on rather than from a marketing page. A model with no live endpoint is listed below as catalogue only, because that is what it is.

This deployment has no production inference backend. Every request is answered by the built-in mock backend in region dev-mock, so nothing is processed in the regions the endpoint rows record. The inference region is shown below as dev-mock for that reason, and no hosting class is claimed for a model that has no runtime here.

This page is documentation, not a certification. It records what each endpoint states about its region, its retention and its hosting class, and names the source. It is not a legal opinion, not a conformity assessment, and not a guarantee about any provider.

0 of 53 catalogue models have a live endpoint that is not the mock backend.
Model Provider Inference region Retention Hosting class Licence
Devstral Small 2 (24B) Mistral AI dev-mock zero no runtime here apache-2.0
EuroLLM-1.7B-Instruct EuroLLM (utter-project) dev-mock zero no runtime here apache-2.0
EuroLLM-22B-Instruct (2512) EuroLLM (utter-project) dev-mock zero no runtime here apache-2.0
EuroLLM-9B-Instruct (2512) EuroLLM (utter-project) dev-mock zero no runtime here apache-2.0
Mistral Small 4 (119B-A6B) Mistral AI dev-mock zero no runtime here apache-2.0
Whisper large-v3 OpenAI dev-mock zero no runtime here apache-2.0

Catalogue only — no endpoint yet

These models are documented in the catalogue but have no endpoint that can answer a request, so there is no region and no retention to report for them. Asking for one returns an error rather than routing somewhere else.

See also: Hosting · Docs · Verified