Retrieval (RAG)
Answering from retrieved passages, where the answer must stay inside the supplied context and a reranker usually matters as much as the generator.
The router maps a task to candidate models, then applies your policy: region first, then price, then the model with the better documented result.
Prefer a generator with documented citation behaviour and a long context window, keep the embedding and reranker models inside the same residency boundary, and route document extraction to a dedicated OCR model.
{"model":"llmeu-auto","messages":[{"role":"user","content":"…"}],"llmeu":{"task":"rag"}}
Recommended
No first-party endpoint yet
No first-party endpoint serves this task yet. The models listed below are catalogued for reference and run under their own provider’s terms.