On 2 September someone in r/AI_Governance posted the question every European platform team eventually has to answer: best AI gateway with EU data sovereignty, what did legal verify? They had started with nine gateways, read every security page, and concluded that "most of them treat EU sovereignty as an enterprise tier conversation" while they needed "traffic staying in EU, subprocessor and auditable, DPA ready to sign" as the bare minimum.
The security page is the wrong place to start. The right place is the model catalog, because the constraint that decides an EU architecture is not the gateway's promise, it is which models exist in an EU region at all and what they cost there. So we ran the check: take the Requesty Intelligence ranking, which orders models by the Artificial Analysis Intelligence Index, and look each of the top 16 up in the live catalog for a deployment flagged geolocation: eu.

The table
Prices are per million tokens. "Global" is the price shown on the ranking page. "EU" is the cheapest EU region deployment in the catalog on 10 September 2026, blended 3:1 input to output for the premium column.
| Rank | Model | Index | Global in / out | Cheapest EU in / out | EU route | Premium |
|---|---|---|---|---|---|---|
| 1 | Claude Fable 5.1 | 53.4 | 10.00 / 50.00 | 11.00 / 55.00 | vertex @eu | +10% |
| 2 | GPT-6 Astra | 52.8 | 10.00 / 50.00 | none | ||
| 3 | Claude Opus 5 | 50.7 | 5.00 / 25.00 | 5.50 / 27.50 | bedrock @eu-central-1, eu-west-1, eu-west-3, eu-north-1; vertex @eu | +10% |
| 4 | Claude Fable 5 | 49.7 | 10.00 / 50.00 | 11.00 / 55.00 | vertex @eu | +10% |
| 5 | GPT-5.6 Sol | 47.1 | 4.00 / 20.00 | 5.50 / 33.00 | azure @francecentral, swedencentral, germanywestcentral, westeurope | +55% |
| 6 | GLM-5.3 | 44.9 | 1.40 / 4.40 | 1.20 / 4.20 | sference, tensorx | -9% |
| 7 | Grok 4.6 | 44.4 | 2.00 / 6.00 | none | ||
| 8 | Kimi K3 | 43.8 | 3.00 / 15.00 | 3.00 / 15.00 | nebius, sference, tensorx | 0% |
| 9 | GPT-5.6 Terra | 42.3 | 2.00 / 12.00 | 2.20 / 13.20 | azure, four EU regions | +10% |
| 10 | Claude Opus 4.8 | 42.0 | 5.00 / 25.00 | 5.50 / 27.50 | bedrock, four EU regions; vertex @eu | +10% |
| 11 | GLM-5.3 Flash | 41.9 | 0.15 / 0.50 | 0.20 / 0.50 | sference, tensorx | +16% |
| 12 | Gemini 3.8 Flash | 41.2 | 0.75 / 3.75 | 0.825 / 4.125 | vertex @eu | +10% |
| 13 | Claude Opus 4.7 | 40.7 | 5.00 / 25.00 | 5.50 / 27.50 | bedrock, four EU regions; vertex @eu | +10% |
| 14 | Qwen3.8 Max | 40.3 | 2.00 / 6.00 | none | ||
| 15 | Qwen3.8 2.4T A95B | 40.0 | 2.50 / 6.00 | 2.50 / 6.00 | tensorx | 0% |
| 16 | Qwen3.8 Flash Next | 39.9 | 0.20 / 0.50 | 0.20 / 0.50 | tensorx | 0% |
Three things fall out of it.
The frontier is available in the EU, with one large exception. Anthropic's entire current line, including the number one model on the ranking, runs in EU regions through Vertex and Bedrock, 47 deployments across six regions at last count. OpenAI's GPT-5.6 family and GPT-5.5 run on Azure in four EU regions. Google's Gemini 3.8 Flash runs on Vertex EU. The exception is GPT-6 Astra, which a week after launch has no EU region deployment anywhere in the catalog. We covered why Astra is not production ready for European teams last week; the ranking makes the cost concrete. If your organisation is EU only, the smartest model you can call is Fable 5.1 at 53.4, and the second smartest is Opus 5 at 50.7, not Astra at 52.8.
The residency premium is 10%, and it is the vendors' number, not ours. Ten of the thirteen EU capable models cost exactly 10% more in their EU region than the global list price: Anthropic on Bedrock and Vertex, OpenAI on Azure, Google on Vertex. OpenAI's own data residency documentation states that "data residency endpoints are charged a 10% uplift for models released on or after March 5, 2026." The hyperscalers apply the same uplift to their EU regions. Budget for it as a line item: an EU only organisation pays 1.1x on the frontier models, and the gateway does not add to that.
Sol's 55% is a promo artefact worth understanding. GPT-5.6 Sol's global price of $4 and $20 is a promotional rate that OpenAI labelled as available at least through 21 November; Azure's EU regions bill Sol at $5.50 and $33, a 10% uplift on the pre-promo $5 and $30. So the EU premium on Sol today is "promo versus no promo plus 10%". We wrote about list prices that depend on the calendar last week; this is the EU version of the same problem. When the promo ends, Sol's EU premium drops back to 10%.
The open weight rows behave differently, and that is the point of open weights. GLM-5.3 is cheaper on its EU flagged hosts than the global list price, Kimi K3 and the Qwen3.8 family are at parity, and GLM-5.3 Flash carries a small premium on one host. Open weights let European hosts compete on price for the same model, and the provider variance that is a headache elsewhere is a discount here.
The cheapest models you can run in the EU
The cheapest models ranking is global. Filter the same catalog to EU flagged deployments and rank by blended 3:1 price, and the EU budget tier on 10 September looks like this:
| Model | Host / region | Input | Output | Context |
|---|---|---|---|---|
| Nemotron 3 Nano Omni | nebius | 0.06 | 0.24 | 300K |
| Qwen3 30B A3B Instruct | nebius | 0.10 | 0.30 | 128K |
| Qwen3 32B | nebius | 0.10 | 0.30 | 128K |
| Gemma 3 27B | nebius | 0.10 | 0.30 | 128K |
| GPT-5 Nano | azure @germanywestcentral | 0.055 | 0.44 | 200K |
| Mistral Small 2503 | mistral | 0.11 | 0.33 | 32K |
| Devstral Small | mistral | 0.11 | 0.33 | 131K |
| Gemini 2.5 Flash Lite | vertex @europe-west4 | 0.10 | 0.40 | 1M |
| GPT-4.1 Nano | azure @swedencentral | 0.11 | 0.44 | 1M |
| Llama 3.3 70B | nebius | 0.13 | 0.40 | 128K |
| DeepSeek V4 Flash 0731 | tensorx | 0.25 | 0.30 | 1M |
| GLM-5.3 Flash | tensorx | 0.20 | 0.50 | 1M |
Two of these deserve a second look against the intelligence ranking. GLM-5.3 Flash scores 41.9, above Gemini 3.8 Flash and one place below Opus 4.8, at $0.20 in and $0.50 out in an EU region. DeepSeek V4 Flash 0731 is a 1M context model at $0.25 and $0.30. Both are the kind of model you route the 80% of traffic that does not need a frontier answer to, and both exist in Europe.
What "in the catalog" means and does not mean
A note on rigour, because the r/AI_Governance poster's complaint was that vendor pages overstate. The geolocation: eu flag in the Requesty catalog records where the model inference runs, as declared by the host: Bedrock and Vertex regions in the model id, Azure regions in the model id, Mistral's La Plateforme in the EU, and EU flagged third party hosts for open weights. It says nothing about where the gateway processed the request. Those are two separate layers, and confusing them is the most common EU compliance mistake we see. The EU routing documentation spells it out: the EU endpoint at router.eu.requesty.ai guarantees that Requesty's processing, logging and caching stay in Frankfurt (AWS eu-central-1); an EU region model guarantees the inference stays in the EU; you need both for end to end residency.
Also: the table is a snapshot from 10 September 2026. Catalog entries change weekly. The /eu/claude, /eu/openai, /eu/gemini and /eu/mistral pages render the same EU rows live from the catalog, so use those for the current list rather than this post.
Making the EU column the only column
The table above tells you what is possible. Turning "possible" into "enforced" is three settings, none of which require code:
- Restrict the serving region. In the Compliance panel, set the organisation's allowed Requesty region to EU. Every request now has to enter through
router.eu.requesty.ai; the global, US and AP endpoints reject the organisation's keys, and the Playground only offers EU. - Build the Approved Models list from EU regions only. Filter the Model Library by EU regions, select, approve. A request for
anthropic/claude-opus-5(the global route) is rejected;bedrock/claude-opus-5@eu-central-1succeeds. This is the control that stops inference leaving the EU by accident, and it is what the r/AI_Governance poster meant by "enforced, not listed". - Turn on Zero Data Retention if the requirement is that prompts and responses are never stored. ZDR is organisation wide and one way: once enabled it disables payload logging on every key and blocks anyone, including admins, from turning it back on.
Then use access lists per team to narrow further: a support automation key that may only call Gemini 2.5 Flash Lite in Belgium, an engineering key that may call Opus 5 in Frankfurt. The team model access post walks through the layering.
For the paperwork, the DPA covers GDPR Article 28 processor obligations and the EU residency commitments, and turns around in a few business days.
The takeaway
Thirteen of the sixteen smartest models in the world have an EU region today, at a 10% premium set by the model vendors rather than by the routing layer. The three that do not (GPT-6 Astra, Grok 4.6, Qwen3.8 Max) are the honest limit of an EU only architecture in September 2026, and the honest answer to "can we use Astra" for a European regulated team is "not yet, use Fable 5.1 or Opus 5."
Everything below the frontier is better than people assume: a budget tier of open weight and small proprietary models exists in Europe at prices that match or beat the global cheapest ranking. What turns that catalog into compliance is enforcement at the organisation level: EU serving region, EU only approved models, ZDR. Start at the /eu page, or sign up and point your base URL at Frankfurt.
Frequently asked questions
- Which frontier models have EU region deployments?
- As of 10 September 2026 the Requesty catalog lists EU region deployments for Claude Fable 5.1, Fable 5, Opus 5, Opus 4.8 and Opus 4.7 (Vertex EU and Bedrock Frankfurt, Ireland, Paris, Stockholm), GPT-5.6 Sol, Terra and Luna, GPT-5.5 and GPT-5.4 (Azure France Central, Sweden Central, Germany West Central, West Europe), Gemini 3.8 Flash (Vertex EU), and open weight models such as GLM-5.3, Kimi K3 and Qwen3.8 on EU flagged hosts. GPT-6 Astra, Grok 4.6 and Qwen3.8 Max have no EU deployment in the catalog.
- How much more do models cost in EU regions?
- For the top ranked models the typical premium is 10%: Claude Opus 5 is $5.00 and $25.00 per million tokens globally and $5.50 and $27.50 on Bedrock or Vertex in the EU, and Gemini 3.8 Flash is $0.75 and $3.75 globally versus $0.825 and $4.125 on Vertex EU. OpenAI documents the same 10% uplift for its own data residency endpoints. GPT-5.6 Sol is the outlier at 55% because its global price is a time limited promotion that Azure does not mirror.
- What is the cheapest model available in an EU region?
- On a blended 3:1 input to output basis the cheapest EU deployments in the catalog are Nemotron 3 Nano Omni on Nebius ($0.06 in, $0.24 out), Qwen3 30B and 32B and Gemma 3 27B on Nebius ($0.10 in, $0.30 out), GPT-5 Nano on Azure Germany West Central ($0.055 in, $0.44 out), Mistral Small ($0.11 in, $0.33 out) and Gemini 2.5 Flash Lite on Vertex EU regions ($0.10 in, $0.40 out).
- How do I make sure my organisation can only call EU models?
- Two settings. Restrict the organisation's serving region to EU in the Compliance panel so every request has to enter through router.eu.requesty.ai, and build the Approved Models list from EU region models only. Any request for a model outside that list is rejected, so no inference can leave the EU by accident.
- MAY '26
EU Compliant AI Routing: Why Your LLM Gateway Needs to Be GDPR and EU AI Act Ready
The EU AI Act's high risk provisions take full effect on August 2, 2026. Edge based routers like OpenRouter on Cloudflare give you no audit trail of where your data went. Here is why Requesty's EU infrastructure in Frankfurt is the compliant choice for AI routing in Europe.
- JUN '26
OpenRouter in Europe: Why EU Teams Are Switching to an EU-Hosted AI Gateway
Searching for OpenRouter with EU data residency? Here is what European teams need from an LLM gateway (GDPR compliance, Frankfurt hosting, zero data retention) and how to get all of it without changing your code.
- SEP '26
GPT-6 Astra is the best model you may not be allowed to call: gated access, the EU gap and safety stops
Astra launched 3 September at $10 and $50 per million tokens with the first Critical cybersecurity rating in OpenAI's framework. Three days later, most developers cannot call it, Foundry customers in the EU cannot keep it in region, and the ones who can call it are learning that a safety monitor can end a job mid-run. The benchmarks are one story. Getting it into production is another.
- SEP '26
The list price is now a range: peak hours, promo windows and host floors
In August the frontier rate card stopped being a number. DeepSeek bills V4 Pro at $1.32 input during seven weekday hours and $0.66 the rest of the time. OpenAI cut GPT-5.6 Sol to $4 and $20 but only through 21 November. Gemini 3.8 Flash doubles on 1 January. And the cheapest GPT-4-class price in the market is set by a reseller, not by the lab. Your cost model needs a clock, a calendar and a host column.
- MAY '26
Give every team exactly the models they need (and nothing more)
Approved Models set the org floor. Access Lists narrow it per team or per key. Expiring keys enforce rotation. Together they give platform engineers a governance stack that scales from 3 people to 300 without a single Slack argument about who broke prod.
- JUN '26
EU AI Compliance in 2026: The 7 Regulations Every Enterprise Now Has to Answer For
In three years the EU enacted seven regulations that govern how companies build, buy, and run AI. GDPR, NIS2, DORA, the AI Act, the Data Act, the Cyber Resilience Act, and the European Health Data Space now stack on top of each other. This is a plain reading of what each one demands, the dates that have already passed, the ones coming next, and the single capability they all converge on: by the end of 2027 you have to prove governance, auditability, and data residency.
