All model families

gpt-4.1-mini

Median total latency per provider over the last 24 hours, from live production traffic. This is the signal Requesty's latency router samples when it picks a provider for gpt-4.1-mini.

Provider latency, last 24 hours

30 minute buckets · UTC
fastest provider, last 24h · hover to inspect8 lead changes
500ms1.0s2.0s5.0s10s00:0004:0008:0012:0016:0020:00azure @eastus2azure @francecentralopenaiazure @westus3

Current routing standings

gpt-4.1-mini

1🏆openai42.4%1.60s12%2azure @westus333.5%1.81s3azure @francecentral24.0%2.34s51%
Requesty policy4% faster
1.85sp509.34s p95
Pick one provider
1.92sp509.78s p95
fastest provider, last 24h · hover to inspect8 lead changes

Win share is simulated with the production router's Thompson Sampling over the last hour of gpt-4.1-mini traffic.