All model families

gpt-4o-mini

Median total latency per provider over the last 24 hours, from live production traffic. This is the signal Requesty's latency router samples when it picks a provider for gpt-4o-mini.

Provider latency, last 24 hours

30 minute buckets · UTC
fastest provider, last 24h · hover to inspect3 lead changes
500ms1.0s2.0s5.0s10s00:0004:0008:0012:0016:0020:00openaiazure @swedencentral

Current routing standings

gpt-4o-mini

1🏆azure @swedencentral74.0%1.37s5%2openai26.0%2.11s7%
Requesty policy10% faster
1.56sp505.00s p95
Pick one provider
1.74sp506.51s p95
fastest provider, last 24h · hover to inspect3 lead changes

Win share is simulated with the production router's Thompson Sampling over the last hour of gpt-4o-mini traffic.