All model families
gpt-4o-mini
Median total latency per provider over the last 24 hours, from live production traffic. This is the signal Requesty's latency router samples when it picks a provider for gpt-4o-mini.
Provider latency, last 24 hours
30 minute buckets · UTCfastest provider, last 24h · hover to inspect3 lead changes
Current routing standings
gpt-4o-mini
1🏆azure @swedencentral74.0%1.37s3.36s5%$0.202openai⚡26.0%2.11s9.67s7%$0.22Requesty policy10% faster
1.56sp505.00s p95
Pick one provider
1.74sp506.51s p95
fastest provider, last 24h · hover to inspect3 lead changes
Win share is simulated with the production router's Thompson Sampling over the last hour of gpt-4o-mini traffic.
