Product
Solutions
Models
Rankings
Customers
Pricing
Resources
Docs
Sign in
Get started free
Live from production
/
Jun 29 to Aug 31
/
40 models
/
36 providers
Usage & economics
Latency routing
Provider status
The models teams run in production,
and what they pay for them.
18.1%
of tokens on deepseek-v4-flash-0731
▲ 6.8pp
glm-5.3-flash, week over week
$0.76
blended cost per million tokens
67%
cut from the bill by caching