2026

12 posts

Why AI gateways should be built in Go: a language audit of every major gateway

OpenRouter Is Down? How to Fail Over to a Backup Provider in 2 Minutes

OpenRouter Rate Limits: Why They Happen and How Multi Provider Fallback Fixes Them

The supply side is exploding: 136 providers, 502 models, and 2,084 client apps in a single…

Why Traditional Latency Routing Fails and How We Fixed It With One Formula

No model stays #1: the leader's share of gateway traffic fell from 29% to 8% in eight mont…

Inside Sakana Fugu Ultra: We Reverse Engineered Its Multi Agent Architecture

How to route LLM requests by cost and latency

What the gateway saw in April 2026: agents live on Anthropic, open-source models got fast,…

Agentic routing, benchmarked: Requesty adds 16ms of overhead, OpenRouter adds 55ms

Designing fallback retries: why Requesty uses 500ms → 4s with jitter

Routing policies 101: fallback, load balancing, and latency in production

2025

17 posts

Case Study: How E-commerce Chatbots Scale to Black Friday Traffic with Requesty

Cross-Provider Caching Deep Dive: Maximize Performance Across Your Stack

LLM Gateway vs Direct API Calls: Benchmarking Latency & Uptime

Rate-Limiting, Retries & 429s: Bullet-Proofing Your AI Pipeline

Smart Routing Demystified: Choosing the Fastest-Cheapest Model per Request

Solving Provider Outages: Real-World Failover War Stories

The Future of LLM Routing: On-device, Edge AI, and Federated Models

Top 7 Smart-Routing Strategies (with YAML/JSON Examples)

Smarter-Than-Human Model Picking: Introducing Requesty Smart Routing

Intelligent LLM Routing in Enterprise AI: Uptime, Cost Efficiency, and Model Selection

Introducing Smart Routing: Smart AI Model Selection!

Supercharging Cline with Requesty: Models, Fallbacks, and Optimizations

Handling LLM Platform Outages: What to Do When OpenAI, Anthropic, DeepSeek, or Others Go D…

Implementing Zero-Downtime LLM Architecture: Beyond Basic Fallbacks

Claude-3-5-Sonnet: Save Over 50% on AI Costs with Cline & Requesty Router

Switching LLM Providers: Why It’s Harder Than It Seems

What is LLM Routing?