Stop calling a dead API. Shed load fast, recover automatically, and stay consistent across restarts with Redis-backed failure state.


Why this matters
Every LLM-powered application depends on an external provider - OpenAI, Anthropic, Google, or a self-hosted model. These providers go down. Rate limits spike. Latency balloons. Without a circuit...