Add llm-gateway module: routing, failover below the tool-calling advisor, per-provider circuit breakers, dollar caps and token limits, tenant-keyed cache; real OpenAI and Anthropic models against a local fake
Co-Authored-By: Claude Sonnet 5.5 <[email protected]> Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
@@ -0,0 +1,11 @@
|
||||
# Circuit breakers: when they open and how they recover
|
||||
|
||||
old yml (window 10, threshold 50%): minimumNumberOfCalls is 100; the first request that skips openai is #11
|
||||
this module (window 10, minimum 5, threshold 50%): the first request that skips openai is #6
|
||||
|
||||
after 8 requests: openai breaker OPEN, claude breaker CLOSED
|
||||
openai was contacted 5 times of 8; claude answered 8 times
|
||||
|
||||
after the wait, one probe goes to openai (it has recovered): breaker CLOSED, openai calls 1
|
||||
|
||||
20 requests rejected with 400: breaker CLOSED, calls the breaker recorded: 0
|
||||
Reference in New Issue
Block a user