Add llm-gateway module: routing, failover below the tool-calling advisor, per-provider circuit breakers, dollar caps and token limits, tenant-keyed cache; real OpenAI and Anthropic models against a local fake
Co-Authored-By: Claude Sonnet 5.5 <[email protected]> Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
@@ -0,0 +1,12 @@
|
||||
# Failover: which statuses move to the next provider
|
||||
|
||||
status second tried? outcome trail
|
||||
408 true answered [first: failed with status 408, second: ok]
|
||||
429 true answered [first: failed with status 429, second: ok]
|
||||
500 true answered [first: failed with status 500, second: ok]
|
||||
503 true answered [first: failed with status 503, second: ok]
|
||||
400 false rethrown [first: rejected with status 400, not retried]
|
||||
401 false rethrown [first: rejected with status 401, not retried]
|
||||
404 false rethrown [first: rejected with status 404, not retried]
|
||||
|
||||
both down: No provider could answer: [a: failed with status 503, b: failed with status 429]
|
||||
Reference in New Issue
Block a user