Add llm-gateway module: routing, failover below the tool-calling advisor, per-provider circuit breakers, dollar caps and token limits, tenant-keyed cache; real OpenAI and Anthropic models against a local fake

Co-Authored-By: Claude Sonnet 5.5 <[email protected]>
Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
Claude
2026-10-09 10:24:57 +00:00
parent 80cd21f89b
commit cdee85d3f4
51 changed files with 2404 additions and 0 deletions
+11
View File
@@ -0,0 +1,11 @@
# Routing: hints, per-target options, unknown hints
hint smart -> answered by claude (claude-model)
model id sent to openai: openai-model
model id sent to claude: claude-model
maxTokens sent to both: 256 and 256
hint local -> answered by local; cloud calls made for it: 0
old registry, hint "locla" (typo for local): goes to claude, a cloud provider
this Routes, hint "locla": Unknown model hint 'locla'. Known hints: [local, smart]