Add providers module: one app on the real OpenAI, Anthropic and Gemini Spring AI models against a local server in three wire formats; options, prompt caching, cost from price sheets, failover with retry layers measured
Co-Authored-By: Claude Sonnet 5.5 <[email protected]> Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
@@ -0,0 +1,20 @@
|
||||
# One TicketService.raw() call, three providers, as the HTTP server saw it
|
||||
|
||||
OPENAI POST /v1/chat/completions
|
||||
top-level keys : messages, model
|
||||
model sent : gpt-6.1-sol
|
||||
system prompt : messages[0], role "system"
|
||||
sampling sent : nothing (the vendor's own defaults apply)
|
||||
|
||||
ANTHROPIC POST /v1/messages
|
||||
top-level keys : max_tokens, messages, model, system
|
||||
model sent : claude-sonnet-5-5
|
||||
system prompt : top-level "system" (a string)
|
||||
sampling sent : max_tokens=1024
|
||||
|
||||
GEMINI POST /v1beta/models/gemini-3.8-flash:generateContent
|
||||
top-level keys : contents, systemInstruction, generationConfig
|
||||
model sent : gemini-3.8-flash
|
||||
system prompt : top-level "systemInstruction".parts[0]
|
||||
sampling sent : temperature=0.7, topP=1.0
|
||||
|
||||
Reference in New Issue
Block a user