Add multimodal module: receipt images to Java records on the real OpenAI, Anthropic and Ollama models against an OCR-backed local server, validation, repair retry, accuracy by photo condition
Co-Authored-By: Claude Sonnet 5.5 <[email protected]> Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
@@ -0,0 +1,13 @@
|
||||
# The same receipt image through three real Spring AI chat models (24530 bytes, PNG)
|
||||
|
||||
openai image as sent:
|
||||
[{"text":"Extract this receipt. Your response shou... (1455 characters)","type":"text"},{"image_url":{"url":"data:image/png;base64,<32708 base64 characters>"},"type":"image_url"}]
|
||||
mime label: image/png | bytes the server decoded: 24530 | sha256 matches: true
|
||||
anthropic image as sent:
|
||||
[{"text":"Extract this receipt. Your response shou... (1455 characters)","type":"text"},{"source":{"data":"<32708 base64 characters>","media_type":"image/png","type":"base64"},"type":"image"}]
|
||||
mime label: image/png | bytes the server decoded: 24530 | sha256 matches: true
|
||||
ollama image as sent:
|
||||
{"role":"user","content":"Extract this receipt. Your response shou... (1455 characters)","images":["<32708 base64 characters>"]}
|
||||
mime label: (none: Ollama sends bare base64) | bytes the server decoded: 24530 | sha256 matches: true
|
||||
|
||||
distinct image hashes seen by the server across the three providers: 1
|
||||
Reference in New Issue
Block a user