Add multimodal module: receipt images to Java records on the real OpenAI, Anthropic and Ollama models against an OCR-backed local server, validation, repair retry, accuracy by photo condition

Co-Authored-By: Claude Sonnet 5.5 <[email protected]>
Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
Claude
2026-10-09 07:07:57 +00:00
parent b0bba995e6
commit d67ac0630b
31 changed files with 1339 additions and 0 deletions
+13
View File
@@ -0,0 +1,13 @@
# The same receipt image through three real Spring AI chat models (24530 bytes, PNG)
openai image as sent:
[{"text":"Extract this receipt. Your response shou... (1455 characters)","type":"text"},{"image_url":{"url":"data:image/png;base64,<32708 base64 characters>"},"type":"image_url"}]
mime label: image/png | bytes the server decoded: 24530 | sha256 matches: true
anthropic image as sent:
[{"text":"Extract this receipt. Your response shou... (1455 characters)","type":"text"},{"source":{"data":"<32708 base64 characters>","media_type":"image/png","type":"base64"},"type":"image"}]
mime label: image/png | bytes the server decoded: 24530 | sha256 matches: true
ollama image as sent:
{"role":"user","content":"Extract this receipt. Your response shou... (1455 characters)","images":["<32708 base64 characters>"]}
mime label: (none: Ollama sends bare base64) | bytes the server decoded: 24530 | sha256 matches: true
distinct image hashes seen by the server across the three providers: 1