Files
spring-ai/llm-gateway/output/09-cache.txt
T

11 lines
459 B
Plaintext

# Semantic cache mechanics (stand-in embeddings, threshold 0.92)
EmbeddingModel.embed(String) returns: float[]
probe cosine hit at 0.92?
same words, new order 0.985 true
one word different (A18 for A17) 0.970 true
different question, shares a few words 0.395 false
tenant globex asks the stored question: miss
a cache keyed by feature only, tenant globex asks: 30 days, original packaging.