ended7월 26일· 1 sources
Kmemo: a semantic cache for LLM calls that refuses to serve you the wrong answer
Why it matters
An exact-match cache misses "how do I reverse a list in Python" when it has already answered "python list reverse". A semantic cache doesn't: it embeds the prompt, finds the closest one it has seen, a...
1
Sources
+0
24h
—
Growth
7d
Active