Expand description
Semantic output deduplication cache with TTL and pluggable eviction policies.
Structs§
- Cache
Key - Key identifying a cached prompt/model/temperature combination.
- Cache
Stats - Statistics snapshot returned by
OutputCache::stats. - Cached
Output - A single cached LLM response entry.
- Output
Cache - In-memory output deduplication cache.
Enums§
- Eviction
Policy - Policy used to select which entry to evict when the cache is full.
Functions§
- cache_
key - Build a
CacheKeyfrom a prompt string, model name, and temperature. - fnv1a_
hash - FNV-1a 64-bit hash of
data.