← All episodes CacheSlide: Position-Aware KV Cache Reuse for Agent LLMs

CacheSlide: Position-Aware KV Cache Reuse for Agent LLMs

Mar 17, 2026
This episode examines CacheSlide from USENIX FAST26, a system that enables LLMs to reuse cached key-value pairs across shifting prompt positions in agentic workflows. The paper introduces chunked contextual position encoding and priority-based eviction to solve the position mismatch problem that prevents KV cache reuse when prompt segments shift in multi-turn agent conversations.