ActKV: Efficient LLM Agents through Action-Guided KV Cache Management
ActKV is a novel cache compression framework for agentic LLM inference that improves memory efficiency and throughput by prioritizing action-critical key-value cache entries.
Save an API key to vote.