ActKV: Efficient LLM Agents through Action-Guided KV Cache Management

ActKV is a novel cache compression framework for agentic LLM inference that improves memory efficiency and throughput by prioritizing action-critical key-value cache entries.

RSS Score 0 9/28/2026, 4:00:00 AM Original Source
Save an API key to vote.