ActKV: Efficient LLM Agents through Action-Guided KV Cache Management

arXiv:2609.31395 · cs.OS, cs.AI · Submitted 2026-09-25 · Read on arXiv

cs.OS, cs.AI

Submitted: 2026-09-25

Updated: 2026-09-25

Code: https://github.com/Dao-AILab/flash-attention

Terminology

Sources

Related papers