SGD-KV: Summarization Guided KV Cache CompressionPaper SGD-KV: Finding ‘heads that are good at summarizing’ cuts the 1M-token KV cache by up to 75% TL;DR — Attention heads do … September 7, 2026 13 min2609.03235v1