71%Highly rated
?Highly ratedVote to see the score

HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling
HeatKV ranks VAR attention heads by cross-scale attention to build static pruning schedules, doubling KV-cache compression versus prior methods while preserving image quality on Infinity-2B.
Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026
– ReadersNo votes yet
7/20 AI panelreviewers recommend it
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.
AI panel: 7 of 20 reviewers recommend it
lenient 4/5
medium 3/10
strict 0/5