Good Papers

Showing papers from Arm Show all papers

71%Highly rated
?Highly ratedVote to see the score

HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling

HeatKV ranks VAR attention heads by cross-scale attention to build static pruning schedules, doubling KV-cache compression versus prior methods while preserving image quality on Infinity-2B.

Jonathan Cederlund, Axel Berg, Durmus Alp Emre Acar, Chuteng Zhou and 1 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 4/5
medium 3/10
strict 0/5