57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026TechnionCrusoe / TechnionTechnion - Israel Institute of THabana LabsEfficient attention & state-space modelsMore Value per Key: Asymmetric Sparse Attention for Faster LLM DecodingNoam Elata, Itay Lamprecht, Mikey Shechter, Daniel Ohayon and 2 moreParis Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
78%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026Technion - Israel Institute of TCrusoe / TechnionHabana LabsTechnionEfficient attention & state-space modelsBlock Sparse Flash AttentionBlock Sparse Flash Attention accelerates long-context inference by computing exact similarities to select top-k value blocks, skipping ~50% of computation for up to 1.38x kernel and 1.24x end-to-end speedups with minimal accuracy loss.Daniel Ohayon, Itay Lamprecht, Itay Hubara, Israel Cohen and 2 moreSydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet11/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 11 of 20 reviewers recommend itlenient 5/5medium 5/10strict 1/5