45%Niche pick?Niche pickVote to see the scoreNeurIPS 2026Johns HopkinsNew YorkNYU Shanghai / NYUTexas A&MJohns HopkinsVision-language modelsReading, Not Thinking: Bridging the Modality Gap When Text Becomes PixelsKaiser Sun, Xiaochuang Yuan, Hongjun Liu, Chen Zhao and 3 moreAtlanta Poster Session 5, Fri, Dec 11, 10:00 AM–1:00 PM, Hall C1 · Published 2026– ReadersNo votes yet0/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 0 of 20 reviewers recommend itlenient 0/5medium 0/10strict 0/5
45%Niche pick?Niche pickVote to see the scoreNeurIPS 2026BloombergStony BrookOptimizationYour Hypergradient is Skewed: Antithetic Neumann Estimation for Bilevel OptimizationJason Bohne, Pawel Polak, Gary Kazantsev, David RosenbergAtlanta Poster Session 3, Thu, Dec 10, 10:00 AM–1:00 PM, Hall C1 · Published 2026– ReadersNo votes yet0/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 0 of 20 reviewers recommend itlenient 0/5medium 0/10strict 0/5
76%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026U CopenhagenCarnegie MellonTexas A&MBloombergAutodeskVideo-language modelsNot Another Text Benchmark: Putting the “Visual" Back in Visual Question Answering for Large Video ModelsThree visual benchmarks for video understanding expose large video model weaknesses when reasoning through visual queries instead of text options.Rwiddhi Chakraborty, Yinong O Wang, Cheng Zhang, Fan Bai and 5 moreParis Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet10/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 10 of 20 reviewers recommend itlenient 4/5medium 5/10strict 1/5
89%Must read?Must readVote to see the scoreNeurIPS 2026YaleU California, BerkeleyBloombergRutgersAgent benchmarks & environmentsPieArena: Ranking and Profiling Language Agents in Realistic Negotiation ScenariosPieArena benchmarks LLM negotiation via multi-agent MBA scenarios, ranking agents with order-invariant payoffs and finding GPT-5 matches trained human baselines while profiling cross-model behavioral heterogeneity.Chris Zhu, Sasha Cui, Will S Dufallo, Runzhi Jin and 3 moreParis Poster Session 6, Fri, Dec 11, 2:30 PM–4:30 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet16/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 16 of 20 reviewers recommend itlenient 5/5medium 7/10strict 4/5