57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026U Illinois at ChicagoNorthwesternIntelU Illinois, ChicagoVision-language modelsVLS: A Vision-Language-Shape Model for Open-Vocabulary Partonomic 3D ReconstructionXiaoqian Ruan, Pei Yu, Dian Jia, Hyeonjeong Park and 2 moreAtlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
67%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026U Illinois ChicagoU Illinois at ChicagoIllinois Institute of TechnologyU WashingtonState University of New York at Agent benchmarks & environmentsJARVIS-Bench: Benchmarking Personal Intelligence Agents on Long-Horizon Real-User Daily TracesWeizhi Zhang, Wei-Chieh Huang, Yueqing Liang, Liwei Jiang and 36 moreAtlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026– ReadersNo votes yet2/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 2 of 20 reviewers recommend itlenient 2/5medium 0/10strict 0/5
86%Must read?Must readVote to see the scoreNeurIPS 2026U Illinois, ChicagoSalesforce AI ResearchSalesforce ResearchSalesForce.comRecursive SuperintelligenceLLM evaluation & benchmarksBehaviorBench: Modeling Real-World User Decisions from Behavioral TracesBehaviorBench evaluates personalized decision modeling using real-world wallet traces across belief and trade prediction tasks, showing personalization improves beliefs more than trades and reveals model failure modes.Liangwei Yang, Jielin Qiu, Zixiang Chen, Ming Zhu and 8 moreAtlanta Poster Session 5, Fri, Dec 11, 10:00 AM–1:00 PM, Hall C1 · Published 2026– ReadersNo votes yet14/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 14 of 20 reviewers recommend itlenient 5/5medium 8/10strict 1/5