45%Niche pick?Niche pickVote to see the scoreNeurIPS 2026McGill University/MilaMcGill University, Mila, Paris-SMcGill University & Quebec AILLS and MILA | CNRS Paris-SaclMila / McGillLLM pretraining & scaling lawsFrom Infrastructure to Interface, the AI Value Chain Drives LLM HomogenizationKhaoula Chehbouni, Cléa Chataigner, Prakhar Ganesh, Pablo Piantanida and 2 moreSydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet0/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 0 of 20 reviewers recommend itlenient 0/5medium 0/10strict 0/5
57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026McGill University, Mila, Paris-SILLS and MILA | CNRS Paris-SaclMcGill University / MilaLLM evaluation & benchmarksAuditing is not Evaluating: LLM Audit Requires Dynamic, Contextual, Budget-aware and Reliable EvidenceCléa Chataigner, Pablo Piantanida, Golnoosh FarnadiParis Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026SpotlightMILA / McGillEPFLETHZ - ETH ZurichApertus ETH Zürich McGill University / MilaVision-language modelsScaling Laws for Multimodal Data MixturesAditi Khandelwal, Ayush Kumar Tarun, Yixuan Xu, Imanol Schlag and 4 moreSydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
88%Must read?Must readVote to see the scoreNeurIPS 2026Mila and McGillMcGill University / MilaMila / McGillElement AI, a ServiceNow companyLLM evaluation & benchmarksForecasting Downstream Performance of LLMs With Proxy MetricsAggregating token-level statistics over expert solutions yields proxy metrics that outperform loss-based baselines for model selection, data selection, and training-time forecasting.Arkil Patel, Siva Reddy, Marius Mosbach, Dzmitry BahdanauSydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 12 on Hugging Face · Code ★ 11– ReadersNo votes yet15/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 15 of 20 reviewers recommend itlenient 5/5medium 8/10strict 2/5
76%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026McGill University / MILAGoogle DeepmindConcordiaÉcole de technologie supérieure,McGill University / MilaLLM evaluation & benchmarksIDEAFix: Evaluation Framework for Creative Defixation Prompting in LLMsIDEAFix evaluates LLM divergent thinking via controlled design scenarios and defixation prompts, showing task formulation and simple prompting boost originality but homogenization persists.Florian Carichon, Soumya Sharma, Meaghan J. Girard, Romain Rampa and 1 moreParis Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet10/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 10 of 20 reviewers recommend itlenient 4/5medium 6/10strict 0/5