83%Must read
?Must readVote to see the score

Benchmark for Assessing Olfactory Perception of Large Language Models
The Olfactory Perception benchmark evaluates LLM smell reasoning across 1,010 questions, finding compound names outperform molecular structures and best accuracy reaches 64.4%.
Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026
– ReadersNo votes yet
13/20 AI panelreviewers recommend it
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.
AI panel: 13 of 20 reviewers recommend it
lenient 3/5
medium 7/10
strict 3/5