Good Papers

Showing papers from Boson AI - we are hiring! Show all papers

57%Worth a look
?Worth a lookVote to see the score

The Pok\'emon Theorem and other Fairness Impossibility Results

Daniel Matsui Smola, Alexander Smola

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 1/5
83%Must read
?Must readVote to see the score

ProactBench: Beyond What The User Asked For

ProactBench measures LLM conversational proactivity via emergent, critical, and recovery inference across 198 dialogues, finding recovery is hard and poorly predicted by standard benchmarks.

Sepehr Harfi Moridani, Ahmad Salimi, Dongming Shen, Alexander Smola

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

Submodular Benchmark Selection

Submodular maximization selects small benchmark subsets to approximate all others, with mutual information outperforming entropy for small-set imputation across public LLM leaderboards.

Alexander Smola

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 1/5