Good Papers

Showing papers from Max Planck Institute for Software Systems Show all papers

45%Niche pick
?Niche pickVote to see the score

Youdunit: Single-Call Counterfactual Necessity in Multi-Agent LLM Systems

Marissa Li, Stephanie Gao, Kenny Guo, Xingjian Li and 2 more

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
69%Highly rated
?Highly ratedVote to see the score

Can an LLM Reason Like a Lawyer? Benchmarking the ability of LLMs to map the facts of a case to the elements of the applicable legal rule

Shounak Paul, Seungeon Lee, Christoph Engel, Krishna Gummadi

Paris Poster Session 1, Wed, Dec 9, 12:30 PM–2:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
3/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 3 of 20 reviewers recommend it
lenient 2/5
medium 1/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

GeoX acquires geospatial reasoning via self-play with executable programs and verifiable rewards, improving base VLMs up to 5.5 points without large-scale human-curated data.

Kyeongjin Ahn, Seungeon Lee, Krishna Gummadi, Meeyoung Cha

Paris Poster Session 6, Fri, Dec 11, 2:30 PM–4:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 0/5