Good Papers

Showing papers from University of California, San Francisco Show all papers

57%Worth a look
?Worth a lookVote to see the score

Efficient evaluation and error pattern discovery for blackbox AI systems

Maxim Rabinovich, Harvineet Singh, Aman Sinha

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
88%Must read
?Must readVote to see the score

Adaptive auditing of AI systems with anytime-valid guarantees

An adaptive auditing framework using anytime-valid betting tests rigorously evaluates AI failure modes with as few as 20 observations and certifies global robustness upon passing stringent audits.

Siyu Zhou, Patrick Vossler, Venkatesh Sivaraman, Yifan Mai and 1 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 3/5