Good Papers

Showing papers from Google UNSW Show all papers

67%Highly rated
?Highly ratedVote to see the score

JARVIS-Bench: Benchmarking Personal Intelligence Agents on Long-Horizon Real-User Daily Traces

Weizhi Zhang, Wei-Chieh Huang, Yueqing Liang, Liwei Jiang and 36 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
2/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 2 of 20 reviewers recommend it
lenient 2/5
medium 0/10
strict 0/5
89%Must read
?Must readVote to see the score

GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring

GlucoFM decomposes CGM data into dual slow and short-term streams for pretraining, improving linear-probe phenotype classification and postprandial response prediction over prior models.

Zechen Li, Keerthana Natarajan, Weizhi Zhang, Simon Lee and 10 more

Atlanta Poster Session 4, Thu, Dec 10, 4:30 PM–7:30 PM, Hall C1 · Published 2026 · ▲ 8 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 3/5
83%Must read
?Must readVote to see the score

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs

MAGE externalizes self-knowledge into co-evolutionary knowledge graphs that guide frozen-learner agents, achieving strong multi-benchmark gains via complementary success and correction memories.

Ruiyi Yang, Zechen Li, Hao Xue, Imran Razzak and 1 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 1/5