Good Papers

Showing Self-supervised vision Show all papers

45%Niche pick
?Niche pickVote to see the score

Shaping Useful Noise: Energy Distributions Predict Visual Pretraining Quality

Ching Lam Choi, Antonio Torralba, Phillip Isola, Stefanie Jegelka

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

HIFC-IQA: Train-Free Cross-Domain Image Quality Assessment via Dual-Process Cognition

Yu Li, Zhengran Shen, Puchao Zhou, Yachun Mi and 4 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Complementing DINO Features with Image Structure for Part Discovery

Samyak Rawlekar, Nikhil C Paleti, Amey Gupta, Narendra Ahuja

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

SemGeo-Gen: Unsupervised Generation of Approximate Cross-Instance Semantic-Geometric Correspondences

Roy Amoyal, Shira Ifergane, Oren Freifeld

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

SICAF: Time-Varying Focus Bottleneck for Self-Supervised Event-Based Optical Flow with Spiking Neural Network

Shuangming Yang, Qing He, Shangqi Guo, Badong Chen

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Your Self-Supervised Projection Head Captures Object Co-Occurrence Statistics

Arthur Aubret, Jochen Triesch, Céline Teulière

Paris Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 1/5
45%Niche pick
?Niche pickVote to see the score

Data Density Scaling Laws for Image Self-Distillation

Alvard Barseghyan, Ani V Vanyan, Hakob Tamazyan, Hrant Khachatrian

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Seeing the World through Any Eyes

Yang Fu, Jianqin Wang, Xiangtai Li, Henghui Ding

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

DynaProto: Dynamic Prototypical Contrast for Temporally Consistent Object-Centric Learning

Tianran Ouyang, Peiqin Xu, Xingping Dong, Kaihao Zhang and 1 more

Paris Poster Session 3, Thu, Dec 10, 12:30 PM–2:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
88%Must read
?Must readVote to see the score

Revisiting Cross-View Completion: Self-Supervised Pre-Training via Reconstruction Error Comparison

Gekko uses relative reconstruction error between cross-view and masked-autoencoder predictions as a self-supervised co-visibility proxy, adding binocular training signals that consistently outperform CroCo on 3D vision tasks while training directly from raw video.

Thibaut Loiseau, Guillaume Bourmaud, Vincent Lepetit

Paris Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 3/5
medium 10/10
strict 2/5
80%Must read
?Must readVote to see the score

FoMo: Forking Moment in Generative Trajectory as a Perceptual Distance

Diffusion trajectory forking moments generate automatic perceptual distance labels that train reference-based image quality assessment metrics outperforming human-annotated datasets.

Jaihyun Lew, Mingi Jung, Minjun Park, Wooseok Song and 1 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 24 on Hugging Face · Code ★ 11

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

Diffusion Masked Pretraining for Dynamic Point Cloud

DiMP applies diffusion modeling to masked tube-center inference and inter-frame motion prediction, eliminating positional leakage and deterministic trajectory collapse to improve dynamic point cloud pretraining.

Zhuoyue Zhang, Yiding Sun, Chaowei Fang, Haozhe Cheng and 3 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 4/5
medium 5/10
strict 0/5
80%Must read
?Must readVote to see the score

RATS! Patches Talk Through Registers: Emergent Parts in Register Attention Transformers

RATS decomposes vision transformers' classification token into learnable register tokens that spontaneously specialize into object parts, improving segmentation by up to 12 mIoU.

Timing Yang, Predrag Neskovic, Jansen Seheult, Wenchao Han and 3 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 1/5
89%Must read
?Must readVote to see the score

I Have a Stream: Making Self-Supervised Learning Work on Continuous Video

Self-supervised video-stream pretraining fails due to intra-batch near-duplicate frames, but proposed StreamMAE with motion-biased crops matches i.i.d. MAE and scales to 95 hours.

Ivan Martinović, Lukas Knobel, Yuki Asano

Paris Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026 · ▲ 14 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 4/5
medium 9/10
strict 3/5
72%Highly rated
?Highly ratedVote to see the score

Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry

A self-supervised encoder learns subject-specific fMRI embeddings from repeated brain responses, and unsupervised orthogonal rotations align them across subjects into a shared geometry, demonstrating approximately isometric cross-subject visual representations.

Pablo Marcos Manchón, Rishi Jha, Lluís Fuentemilla

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 2 on Hugging Face · Code ★ 2

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 4/5
medium 3/10
strict 1/5
83%Must read
?Must readVote to see the score

SPHERE-JEPA: Spherical Prediction with Homogeneous Embeddings

SPHERE-JEPA proves hyperspherical uniformity minimizes downstream prediction risk on manifolds and improves SSL retrieval and ImageNet linear probing over LeJEPA.

Léo Nicollier, Max Dunitz, Marc Pic, Pablo Muse and 2 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 3/5
medium 9/10
strict 1/5
89%Must read
?Must readVote to see the score

CanViT: Toward Active-Vision Foundation Models

CanViT introduces the first active-vision foundation model with a retinotopic backbone and scene-wide canvas, achieving 38.5% ADE20K mIoU with one glimpse and 84.5% ImageNet accuracy.

Yohaï-Eliel BERREBY, Sabrina Du, Audrey Durand, B. S Krishna

Paris Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026 · ▲ 13 on Hugging Face · Code ★ 23

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 4/5
medium 9/10
strict 3/5