
Optimal Representation Size: High-Dimensional Analysis of Pretraining and Linear Probing
High-dimensional analysis of pretraining via PCA and linear probing derives exact errors versus representation size, showing compression helps with abundant unlabeled but scarce labeled data.
Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.