Good Papers

MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning

MAPLE uses influence-based pseudo-labeling to adaptively select many-shot ICL demonstrations, boosting LLM performance without extensive labeling costs.

Zihan Chen, Song Wang, Zhen Tan, Jundong Li, Cong Shen

Published May 22, 2025arXiv ↗

70%
OverallHighly rated
?
OverallHighly ratedVote to see the scoreThe exact score shows once you've voted, so every vote is your own call. The first half of each home page shelf shows its scores.
Readers
–

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel5/20reviewers recommend it
lenient 4/5
medium 1/10
strict 0/5
AI panel?Vote to see what the 20 AI reviewers said
Panel consensus
MAPLE's adaptive pseudo-labeling delivers compelling many-shot ICL gains from unlabeled data without heavy labeling costs, though it remains unclear whether influence selection outperforms retrieval, pseudo-label errors are quantified, or computation overhead erodes savings under small pools…

Abstract

In-Context Learning (ICL) empowers Large Language Models (LLMs) to tackle diverse tasks by incorporating multiple input-output examples, known as demonstrations, into the input of LLMs. More recently, advancements in the expanded context windows of LLMs have led to many-shot ICL, which uses hundreds of demonstrations and outperforms few-shot ICL, which relies on fewer examples. However, this approach is often hindered by the high cost of obtaining large amounts of labeled data. To address this challenge, we propose Many-Shot Adaptive Pseudo-LabEling, namely MAPLE, a novel influence-based many-shot ICL framework that utilizes pseudo-labeled samples to compensate for the lack of label information. We first identify a subset of impactful unlabeled samples and perform pseudo-labeling on them by querying LLMs. These pseudo-labeled samples are then adaptively selected and tailored to each test query as input to improve the performance of many-shot ICL, without significant labeling costs. Extensive experiments on real-world datasets demonstrate the effectiveness of our framework, showcasing its ability to enhance LLM adaptability and performance with limited labeled data.