Good Papers

Showing papers from AIR Show all papers

71%Highly rated
?Highly ratedVote to see the score

ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder

ProCLIP progressively aligns CLIP image encoders with LLM-based embedders via curriculum distillation and contrastive tuning to support long multilingual texts without disrupting pretrained vision-language alignment.

Xiaoxing Hu, Kaicheng Yang, Ziqi Ye, Ziyang Gong and 6 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 11 on Hugging Face · Code ★ 27

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 4/5
medium 3/10
strict 0/5