Good Papers

Showing papers from The Hong Kong University of Science and Technology (GuangZhou) Show all papers

83%Must read
?Must readVote to see the score

CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models

CodeScaler uses a reward model to scale code LLM training and inference without test cases, improving benchmarks by up to 14.64 points and cutting latency tenfold.

Xiao Zhu, Xinyu Zhou, Boyu Zhu, Hanxu Hu and 4 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 24 on Hugging Face · Code ★ 50

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 2/5