78%Highly rated
?Highly ratedVote to see the score

PCEval: A Benchmark for Evaluating Physical Computing Capabilities of Large Language Models
PCEval introduces an automatic benchmark revealing LLMs generate physical computing code well but fail at breadboard layouts and pin connections.
Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026
– ReadersNo votes yet
11/20 AI panelreviewers recommend it
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.
AI panel: 11 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 1/5