
PCEval: A Benchmark for Evaluating Physical Computing Capabilities of Large Language Models
PCEval introduces an automatic benchmark revealing LLMs generate physical computing code well but fail at breadboard layouts and pin connections.
Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.