
Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers
A benchmark of 390 AI papers measures scientific slop across structure, argument, and artifacts; a harness reduces the AI-human gap by 63% via evidence-grounded revision.
Published Sep 30, 2026 · 0 citations · ▲ 55 on Hugging Face · Code ★ 13
Only vote on papers you've read. Sign in with GitHub to vote.


























































