
LexReward: A Taxonomy-Driven Reward Framework for Legal Language Models
LexReward introduces taxonomy-driven rubric-based rewards for legal language models across style, element, and reasoning dimensions, improving DPO and reinforcement learning performance.
Published Sep 30, 2026 · 0 citations · ▲ 51 on Hugging Face
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.






