arXiv paper proposes reward models for legal grounding and abstention
A new arXiv preprint introduces reward models designed to train language models to cite retrieved evidence when answering legal questions and to decline when that evidence is inadequate. The authors note that most existing reward models are not built for these high-stakes requirements. The work targets legal applications where unsupported or fabricated answers carry significant risk.