כתבה
arXiv cs.CL ·
LexReward: פרקטיקה חדשה לעיצוב תגמולים למודלי שפה חוקי
LexReward: A Taxonomy-Driven Reward Framework for Legal Language Models
מאמר חדש מציג פרקטיקה חדשה לעיצוב תגמולים למודלי שפה חוקי, שמספקת תיאור עשיר יותר של תכונות התגובות החוקיות. הפרקטיקה, שנקראת LexReward, מאפשרת למודלי שפה להתאים את תגובותיהם לתכונות חוקיות שונות, כגון סגנון, תוכן וסדר.
תקציר מקורי באנגליתarXiv:2609.39071v1 Announce Type: new Abstract: Legal language models require reward signals that capture not only answer correctness but also the multidimensional quality of legal responses. Existing reward methods, however, often rely on coarse-grained holistic judgments, providing limited domain specificity and interpretability. We introduce LexReward, a taxonomy-driven framework for legal reward modeling. LexReward characterizes legal response quality along three complementary dimensions: Style, covering lexical and syntactic quality; Element, assessing legal subjects, facts, statutes, and decisions; and Chain, evaluating the order, completeness, correctness, and non-redundancy of legal reasoning. For each dimension, we develop rubrics that specify evaluation criteria and quality level
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית