יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

אופטימיזציה של מדיניות Juris לתהליכי היגיון משפטי מובנים

JPO: Juris Policy Optimization for Structured Legal Reasoning in Criminal Judgment Prediction
חוקרים מציגים את Juris Policy Optimization (JPO), שיטה לאופטימיזציה של היגיון משפטי מובנה. JPO משתמשת ברציונלים מורים ולמידת חיזוק עם פרס מרכב. השיטה מראה שיפור באיכות הניבוי וההיגיון בהשוואה לשיטות קודמות.
תקציר מקורי באנגליתarXiv:2608.29616v2 Announce Type: replace Abstract: Criminal judgment prediction requires models to infer statutory articles, charges, and sentencing outcomes from case facts. Unlike standard classification tasks, it involves a structured reasoning process in which statutes should be matched with facts, charges should be justified by statutes, and sentencing outcomes should remain consistent with charges. Existing approaches optimize final labels, and while some have attempted to evaluate reasoning quality, their evaluations are indirect, often relying on LLM-generated rubrics that reflect model-internal preferences rather than the inherent logical structure of legal adjudication. We propose Juris Policy Optimization (JPO), a post-training framework for structured legal reasoning in Chines
קרא במקור המקורי