כתבה
arXiv cs.CL ·
ביקורת רב-ממדית לביצועי תרגום זמני
Rubric-Aligned Disentangled Evaluation of Human Simultaneous Interpreting
נחשפה חידוש בביצועי תרגום זמני: חישוב רב-ממדי לביקורת תרגום. המאמר עוסק בפיתוח קורפוס מאומת של 1,101 תרגומים זמניים, ובפיתוח דגם לביצועי תרגום. הדגם משתמש ב-LM ובפקודות סקלריות, ומציג תוצאות טובות יותר מאשר דגם COMET-KIWI.
תקציר מקורי באנגליתarXiv:2609.11131v1 Announce Type: new Abstract: Human simultaneous interpreting (SI) is commonly assessed with analytic rubrics separating meaning transfer, delivery quality, and temporal synchrony, yet no automatic metric is designed for rubric-aligned segment-level SI evaluation. We construct a professionally annotated corpus of 1,101 SI segments with scores for meaning transfer (LQ), delivery quality (EXP), and perceived latency (LAT). We show that structured LLM prompting and scalar supervision collapse rubric dimensions, yielding near-zero correlation with human ratings and strong cross-dimension coupling. To isolate supervision structure under identical backbone capacity, we introduce dual regression heads on a LoRA-adapted COMET-KIWI encoder. On a held-out talk-level test set, the m
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית