יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

התאמה דינמית של רובריקה לתיאור תמונות מפורט

Momentum-Coupled Rubric Adaptation for Detailed Image Captioning
MoCo Rubric הוא שיטה חדשה לתיאור תמונות מפורט, המשתמשת בלמידת חיזוק ורובריקה דינמית. השיטה משפרת את איכות התיאורים על ידי קואורדינציה בין המדיניות, הרובריקה והשופט. היא מגיעה לתוצאות טובות בחמישה מבחני תיאור תמונות.
תקציר מקורי באנגליתarXiv:2609.36893v1 Announce Type: new Abstract: Detailed image captioning requires accurate and comprehensive descriptions of fine-grained visual content, yet caption quality spans factual accuracy, information coverage, and clarity. Compared with conventional methods that rely mainly on high-quality supervision or holistic rewards, rubric-based reinforcement learning decomposes these requirements into explicit criteria and provides targeted, structured feedback. However, existing methods often use separate models for caption generation, rubric construction, and judging, which may lead to inconsistent interpretations across roles. Some dynamic rubric methods alternate updates between the caption policy and rubric generator while keeping the judge fixed, but staged optimization may still le
קרא במקור המקורי