יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

CoEvolve: קירוב ויזואלי עם שיפור מצב דו-כיווני

CoEvolve: Construct-to-Edit Visual Grounding with Bidirectional State Refinement
CoEvolve הוא כלי לקירוב ויזואלי שמאפשר שיפור מצב דו-כיווני. הוא משלב רכיבי ריפוי ושיפור מצב כדי לשפר את דיוק הקירוב. CoEvolve מתאים ליישומים שונים, כולל קירוב בתמונות טבעיות ובניתוח תמונות לוויין.
תקציר מקורי באנגליתarXiv:2610.01710v1 Announce Type: new Abstract: Visual grounding localizes an object described by language with a bounding box. Most multimodal grounding models compress target identification, spatial reasoning, and boundary estimation into one terminal prediction. Free-form rationales make reasoning linguistically explicit but do not necessarily expose measurable, editable spatial states. Intermediate localization errors are therefore difficult to diagnose and correct, allowing incorrect region choices and imprecise boundaries to persist in the final box. We introduce CoEvolve, a construct-to-edit framework that separates grounding into explicit state construction and state editing. Region-Evolution Reinforcement (RER) organizes grounding analysis into a progressive semantic--spatial traj
קרא במקור המקורי