יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

בדיקת החלפת זהויות בצינורות מודלי שפה מבוססי קרקע

Does the Selected Object Reach the Reader? Auditing Identity Handoffs in Grounded Language-Model Pipelines
חוקרים בדיקת החלפת זהויות בצינורות מודלי שפה מבוססי קרקע. המחקר בודק 600 שאילתות HybridQA ומוצא כי שיטות שונות לאיתור עצמים יכולות להחמיץ את העצם הנבחר. החוקרים משחררים כלי בשם Returned-Object Profile (ROP) לבדיקת החלפת זהויות.
תקציר מקורי באנגליתarXiv:2609.04579v1 Announce Type: cross Abstract: Grounded language-model pipelines can be divided into three stages: selecting an object, retrieving passages for it, and using that evidence to answer. If the selected object must reach the reader, losing it breaks the handoff. Benchmark recall checks the dataset-linked object, which can differ. We audit 600 HybridQA questions across three selector families. On 1,463 resolvable records where the selected object matches the dataset-traced passage, exact key lookup and exact title matching return the object every time. With every ranked rule given the same decoded selected title, body-only BM25 omits it on 389 records (26.6%) at cutoff five, while hybrid retrieval with reranking omits it on 14 (1.0%). The two identities differ on 329 of 1,792
קרא במקור המקורי