יום חמישי, 8 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

Hallucination Self-Play

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator
Hallucination Self-Play הוא כלי חדש לשיפור מודלי LLM. הוא מאפשר למנגנון הגילוי ללמוד ממנגנון יצירה משופר. המחקר הראה כי השיטה יכולה לשפר מודלים קטנים לרמה של מודלים מתקדמים.
תקציר מקורי באנגליתarXiv:2607.07993v2 Announce Type: replace Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work relies on advanced LLMs to synthesize training data, including rationales, labels, and hallucinated claims. However, these methods treat the generator as a static component, limiting iterative improvement of the detector. To address this limitation, we introduce Hallucination Self-Play (HSP), a novel framework that enables the detector to bootstrap with an evolved generator. HSP involves two roles initialized from the same base model, a detector that assesses the faithfulness of model outputs, and a generator that produces increasingly hard-to-detect hallucinated responses. Specifically, the
קרא במקור המקורי