יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

מעבר לחיקוי: תשתית וביצועים לביקורת עמיתים בסיוע LLM

Beyond Imitation: A Framework and Benchmark for LLM-Assisted Peer Review
חוקרים הציגו תשתית לביקורת עמיתים בסיוע מודלי LLM. התשתית מתמקדת בגילוי שגיאות ומשפרת את יעילות הביקורת. המחקר הראה תוצאות מבטיחות, אך גם הדגיש את הצורך בעמידות בפני תקיפות אדברסריות.
תקציר מקורי באנגליתarXiv:2610.11087v1 Announce Type: new Abstract: The rapid growth of scientific publishing has strained peer review, particularly in machine learning, raising concerns about declining review quality and increasing reviewer workload. Large language models (LLMs) have been proposed as automated review assistants, yet their evaluation has focused largely on imitating human-written reviews rather than supporting the core functions of peer review. Here, we introduce a verification-centric perspective on LLM-assisted peer review, emphasizing error detection as a critical and resource-intensive task. We present a scalable benchmark that evaluates review systems' ability to identify logical contradictions, constructed through synthetic insertion of errors into conference papers, yielding unambiguou
קרא במקור המקורי