כתבה
arXiv cs.LG ·
ReHoPER: תכנון-לקראת-הגבול לשיפור תפיסה
ReHoPER: Receding-Horizon Planning for Enhanced Reasoning
ReHoPER משפר את תפיסת המודלים הגדולים של השפה הטבעית על ידי יצירת שאלות ביניים.
תקציר מקורי באנגליתarXiv:2610.00940v1 Announce Type: cross Abstract: We propose ReHoPER, an inference-only, zero-shot method that improves large language models' reasoning by generating and answering intermediate questions along multiple paths before the final answer. It iteratively plans a horizon of candidate intermediate questions, selects one to answer, and replans from the updated history. ReHoPER is task-agnostic, using the same generic instructions across datasets and models without labeled data or task-specific prompt design. Across multiple datasets, including iLLC, a new controlled benchmark for compositional reasoning, ReHoPER outperforms strong baselines, with the largest gains in the most compositional settings. Our implementation and the iLLC generator are publicly available to support future w
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית