כתבה
arXiv cs.AI ·
מה ישראליות במשימות הכלליות של הכללה? ניתוח מרכזי של הסיבות
What Do Current Systematic Generalization Tasks Miss? A Reasoning-Centered Analysis
חידוש: חוקרים גילו חסרונות במשימות הכלליות של הכללה. המחקר, שפורסם ב-arXiv, חושף חסרונות במשימות הכלליות של הכללה, כולל חסרונות במשימות הכלליות של הכללה.
תקציר מקורי באנגליתarXiv:2609.19212v2 Announce Type: replace Abstract: Systematic generalization, the ability to solve novel problems by recombining known atomic elements, is central to human intelligence but difficult to study rigorously under controlled settings. Existing studies therefore rely on simplifications such as elemental composition, productivity-based tests, and action-explicit goals, which make systematic generalization easier to study but omit some essential aspects of this capability. To characterize what these simplifications miss, we adopt a reasoning-centered lens and introduce TranSGrid, a testbed that brings deductive, inductive, and abductive reasoning together within a unified task. Experiments with seven Transformer models on 4,800 TranSGrid instances show that all models perform much
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית