יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

כלי טולים: האם להסתמך עליהם?

Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools
במחקר זה נבחנה תלותם של כלי טולים במערכות AI. נמצא כי כלי אלה עשויים לחזור תוצאות שגויות, וכי כלי טולים עשויים להיות תלויים יתר על המידה בהם.
תקציר מקורי באנגליתarXiv:2609.05587v1 Announce Type: new Abstract: Existing evaluations of tool-using agents primarily measure whether an agent can successfully complete diverse tasks with tools. These evaluations generally assume that tools return reliable information. However, tool returns in real-world systems can be plausible yet incorrect. We investigate how agents respond to unreliable tool returns by evaluating fourteen LLMs using three tools-web search, LLM sub-agent delegation, and code execution. For each tool, we corrupt its returns and measure whether agents adopt the corrupted content in their final answers. Agents exhibit high levels of overtrust across all three settings: the mean adoption rate exceeds one third for every tool and reaches 68.0% for web search. Analysis of reasoning traces reve
קרא במקור המקורי