יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

פטנטים: בדיקת שופטי LLM לסוכני פטנט-דרפטינג מקצועיים

Vibe Patenting: Evaluating LLM Judges for Professional Patent-Drafting Agents
במאמר זה, נחקרה יעילות שופטי LLM בבדיקת סוכני פטנט-דרפטינג. התוצאות הראו ששופטים משפרים את תפוקת הסוכנים, אך ישנם גבולות לשימוש בשופטים.
תקציר מקורי באנגליתarXiv:2609.13422v1 Announce Type: cross Abstract: LLM judges are increasingly used to evaluate and improve AI-generated outputs, yet their reliability for complex professional work remains unclear. We study this problem through Vibe Patenting, an end-to-end patent-drafting testbed for AI-agent evaluation. A separately-invoked LLM judge evaluates generated patent drafts and provides structured feedback for iterative revision. Across multiple inventions and drafting-agent configurations, judge-guided revision consistently improves judge-assessed quality, while unguided revision tends to saturate. Notably, iterative judge feedback enables a low-reasoning agent to approach the performance of a substantially more expensive high-reasoning agent. Stronger models and increased reasoning generally
קרא במקור המקורי