כתבה
arXiv cs.AI ·
קירור הקונטקסט לסוכנים LLM
Curating Always-Loaded Context for LLM Agents: A Capacitated Assortment Model with Censored Feedback
חוקרים פיתחו מודל לקירור הקונטקסט של סוכנים LLM. המודל מאפשר לטעון קונטקסט רלוונטי ולמנוע טעינה מיותרת. החוקרים בדקו את השיטה על קונטקסט אמיתי ומצאו שהיא משפרת את ביצועי המודל.
תקציר מקורי באנגליתarXiv:2610.11007v1 Announce Type: new Abstract: At the start of every session, LLM agents load a fixed context file, such as $\texttt{AGENTS.md}$. Each loaded token in the file is charged again in every later round of the session, and these files can degrade performance as they grow in size. However, in practice, human or automated curators usually grow these files by appending. We formulate context curation as a capacitated assortment problem. Instructions consume tokens under a finite attention capacity; adding an instruction never raises the compliance of the others, while retained instructions incur a per-session setup cost. We prove an upper bound on the optimal file size, regardless of the number of available candidate instructions, and that appending every instruction with positive
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית