יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

ServeLearnBench: כיצד יכולים Agents לשפר את עצמם מתוך ניסיון שירות?

ServeLearnBench: How Well Can Agents Self-Improve from Serving Experience?
במאמר זה, ServeLearnBench, נבחנה יכולתם של Agents לשפר את עצמם מתוך ניסיון שירות. המחברים פיתחו נתוני זרימה של סביבה המשתנה, ServeLearnBench, כדי לבחון את יכולתם של Agents ללמוד מתוך ניסיון שירות.
תקציר מקורי באנגליתarXiv:2610.07792v1 Announce Type: new Abstract: Large language model agents are increasingly deployed to perform complex tasks in real-world environments. However, the knowledge required for correct behavior in these environments is often implicit, undisclosed, and subject to change over time. Recent continual-learning harnesses seek to address this challenge by enabling agents to improve from serving experience. Yet the effectiveness and limitations of these methods are not yet well characterized. Existing benchmarks provide only partial coverage: some explicitly provide the target knowledge, others assume a static environment, and those that support continual adaptation remain limited in scale and knowledge diversity. To enable systematic evaluation, we formalize an evolving-environment
קרא במקור המקורי