כתבה
arXiv cs.AI ·
AgentFly: Scaling Agentic Reinforcement Learning with Unified Resource System
תקציר מקורי באנגליתarXiv:2507.14897v2 Announce Type: replace Abstract: Methods to build LLM agents have evolved from prompt engineering and supervised finetuning to agentic reinforcement learning (agentic RL). However, agentic RL remains bottlenecked by its surrounding systems: agents must interact with heterogeneous environments, such as sandboxes, model services, and external APIs. Their allocation, reuse, and lifecycle dominate rollout cost and cap the scale at which training becomes practical. In this work, we present AgentFly, an agentic RL framework built with a unified resource layer that treats each of these environments as a distinct, typed resource scheduled through one engine, with per-tool acquisition for multi-turn reuse, asynchronous backpressure, and rollout versus global-scoped lifecycles. Ag
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית