כתבה
MarkTechPost ·
Prime Intellect משחרר Verifiers v1
Prime Intellect Releases Verifiers v1: Composable Tasksets, Harnesses, and Runtimes for Agentic RL Training and Evaluations
Prime Intellect השיקה Verifiers v1, כלי לאימון והערכה של למידת חיזוק עם סוכנים. Verifiers v1 מאפשר ריצה של סוכנים עם כלים ותשתיות שונות, ומספק יכולת להרצת ניסויים בקנה מידה גדול.
תקציר מקורי באנגליתPrime Intellect launched verifiers 0.2.0 . It previews a rewritten core, shipped under the new verifiers.v1 namespace. Modern evaluations now run coding agents with tools, compaction, and subagents. Accordingly, v1 rebuilds environments to run these agentic workloads at scale. What is verifiers v1? First, consider what verifiers is: Prime Intellect’s environment stack for agentic reinforcement learning and evaluations. Previously, an environment bundled its data, agent logic, and infrastructure together. In contrast, v1 breaks that bundle into three composable pieces. A taskset defines the work: the data, tools, and scoring. A harness solves the task and produces a rollout. That harness can be a ReAct loop, a CLI agent, or your own. The rollout then runs inside a runtime , either loc
קרא במקור המקורי
marktechpost.com
פתח כתבה מקורית