כתבה
arXiv cs.AI ·
אין בודק חינם: סקירה של מאמתים למדיניות רובוטים
No Free Checker: A Survey of Verifiers for Robot Policies
סקירה של כ-150 מאמתים למדיניות רובוטים, המשווים בין שני מאפיינים: זמינות ואמינות. המחקר מצא כי אמינות יורדת ככל שזמינות עולה.
תקציר מקורי באנגליתarXiv:2609.09250v1 Announce Type: cross Abstract: A verifier for robot policies reads a candidate behavior and returns a score for how well it did, used both to evaluate vision-language-action policies and to train them. Verifiers range from success detectors and reward models to runtime monitors, safety filters, and temporal-logic specifications. We survey roughly 150 verifiers and compare them along two properties. Availability is how much a verdict costs, how early in a rollout the verdict arrives, and how often a verdict can be asked for. Availability rises as verdicts get cheaper, earlier, and denser. Credibility is how much a high score tells us about the task. Credibility falls as the judgment becomes gameable and self-serving. We group the verifiers by who supplies the judgment: hu
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית