כתבה
arXiv cs.AI ·
CRAX: מבחן רפורמינציה מהיר ובטוח
CRAX: Fast Safe Reinforcement Learning Benchmarking
CRAX הוא מבחן רפורמינציה מהיר ובטוח שמאפשר ניסויים גדולי-ממד ופרוטוטיפ תקין. המבחן כולל שמונה משימות שונות ומציג פתרונות חדשים לבעיות בטיחותיות.
תקציר מקורי באנגליתarXiv:2606.20376v4 Announce Type: replace-cross Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While benchmarks have been central to progress in RL, existing 3D physics-based safety benchmarks remain computationally slow, limiting large-scale experimentation and rapid prototyping. To address this gap, we propose CRAX (Constrained RL Accelerated with JAX). Built on top of the MuJoCo XLA (MJX) physics engine, CRAX leverages vectorized operations and hardware acceleration, yielding up to 200x faster training over comparable CPU-based safety benchmarks. The benchmark features eight tasks spanning three difficulty levels and multiple agent morphologies. Evaluating seven popular safe RL methods, w
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית