יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

STARS: דינמיקה מרחבית-זמנית לייצוגים חברתיים

STARS: From Spatiotemporal Dynamics to Social Representations in Human-Robot Interaction
STARS הוא בנך' להבנת סצנות ניווט חברתיות. הוא בודק את היכולת של מודלים ויז'ואלי-לשוניים להבין סצנות ניווט חברתיות מורכבות. הבנך' כולל מאגר נתונים ומסגרת להערכת מודלים.
תקציר מקורי באנגליתarXiv:2609.40245v2 Announce Type: cross Abstract: Robot navigation in dynamic, human-centered environments requires socially-compliant decisions grounded in robust scene understanding. Recent Vision-Language Models (VLMs) exhibit promising capabilities such as object recognition, common-sense reasoning, and contextual understanding, capabilities that align with the nuanced requirements of social robot navigation. However, it remains unclear whether VLMs can accurately understand complex social navigation scenes (e.g., inferring the spatial-temporal relations among agents and human intentions), which is essential for safe and socially compliant robot navigation. While some recent works have explored the use of VLMs in social robot navigation, no existing work systematically evaluates their
קרא במקור המקורי