כתבה
arXiv cs.AI ·
MulRobBench: תקן-החלטה לבקרת אבטחה ופקודת-בטיחות עבור סוכני UAV מודאליים
MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents
תקן-החלטה לבקרת אבטחה ופקודת-בטיחות עבור סוכני UAV מודאליים, המבוסס על 3,024 דגימות ו-12 מטריקס סורינג.
תקציר מקורי באנגליתarXiv:2607.23870v2 Announce Type: replace-cross Abstract: In IoT-enabled smart-city settings, Uncrewed Aerial Vehicles (UAVs) are evolving from passive sensing platforms into cyber-physical decision makers that must respect operational rules under degraded observations and ambiguous language. Existing UAV and multimodal benchmarks cover aerial perception, navigation, collaboration, and task reasoning, but rarely test whether physical evidence, protocol constraints, and action risk stay coupled at critical decisions. We introduce MulRobBench, an offline, protocol-conditioned benchmark for Vision-Language-Action (VLA) UAV agents that links real UAV multimodal observations, protocol-level security-policy constraints, and action-level cyber-physical safety within an auditable decision contract
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית