יום חמישי, 8 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

לימוד חיזוק עם הדרכת LLM להגנה סייבר אוטונומית

Ask the Expert: LLM-Guided Reinforcement Learning for Autonomous Cyber Defense
חוקרים פיתחו שיטה חדשה להגנה סייבר אוטונומית באמצעות לימוד חיזוק מונחה על ידי מודל LLM. השיטה משתמשת במודל LLM כדי לספק המלצות לפעולות הגנה, ואז משתמשת ב-PPO כדי ללמוד מדיניות הגנה אופטימלית. השיטה הראתה תוצאות טובות יותר מאשר שיטות אחרות בתחום.
תקציר מקורי באנגליתarXiv:2610.09337v1 Announce Type: cross Abstract: Policy-based reinforcement learning (RL) approaches have produced promising results for autonomous cyber defense; however, they are sample-inefficient in settings where defenders must respond under delayed, partial observations with actions from large action spaces. While large language models (LLMs) may reason semantically about security state space, high latency and trust assumptions prevent attractive in-line deployment models. We introduce Ask the Expert, a training-time guidance framework which first summarizes hard cyber-defense states, then intermittently queries an LLM for host-level defensive recommendations via a constrained action interface, and finally transforms those recommendations into tiered reward shaping for use with PPO.
קרא במקור המקורי