כתבה
arXiv cs.CL ·
PolicyMem: זיכרון מדיניות גאומטרי
PolicyMem: Geometric Policy Memory for LLM Governance
PolicyMem הוא זיכרון מדיניות גאומטרי שמאפשר שימוש חוזר בראיות מדיניות. הוא משמש לגילוי התנהגות לא בטוחה של מודלים גדולים ומאפשר כתיבה מחדש ואימות.
תקציר מקורי באנגליתarXiv:2609.13734v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in real-world high-stakes applications, effective governance has become essential. Existing safeguards largely follow two paradigms: learning-based guards provide strong semantic discrimination but couple policy behavior to trained models and taxonomies, while programmable frameworks offer flexible control but require substantial manual prompt and workflow engineering. Neither externalizes policies as reusable operational states, making it difficult to consistently reuse policy evidence across detection, intervention, and verification. In this paper, we introduce PolicyMem, a geometric policy memory that externalizes natural-language policies as reusable geometric memory objects repres
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית