כתבה
arXiv cs.LG ·
A Controlled Audit of Personal AI Memory for Rating Prediction
תקציר מקורי באנגליתarXiv:2610.02764v1 Announce Type: new Abstract: In structured rating prediction, does a personal AI use historical item-rating associations, or mainly the user's rating tendencies? We audit this distinction by permuting historical ratings within each user while preserving the exact rating distribution, item support, and metadata. We combine this control with full history, native memory extraction, and matched numerical readers in a publicly frozen evaluation of 400 held-out user profiles and 6,160 target ratings across Coat and MovieLens. On Coat, the tested Qwen-written Mem0 pipeline increases user-macro mean absolute error relative to full history by 0.084 for Qwen and 0.149 for Phi; both family-adjusted bootstrap intervals exclude zero. Correct historical assignments help both readers o
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית