כתבה
arXiv cs.LG ·
SalesLoop: Reinforcement Learning from Performance Feedback for Sales Lead Ranking
תקציר מקורי באנגליתarXiv:2607.20655v1 Announce Type: new Abstract: Lead ranking in Customer Relationship Management (CRM) systems faces a persistent challenge: models achieving high offline accuracy often underperform in production. We identify three fundamental gaps responsible for this disconnect: offline-online metric mismatch, pointwise-listwise objective misalignment, and temporal distribution drift. To address these gaps, we propose SalesLoop, a reinforcement learning framework that establishes a closed feedback loop between model predictions and real-world business outcomes. Our approach introduces (1) a performance-aware reward that encodes conversion outcomes weighted by ranking position and conversion velocity, and (2) Discriminative GRPO, a listwise optimization objective that adapts Group Relativ
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית