כתבה
arXiv cs.AI ·
RealCompanion: בדיקת הבנת האדם מתוך דיון ארוך-זמן בשיחה ריאל-עולם
RealCompanion: Benchmarking Human Understanding from Reasoning over Longitudinal Real-World Conversations
בדיקת הבנת האדם מתוך דיון ארוך-זמן בשיחה ריאל-עולם. המאמר עוסק בפיתוח בסיס נתונים לבדיקת הבנת האדם על ידי חברת RealCompanion.
תקציר מקורי באנגליתarXiv:2610.01780v1 Announce Type: new Abstract: A companion that talks with a person for months should come to understand them. It should remember what they said, infer who they are, and know when the past bears on the message in front of it. Testing this requires a real person's record, and such records are private, so benchmarks generate the person and the questions and settle in advance what matters. We release \bench, ten real relationships with an AI companion: 27,218 messages over up to 120 days, released as the conversation and four files derived from it, a profile, a persona, a chat ground truth and a question set, each citing the messages it rests on. Every chat label carries the reasoning trace that produced it, checked stage by stage against the conversation. Three findings foll
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית