יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

בדיקת הבנה חברתית בתגובות מקוונות סיניות

You Really Didn't Get That? Benchmarking Social Pragmatic Inference for Indirect and Playful Chinese Online Comments
חוקרים פיתחו בדיקה להערכת יכולתם של מודלים לשפה טבעית להבין תגובות מקוונות סיניות. הבדיקה כוללת 4,735 פריטים שנוצרו מתוך יותר מ-200,000 רשומות אינטראקציה ברשתות חברתיות סיניות. התוצאות הראו שהמודלים מצליחים לזהות עירוניות ומשחקיות, אך לעיתים קרובות מזהים לא נכון את המנגנון או המהלך הבין-אישי.
תקציר מקורי באנגליתarXiv:2609.04384v1 Announce Type: new Abstract: Chinese online comments often convey social meaning through indirect and playful language that is hard to interpret without context. Existing evaluations largely organize items around predefined phenomena or controlled pragmatic categories, leaving open whether models can distinguish plausible readings of what a naturally occurring comment is doing in a particular exchange. We introduce a benchmark for evaluating whether LLMs can recover such situated pragmatic meanings. From more than 200,000 public Chinese social media interaction records, we construct 4,735 human-validated diagnostic items, each pairing a target comment with reconstructed preceding context and plausible misreadings. We evaluate eight LLMs as both question writers and solve
קרא במקור המקורי