כתבה
arXiv cs.LG ·
בחינת הבנת הקשר במודלי שפה גדולים
Evaluation of Contextual Understanding in Large Language Models
במאמר זה, נבחן את יכולתם של מודלי שפה גדולים להבין קשרים. נציג פרקטיקה חדשה לבחינת הבנת הקשר, המשלבת דומיננטיות ודומיננטיות סמנטיות.
תקציר מקורי באנגליתarXiv:2609.09004v1 Announce Type: cross Abstract: Large Language Models (LLMs) demonstrate impressive performance across diverse NLP tasks, yet their ability to exhibit genuine contextual understanding remains uncertain. Traditional evaluation metrics such as perplexity, BiLingual Evaluation Understudy (BLEU), or surface-level accuracy fail to reveal how well LLMs extract, integrate, and reason over contextual information--a gap particularly critical in question answering, where models must align responses with contextually grounded knowledge rather than memorized associations. We propose a novel knowledge graph-based evaluation framework introducing Semantic Structural Similarity for KGs (S3KG), a hybrid similarity measure integrating structural and semantic similarity into a continuous e
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית