כתבה
arXiv cs.LG ·
חזרה למערה של פלאטון: בחינה של קונברגנציה נגזרת-תרבותית בקנה מידה גדול
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
לפי מחקר, אין ראיה לקונברגנציה נגזרת-תרבותית דק-מבנית במודלי רשת נוירונים-תרבותיים-תרבותיים.
תקציר מקורי באנגליתarXiv:2604.18572v3 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis posits that neural networks trained on different modalities (e.g., text and images) converge toward a shared representation of reality. If true, this has significant implications for whether modality choice matters at all. In this paper, we show that the evidence for this claim is substantially weaker than subsequent work suggests. The mutual $k$-nearest-neighbor metric used on 1024 text-image pairs in the original study captures only coarse structure. To keep the alignment from collapsing as one scales up the data, $k$ has to grow proportionally, undercutting the argument for fine-grained representational convergence. The reported increase in alignment with language model strength saturates fo
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית