יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

אטומים קבועים: קוארדינטות דקות של גאומטריה סמנטית מקומית בייצוגי דגלי שפה

Invariant Atoms: Sparse Coordinates of Local Semantic Geometry in Language Model Representations
דגלי שפה שומרים על משמעותם גם לאחר שינויים במילים, סגנון וסינטקס, וזה מעיד על כך שהתנועה הסמנטית המקומית נשארת קבועה.
תקציר מקורי באנגליתarXiv:2609.36451v1 Announce Type: new Abstract: Large language models often preserve meaning despite substantial changes in wording, style, and syntax, while small semantic edits can systematically alter their hidden representations. This suggests that semantic variation may be organized along recurring local directions. We propose the Invariant Atom Hypothesis: local semantic motion admits preferred sparse coordinates along directions that remain stable under meaning-preserving transformations. We learn a shared semantic frame and sparse coordinates that reconstruct semantic displacements while suppressing nuisance variation, with anchor-dependent diagonal modulation adjusting atom strengths without sample-specific rotations. Empirically, the atoms exhibit strong semantic--nuisance separa
קרא במקור המקורי