כתבה
arXiv cs.CL ·
Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs
תקציר מקורי באנגליתarXiv:2607.03936v2 Announce Type: replace Abstract: Dialectal data are scarce relative to Modern Standard Arabic (MSA), causing Arabic LLMs to overproduce MSA and struggle with dialectally accurate generation. This raises a fundamental interpretability question about where and how dialectal features are encoded within model internals and whether these representations can improve dialect generation without fine-tuning. We study two inference-time approaches as interpretability probes and control mechanisms. First, neuron-level analysis identifies sparse populations that encode dialect-specific features and tests whether amplifying or suppressing them steers model outputs toward target dialects. Second, vector steering extracts dialect-specific activation directions and injects them during i
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית