כתבה
arXiv cs.CL ·
The Impact of Editorial Intervention on Detecting Native Language Traces
תקציר מקורי באנגליתarXiv:2605.10216v2 Announce Type: replace Abstract: Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writing. With the advent of human-AI co-authorship, learner texts are routinely corrected and rewritten by large language models, fundamentally altering the linguistic features NLI approaches depend on. In this paper, we investigate the robustness of L1 traces across increasing degrees of editorial intervention. By processing 450 essays from the Write & Improve 2024 (W&I) corpus through varying levels of grammatical error correction and paraphrasing, we demonstrate that L1 attribution does not depend solely on surface-level errors. Instead, the detection models appear to leverage deeper L1-related features, including unid
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית