כתבה
arXiv cs.CL ·
צביעת חתימת LLM לפי כוונת כתיבה
Forging LLM Authorship Fingerprints with Targeted Rewriting
חוקרים פיתחו שיטה לצביעת חתימת LLM לפי כוונת כתיבה. השיטה, ForgePrint, מאפשרת לכתוב מחדש טקסט כך שמערכת הזיהוי של ה-LLM תזהה אותו כמקורו של עמוד כתב נתון. השיטה נבדקה במספר תרחישים, כולל כתיבה מחדש של סיכומים וטקסטים רגילים.
תקציר מקורי באנגליתarXiv:2609.38831v1 Announce Type: new Abstract: Model-attribution classifiers can often identify which language model produced a text, making model-specific writing patterns a signal of provenance. Accurate attribution on unmodified text, however, does not show whether the prediction still identifies the original source after deliberate rewriting. We formulate this problem as targeted fingerprint transfer: rewriting one model's output so that attribution classifiers assign it to a chosen target model. We study summarization, where different models receive the same document and express the same underlying content, providing a controlled setting for conditional generation. We introduce ForgePrint, a search-then-distil framework that first searches for rewrites that move attribution toward a
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית