יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

תקציבי זיוף מאומתים: טבלאות-מנצחים-זמן-אחידות תחת רצועת-זיוף-תקיפה

Certified Corruption Budgets: Anytime-Valid Leaderboard Claims under Adaptive Rigging
מאמר זה עוסק בהבטחות לטבלאות מודלי AI נגד זיוף והטיית קולות. המחברים הציגו תקציבי זיוף מאומתים, שהם טולנט של תקציבי זיוף רגילים, שמאפשרים לטבלאות להיות תקינות גם במקרה של זיוף. התקציב המאומת מתחלק לשני סוגים: תקציב לזיופים של קולות ותקציב לזיופים של קולות שנראים כאילו הם קולות חדשים. התקציב המאומת נחשב לבטוח יותר מתקציבי זיוף רגילים, כיוון שהוא מאפשר לטבלאות להיות תקינות גם במקרה של זיוף.
תקציר מקורי באנגליתarXiv:2610.10597v1 Announce Type: cross Abstract: Public leaderboards for AI models are read continuously, and attackers can see every published standing. Vote rigging, selective disclosure of private variants, and benchmark contamination can each move a ranking. Existing guarantees assume genuine records or bound the corruption per step, which an attacker who corrupts in bursts evades. We introduce the certified corruption budget, a tolerance $\widehat{B}_t$ computed after $t$ records and published with each pairwise claim. With probability at least $1-\alpha$, simultaneously at all times, the claim is correct or more than $\widehat{B}_t$ records were corrupted. It holds against attackers who watch every certificate, with no bound on their budget. Forged records and records altered once s
קרא במקור המקורי