יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

עבר את נקודות הסיום: תנאי זמן וקיבולת מותאמים לביקורת עדכון ידע רציף

Beyond Endpoint Scores: Time- and Capacity-Conditioned Evaluation of Continual Knowledge Updating
מחקר חדש מציג שיטה לביקורת עדכון ידע רציף, ומגלה שנקודות סיום עשויות להיות מוטעות. המחקר משווה בין שיטות שונות ומציע דרך חדשה לבחור בשיטה הטובה ביותר. המחקר כולל ניסויים על מודלי Qwen ו-Llama.
תקציר מקורי באנגליתarXiv:2609.03900v1 Announce Type: new Abstract: Continual knowledge-updating methods are often declared superior from one final checkpoint and one conventional adapter rank. We show that this can be insufficient to identify the better operating point. Holding a periodic hierarchy fixed, we compare it with cumulative replay over a 24-month Wikidata stream while varying evaluation month, replay LoRA rank, and query formulation. The apparent winner changes across this region: on Qwen2.5-1.5B, the hierarchy's 5.0-point advantage over rank-8 replay becomes an 11.6-point deficit against rank-72 replay, and at high ranks a consolidation-aligned endpoint can suggest a tie while time-averaged replay leads by 9-13 points. The same rank-conditioned reversal appears on Llama-3.2-1B and held-out paraph
קרא במקור המקורי