יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

בחינה של ASR ללא התאמה לשפה קורדית ג'ארוסי: ניתוח של תקיפה נפשית משותפת

Unadapted Multilingual ASR on a Garrusi Kurdish Evaluation Set: A Common-Reference Staged Normalization Analysis
בחינה של ASR ללא התאמה לשפה קורדית ג'ארוסי, באמצעות מודל MMS-1B-all. המחקר עוסק בבחינה של ASR ללא התאמה לשפה קורדית ג'ארוסי, באמצעות מודל MMS-1B-all. המחקר עוסק בבחינה של ASR ללא התאמה לשפה קורדית ג'ארוסי, באמצעות מודל MMS-1B-all.
תקציר מקורי באנגליתarXiv:2608.16379v2 Announce Type: replace Abstract: Evaluating speech recognition for a Kurdish variety written in a Latin field orthography, using a model that outputs Arabic script, creates a measurement problem before a modelling one: direct scoring treats writing-system differences as recognition errors. Jointly normalizing reference and hypothesis avoids this, but also changes reference tokenization, mixing agreement gains with a change in the scoring denominator. I evaluate MMS-1B-all with the Central Kurdish (ckb) adapter, used as released without adaptation, on 1,722 Garrusi questionnaire segments from five speakers (9,763 reference word tokens; 117.9 minutes). I use a common-reference design: the reference is folded once and fixed at 9,763 tokens, while only the hypothesis represe
קרא במקור המקורי