יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

התאמה יעילה של LLMs לזיהוי שיחה גזענית בשפות עם משאבים נמוכים: מחקר השוואתי על רומאני עורבי

Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu
המאמר עוסק בהתאמה יעילה של LLMs לזיהוי שיחה גזענית בשפות עם משאבים נמוכים. המחברים השוו בין דגמי LLMs שונים, כולל Mistral ו-LLaMA, ומצאו שהתאמה יעילה של פרמטרים יכולה לשפר את הביצועים באופן יעיל.
תקציר מקורי באנגליתarXiv:2608.18142v2 Announce Type: replace Abstract: It is challenging to detect hate speech in Low Resource Languages (LRLs) because of the absence of annotated data, the informality of its language structure, and the lack of standardized grammar. A good example of such a challenge is Roman Urdu which is broadly used by South Asians on social media and has a high variation while lacking contextually consistent spellings. The objective of this paper is to conduct a comprehensive assessment of Large Language Models (LLMs) for Hate Speech Detection (HSD) in Roman Urdu script and fine-tune these models using the Parameter-Efficient Fine-Tuning (PEFT) method called Low-Rank Adaptation (LoRA). To evaluate zero-shot inference, we benchmarked it against PEFT on different transformer models, includ
קרא במקור המקורי