כתבה
arXiv cs.CL ·
PlurVA-LLM-2026: הסתגלות ערכים רב-לשונית
PlurVA-LLM-2026 Shared Task Track-1: Pluralistic Value Alignment in LLMs via Multilingual Fine-Tuning and Threshold Calibration
PlurVA-LLM-2026 הוא מיזם להסתגלות ערכים רב-לשונית במודלים LLM. המחקר משתמש ב-Llama 3.1 8B Instruct ומיישם טכניקות עיבוד וכיול סף. התוצאות מראות דיוק גבוה בשלוש שפות: סינית, אינדונזית וסינהלית.
תקציר מקורי באנגליתarXiv:2609.32382v2 Announce Type: replace Abstract: We present our system for the PlurVA-LLM 2026 Shared Task Track-1, which focuses on pluralistic value alignment in the contexts of China, Indonesia, and Sri Lanka. For this resource-constrained track, we fine-tuned Llama 3.1 8B Instruct using 4-bit QLoRA. Our approach combines option-permutation augmentation for Chinese data, annotator vote expansion for Indonesian data, and binary reformulation with SinhalaMMLU augmentation for Sri Lankan data. We further applied conditional threshold calibration to the predictions for the Sri Lankan data. The final system achieved accuracies of 0.785 for Chinese, 0.715 for Indonesian, and 0.916 for Sri Lankan, resulting in an overall macro-average accuracy of 0.805.
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית