יום שני, 5 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

CoLMbo-SV: מודל שפה חדש לאימות דיבור

CoLMbo-SV: A Grounded Language Model for Explainable Speaker Verification
מודל CoLMbo-SV מציע אימות דיבור ניתן להסבר, עם תיעוד קולי מבוסס. המודל CoLMbo-SV משלב זיהוי דיבור חזק עם דיווחים מבוססי קול, ומציע תיעוד קולי מבוסס. המודל CoLMbo-SV נבחן על VoxCeleb1-O והציג 0.99% EER, והציג תיעוד קולי מבוסס.
תקציר מקורי באנגליתarXiv:2609.33212v2 Announce Type: replace Abstract: Speaker verification systems achieve high accuracy but provide little account of the acoustic evidence behind their judgments. Making these systems inspectable requires exposing interpretable evidence while retaining the richer information on which their decisions depend. We present \textbf{CoLMbo-SV}, a speaker language model that combines strong speaker discrimination with structured, acoustically grounded comparison reports. By connecting a pretrained speaker encoder to a language model and supplying explicit acoustic measurements, CoLMbo-SV makes voice comparisons inspectable without restricting verification to the evidence verbalized in its reports. We additionally introduce \textbf{VoxReason}, paired recordings with measured acousti
קרא במקור המקורי