יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

CRAF: חיבור תחומי-שונה לזיהוי דיבור-פסאודו

CRAF: Cross-View Residual-Aware Fusion for Deepfake Speech Detection
CRAF היא פלטפורמה של חיבור תחומי-שונה לזיהוי דיבור-פסאודו. היא משתמשת ב-ALLM כדי להעשיר את ה-SLL ולהבחין בין המידע המשותף לבין המידע המשותף.
תקציר מקורי באנגליתarXiv:2609.13842v1 Announce Type: cross Abstract: Recent advances in speech synthesis and voice conversion have made deepfake speech increasingly realistic, making generalization to unseen spoofing attacks a critical challenge. Pretrained speech and audio models offer a promising direction for improving robustness to such unseen attacks. Self-supervised learning (SSL) models capture fine-grained, low-level acoustic characteristics, whereas Auditory Large Language Models (ALLMs) provide higher-level contextual representations. These complementary views can provide useful cues for improving generalization to unseen attacks. However, direct fusion does not explicitly disentangle information shared across the two views from view-specific complementary information, limiting effective cross-view
קרא במקור המקורי