כתבה
arXiv cs.CL ·
Selective Listening: Mechanism-Guided Control of Audio Influence in Large Audio-Language Models
תקציר מקורי באנגליתarXiv:2610.11196v1 Announce Type: cross Abstract: Large audio-language models (LALMs) exploit multimodal evidence, yet task-irrelevant audio can alter text-reasoning decisions when listening is unnecessary. Aggregate Accuracy can hide this paired drift because audio-induced repairs and damages may cancel. Paired drift analysis and targeted interventions identify architecture-specific, intervention-sensitive late audio pathways as actionable control points. We introduce ICAP-Gate, which applies mechanism-guided, task-conditioned control to each model's pathway. Across four LALMs, two reasoning benchmarks, and environmental-sound and natural-speech interference, ICAP-Gate has lower point estimates for Influence Rate and Answer Flip than ungated inference in all 16 full-split model--condition
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית