כתבה
arXiv cs.AI ·
Adaptive Perturbation Selection for Contrastive Audio Decoding
תקציר מקורי באנגליתarXiv:2607.00247v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) frequently hallucinate by overriding acoustic evidence with language priors. While contrastive decoding (CD) offers training-free mitigation, existing methods rely on blunt perturbations like masking or noise, leaving structured audio transformations unexplored. We explore this design space by evaluating a diverse library of targeted audio perturbations and adaptively selecting the optimal negative branch for each task and example. First, we improve upon earlier prompt engineering by showing that a simple binary yes/no constraint reduces the model's tendency to falsely confirm absent audio features. Second, evaluating our library across temporal, spectral, frequency, and amplitude domains reveals
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית