יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

חשוב מחדש: קול תהליכי - סקאלה ואובייקטיב תכנון

Rethinking Procedural Audio Pre-training: Source Scaling and Objective Adaptation
חשוב מחדש: קול תהליכי - סקאלה ואובייקטיב תכנון. ניתוח של סקאלה ואובייקטיב תכנון בקול תהליכי.
תקציר מקורי באנגליתarXiv:2609.15067v1 Announce Type: cross Abstract: Procedural audio has emerged as a viable source for transferable audio representation learning, but its design principles remain unclear.We revisit two questions: how a procedural source should be scaled, and whether training choices developed on natural audio should transfer unchanged to procedural data.Using a controlled source, we separate scale into formula-class coverage C and within-class rendering diversity I.Experiments with FDSL and AudioMAE show that these two forms of scale provide different benefits and depend on the learning formulation and downstream task. A matched AudioMAE study further shows that procedural audio favors low mask ratios (10%--25%), whereas AudioSet-28K favors 50%--75%. Shared-codebook analysis reveals lower
קרא במקור המקורי