יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

ARIA: מודל ליצירת מילים לשירים קנטונזים

ARIA: Audio-Driven Melody-Tone Relation Modeling for Cantonese Lyric Authoring
ARIA הוא מודל ליצירת מילים לשירים קנטונזים מתוך הקלטות שירה. המודל משתמש באותות קוליים רב-ערוציים ומבנה טונאלי על מנת ליצור מילים עם התאמה טובה למנגינה. נוסף על כך, ARIA משתמש במנגינות מוכוונות על מנת לשפר את איכות הטקסט.
תקציר מקורי באנגליתarXiv:2610.07902v1 Announce Type: new Abstract: Cantonese lyric writing requires close alignment between lexical tones and melodic pitch. Existing melody-guided lyric generation methods typically rely on symbolic melody to generate lyrics. However, in real songwriting scenarios, melodies are often expressed as raw singing audio or hummed recordings, where pitch is implicit, noisy, and unstructured, making these methods difficult to apply directly. To address this limitation, we propose ARIA, a two-stage audio-driven melody-tone relation modeling framework for Cantonese lyric authoring that generates Cantonese lyrics from singing recordings with provided character-level timestamps. Specifically, we first design a Tri-Stream Relation-Aware Tone Estimator (TRATE) to predict 0243 sequences fro
קרא במקור המקורי