כתבה
arXiv cs.LG ·
תיאום צבע במרחבים סתומים של VAE ויישומיו
On Color Alignment in VAE Latent Spaces and Its Applications
מחקר זה מציג תיאום צבע במרחבים סתומים של VAE, ומציע שלושה יישומים: ColorTuning, שיפור טובה בצבע נומרי, שליטה בצבע, והעברת צבע.
תקציר מקורי באנגליתarXiv:2610.07072v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) are a key part of modern text-to-image models, which generate images within their latent space. VAEs are known to disentangle the main factors of variation in the data, and color is known to be one of the most structured of these in natural images: decorrelating it yields one luminance axis and two opponent-color axes. Color should therefore be expected to emerge as a distinct factor in the VAE latent space. Yet how these latent spaces represent color remains largely unexplored. In this work, we show that the VAEs of text-to-image models share a color subspace aligned with brightness and opponent-colors. Through a linear approximation of the encoder and targeted latent steering, we find this subspace consiste
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית