יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

RelightFormer: מעבר תאורה רב-מבטים

RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting
RelightFormer הוא מודל גנרטיבי מבוסס טרנספורמר לתאורה רב-מבטית. הוא מאפשר תאורה ישירה ורב-מבטית ללא צורך בחישוב תכונות פנימיות. המודל מבוסס על מודל וידאו ומשתמש בקידוד מיקום תלוי-הילוך.
תקציר מקורי באנגליתarXiv:2609.07414v1 Announce Type: cross Abstract: Image relighting is traditionally tackled via complex inverse rendering pipelines, which suffer from ill-posed optimization, or single-image generative models that ignore crucial multi-view cues necessary for understanding 3D geometry and material interactions. To address these limitations, we introduce a feed-forward generative Transformer for direct single- and multi-view image relighting that entirely bypasses explicit intrinsic property estimation. Adapted from a video foundation model, our architecture features a latent illumination module that dynamically injects target environment maps into spatial features via cross-attention. Furthermore, we employ permutation-invariant positional encodings to symmetrically process unordered multi-
קרא במקור המקורי