כתבה
arXiv cs.AI ·
AcFlow: בקרת מחוללי תמונות מטקסט באמצעות זרימת הפעלה מותנית
AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow
AcFlow הוא אלגוריתם ששולט במחוללי תמונות מטקסט מסוג diffusion transformers. הוא מאפשר שליטה על עוצמת הסגנון ומונע הופעת מושגים לא רצויים. AcFlow פועל על ידי תיאור קונספטואלי של ההתערבות הרצויה.
תקציר מקורי באנגליתarXiv:2609.10723v1 Announce Type: cross Abstract: Text-to-image diffusion transformers (DiTs) are powerful generators, yet direct prompting provides limited control interface for style intensity and can fail to suppress unwanted concepts. To enable these controls, we introduce AcFlow, an inference-time controller that transports intermediate layer image-token activations through a learned concept-conditioned velocity field while keeping the base DiT frozen. A textual concept description specifies the desired intervention, while the integration horizon provides a continuous control parameter. The field produces token-varying, activation-dependent updates. With parameters shared across concepts within each task family, the field supports fine-grained descriptions and generalizes to concepts
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית