返回全部动态

AcFlow:通过学习条件激活流控制文生图扩散Transformer

原标题:AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow

arXiv cs.CV一手来源研究质量 88

AI 摘要

论文提出 AcFlow,一种在推理阶段控制文本到图像扩散 Transformer(DiT)的方法。它通过一个学习到的、以概念描述为条件的速度场,对中间层图像 token 激活进行传输,同时保持基础 DiT 冻结。该方法支持连续调节风格强度并抑制不需要的概念,在风格-内容权衡上优于基线,且无需针对每个概念单独拟合即可泛化到未见概念。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow Abstract Text-to-image diffusion transformers (DiTs) are powerful generators, yet direct prompting provides limited control interface for style intensity and can fail to suppress unwanted concepts. To enable these controls, we introduce AcFlow, an inference-time controller that transports intermediate layer image-token activations through a learned concept-conditioned velocity field while keeping the


发布时间:2026-09-11 12:00
抓取时间:2026-09-11 12:08
来源机构:arXiv
阅读原文arxiv.org