Skip to content
AI Atlas
PaperActive

AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow

arxiv.org/abs/2609.10723

Updated 50 min ago · first seen 11 Sept 2026

paper_01M294H1MB0R1TEZ882VDV55SE

Published
11 Sept 2026
T1 · 50 min ago
arXiv
2609.10723
T1 · 50 min ago
Category
cs.CV
T1 · 50 min ago

Abstract

Text-to-image diffusion transformers (DiTs) are powerful generators, yet direct prompting provides limited control interface for style intensity and can fail to suppress unwanted concepts. To enable these controls, we introduce AcFlow, an inference-time controller that transports intermediate layer image-token activations through a learned concept-conditioned velocity field while keeping the base DiT frozen. A textual concept description specifies the desired intervention, while the integration horizon provides a continuous control parameter. The field produces token-varying, activation-dependent updates. With parameters shared across concepts within each task family, the field supports fine-grained descriptions and generalizes to concepts unseen during training without per-concept fitting. On style control, AcFlow achieves the best style--content trade-off among the evaluated baselines in the high-style-alignment regime. At a fixed operating point, AcFlow attains style--content alignment of 0.5365/0.2860, compared with 0.4397/0.2684 for the baseline with the highest style alignment. Qualitative results demonstrate suppression of diverse concepts, including cases where direct prompting fails. Our analyses support the learned velocity field as an adaptive control mechanism, with update directions varying across tokens and depend on their activation states. Our code is available at https://github.com/Nove1yst/AcFlow.

Authors 4

Junran Wang, Zehao Jin, Tianyu Luan, Xinjie Shen

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

arXiv id
2609.10723

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Categories
cs.CV, cs.AI

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Primary category
cs.CV

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

50 min ago

Conflicts

None