Machine learningML & AI Methods
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation
Attention Control in Video Generation: DiTCtrl is a new training-free method for multi-prompt video generation under MM-DiT architectures, allowing for mask-guided precise semantic control across different prompts.
Featured in No. 80 on 1 Jan 2025 · 8 days after release · 81 citations today · published in 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- Released
- 24 Dec 2024
- First featured
- No. 80 · 1 Jan 2025
- Citations (Semantic Scholar)
- 81
- Influential citations
- 10
- Published in
- 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- Shares when featured
- 10
- Identifier
- arXiv:2412.18597
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).