ML-QuantSubscribe

Machine learningML & AI Methods

DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation

Attention Control in Video Generation: DiTCtrl is a new training-free method for multi-prompt video generation under MM-DiT architectures, allowing for mask-guided precise semantic control across different prompts.

Featured in No. 80 on 1 Jan 2025 · 8 days after release · 81 citations today · published in 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Released
24 Dec 2024
First featured
No. 80 · 1 Jan 2025
Citations (Semantic Scholar)
81
Influential citations
10
Published in
2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Shares when featured
10
Identifier
arXiv:2412.18597

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page