Tora: Trajectory-oriented Diffusion Transformer for Video Generation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Zhenghao, Liao, Junchao, Li, Menghao, Dai, Zuozhuo, Qiu, Bingxue, Zhu, Siyu, Qin, Long, Wang, Weizhi |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
par: Zhang, Zhenghao, et autres
Publié: (2025)
par: Zhang, Zhenghao, et autres
Publié: (2025)
Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence
par: Liao, Junchao, et autres
Publié: (2026)
par: Liao, Junchao, et autres
Publié: (2026)
EffiVED:Efficient Video Editing via Text-instruction Diffusion Models
par: Zhang, Zhenghao, et autres
Publié: (2024)
par: Zhang, Zhenghao, et autres
Publié: (2024)
UVOSAM: A Mask-free Paradigm for Unsupervised Video Object Segmentation via Segment Anything Model
par: Zhang, Zhenghao, et autres
Publié: (2023)
par: Zhang, Zhenghao, et autres
Publié: (2023)
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
par: Zhang, Zhenghao, et autres
Publié: (2025)
par: Zhang, Zhenghao, et autres
Publié: (2025)
Identity-GRPO: Optimizing Multi-Human Identity-preserving Video Generation via Reinforcement Learning
par: Meng, Xiangyu, et autres
Publié: (2025)
par: Meng, Xiangyu, et autres
Publié: (2025)
TransVDM: Motion-Constrained Video Diffusion Model for Transparent Video Synthesis
par: Li, Menghao, et autres
Publié: (2025)
par: Li, Menghao, et autres
Publié: (2025)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
par: Zhu, Qiang, et autres
Publié: (2025)
par: Zhu, Qiang, et autres
Publié: (2025)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
par: Zhu, Jingyuan, et autres
Publié: (2026)
par: Zhu, Jingyuan, et autres
Publié: (2026)
Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
par: Zhu, Shenhao, et autres
Publié: (2024)
par: Zhu, Shenhao, et autres
Publié: (2024)
EnvSocial-Diff: A Diffusion-Based Crowd Simulation Model with Environmental Conditioning and Individual-Group Interaction
par: Zhao, Bingxue, et autres
Publié: (2026)
par: Zhao, Bingxue, et autres
Publié: (2026)
Video Diffusion Transformers are In-Context Learners
par: Fei, Zhengcong, et autres
Publié: (2024)
par: Fei, Zhengcong, et autres
Publié: (2024)
Steering Video Diffusion Transformers with Massive Activations
par: Cheng, Xianhang, et autres
Publié: (2026)
par: Cheng, Xianhang, et autres
Publié: (2026)
GenCompositor: Generative Video Compositing with Diffusion Transformer
par: Yang, Shuzhou, et autres
Publié: (2025)
par: Yang, Shuzhou, et autres
Publié: (2025)
FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control
par: Zhang, Zhiyuan, et autres
Publié: (2025)
par: Zhang, Zhiyuan, et autres
Publié: (2025)
ARLON: Boosting Diffusion Transformers with Autoregressive Models for Long Video Generation
par: Li, Zongyi, et autres
Publié: (2024)
par: Li, Zongyi, et autres
Publié: (2024)
Ingredients: Blending Custom Photos with Video Diffusion Transformers
par: Fei, Zhengcong, et autres
Publié: (2025)
par: Fei, Zhengcong, et autres
Publié: (2025)
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation
par: Wang, Yuelei, et autres
Publié: (2024)
par: Wang, Yuelei, et autres
Publié: (2024)
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
par: Hu, Runyi, et autres
Publié: (2025)
par: Hu, Runyi, et autres
Publié: (2025)
Latte: Latent Diffusion Transformer for Video Generation
par: Ma, Xin, et autres
Publié: (2024)
par: Ma, Xin, et autres
Publié: (2024)
FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models
par: Qiu, Haonan, et autres
Publié: (2024)
par: Qiu, Haonan, et autres
Publié: (2024)
TrailBlazer: Trajectory Control for Diffusion-Based Video Generation
par: Ma, Wan-Duo Kurt, et autres
Publié: (2023)
par: Ma, Wan-Duo Kurt, et autres
Publié: (2023)
Spectral and Trajectory Regularization for Diffusion Transformer Super-Resolution
par: Wang, Jingkai, et autres
Publié: (2026)
par: Wang, Jingkai, et autres
Publié: (2026)
Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion
par: Chen, Jingyuan, et autres
Publié: (2025)
par: Chen, Jingyuan, et autres
Publié: (2025)
DVD-Quant: Data-free Video Diffusion Transformers Quantization
par: Li, Zhiteng, et autres
Publié: (2025)
par: Li, Zhiteng, et autres
Publié: (2025)
Harmonious Group Choreography with Trajectory-Controllable Diffusion
par: Dai, Yuqin, et autres
Publié: (2024)
par: Dai, Yuqin, et autres
Publié: (2024)
TrajLoom: Dense Future Trajectory Generation from Video
par: Zhang, Zewei, et autres
Publié: (2026)
par: Zhang, Zewei, et autres
Publié: (2026)
MagicMirror: ID-Preserved Video Generation in Video Diffusion Transformers
par: Zhang, Yuechen, et autres
Publié: (2025)
par: Zhang, Yuechen, et autres
Publié: (2025)
Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers
par: Feng, Weilun, et autres
Publié: (2025)
par: Feng, Weilun, et autres
Publié: (2025)
Adaptive Caching for Faster Video Generation with Diffusion Transformers
par: Kahatapitiya, Kumara, et autres
Publié: (2024)
par: Kahatapitiya, Kumara, et autres
Publié: (2024)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
par: Zhang, Guiyu, et autres
Publié: (2026)
par: Zhang, Guiyu, et autres
Publié: (2026)
UltraViCo: Breaking Extrapolation Limits in Video Diffusion Transformers
par: Zhao, Min, et autres
Publié: (2025)
par: Zhao, Min, et autres
Publié: (2025)
Transforming Noise Distributions with Histogram Matching: Towards a Single Denoiser for All
par: Fu, Sheng, et autres
Publié: (2025)
par: Fu, Sheng, et autres
Publié: (2025)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
par: Zhao, Tianchen, et autres
Publié: (2024)
par: Zhao, Tianchen, et autres
Publié: (2024)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
par: Zhao, Min, et autres
Publié: (2025)
par: Zhao, Min, et autres
Publié: (2025)
Single-Stage Signal Attenuation Diffusion Model for Low-Light Image Enhancement and Denoising
par: Liu, Ying, et autres
Publié: (2026)
par: Liu, Ying, et autres
Publié: (2026)
SynWeather: Weather Observation Data Synthesis across Multiple Regions and Variables via a General Diffusion Transformer
par: Xu, Kaiyi, et autres
Publié: (2025)
par: Xu, Kaiyi, et autres
Publié: (2025)
Adaptive Noise-Tolerant Network for Image Segmentation
par: Li, Weizhi
Publié: (2025)
par: Li, Weizhi
Publié: (2025)
PoseTraj: Pose-Aware Trajectory Control in Video Diffusion
par: Ji, Longbin, et autres
Publié: (2025)
par: Ji, Longbin, et autres
Publié: (2025)
WiT: Waypoint Diffusion Transformers via Trajectory Conflict Navigation
par: Wang, Hainuo, et autres
Publié: (2026)
par: Wang, Hainuo, et autres
Publié: (2026)
Documents similaires
-
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
par: Zhang, Zhenghao, et autres
Publié: (2025) -
Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence
par: Liao, Junchao, et autres
Publié: (2026) -
EffiVED:Efficient Video Editing via Text-instruction Diffusion Models
par: Zhang, Zhenghao, et autres
Publié: (2024) -
UVOSAM: A Mask-free Paradigm for Unsupervised Video Object Segmentation via Segment Anything Model
par: Zhang, Zhenghao, et autres
Publié: (2023) -
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
par: Zhang, Zhenghao, et autres
Publié: (2025)