BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hosseyni, S. Rohollah, Rahmani, Ali Ahmad, Seyedmohammadi, S. Jamal, Seyedin, Sanaz, Mohammadi, Arash |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Human Action Recognition in Still Images Using ConViT
von: Hosseyni, Seyed Rohollah, et al.
Veröffentlicht: (2023)
von: Hosseyni, Seyed Rohollah, et al.
Veröffentlicht: (2023)
Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation
von: Zhang, Yuhui, et al.
Veröffentlicht: (2023)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2023)
MoTe: Learning Motion-Text Diffusion Model for Multiple Generation Tasks
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
AAMDM: Accelerated Auto-regressive Motion Diffusion Model
von: Li, Tianyu, et al.
Veröffentlicht: (2023)
von: Li, Tianyu, et al.
Veröffentlicht: (2023)
SELMA: Learning and Merging Skill-Specific Text-to-Image Experts with Auto-Generated Data
von: Li, Jialu, et al.
Veröffentlicht: (2024)
von: Li, Jialu, et al.
Veröffentlicht: (2024)
Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
Bidirectional Long-Range Parser for Sequential Data Understanding
von: Leotescu, George, et al.
Veröffentlicht: (2024)
von: Leotescu, George, et al.
Veröffentlicht: (2024)
Bilingual Text-to-Motion Generation: A New Benchmark and Baselines
von: Weng, Wanjiang, et al.
Veröffentlicht: (2026)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2026)
ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models
von: Patel, Maitreya, et al.
Veröffentlicht: (2023)
von: Patel, Maitreya, et al.
Veröffentlicht: (2023)
Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation
von: Yuan, Huizhuo, et al.
Veröffentlicht: (2024)
von: Yuan, Huizhuo, et al.
Veröffentlicht: (2024)
SurGen: Text-Guided Diffusion Model for Surgical Video Generation
von: Cho, Joseph, et al.
Veröffentlicht: (2024)
von: Cho, Joseph, et al.
Veröffentlicht: (2024)
MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text
von: Zhang, Junzhe, et al.
Veröffentlicht: (2025)
von: Zhang, Junzhe, et al.
Veröffentlicht: (2025)
Words in Motion: Extracting Interpretable Control Vectors for Motion Transformers
von: Tas, Omer Sahin, et al.
Veröffentlicht: (2024)
von: Tas, Omer Sahin, et al.
Veröffentlicht: (2024)
Many-to-many Image Generation with Auto-regressive Diffusion Models
von: Shen, Ying, et al.
Veröffentlicht: (2024)
von: Shen, Ying, et al.
Veröffentlicht: (2024)
Text-to-Image Cross-Modal Generation: A Systematic Review
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
von: Cheng, Jiale, et al.
Veröffentlicht: (2025)
von: Cheng, Jiale, et al.
Veröffentlicht: (2025)
DreamReward: Text-to-3D Generation with Human Preference
von: Ye, Junliang, et al.
Veröffentlicht: (2024)
von: Ye, Junliang, et al.
Veröffentlicht: (2024)
Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models
von: Motamed, Saman, et al.
Veröffentlicht: (2023)
von: Motamed, Saman, et al.
Veröffentlicht: (2023)
Controlled Training Data Generation with Diffusion Models
von: Yeo, Teresa, et al.
Veröffentlicht: (2024)
von: Yeo, Teresa, et al.
Veröffentlicht: (2024)
One Category One Prompt: Dataset Distillation using Diffusion Models
von: Abbasi, Ali, et al.
Veröffentlicht: (2024)
von: Abbasi, Ali, et al.
Veröffentlicht: (2024)
T$^3$Bench: Benchmarking Current Progress in Text-to-3D Generation
von: He, Yuze, et al.
Veröffentlicht: (2023)
von: He, Yuze, et al.
Veröffentlicht: (2023)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
PromptTA: Prompt-driven Text Adapter for Source-free Domain Generalization
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
IMAGINE-E: Image Generation Intelligence Evaluation of State-of-the-art Text-to-Image Models
von: Lei, Jiayi, et al.
Veröffentlicht: (2025)
von: Lei, Jiayi, et al.
Veröffentlicht: (2025)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
von: Singhal, Raghav, et al.
Veröffentlicht: (2025)
von: Singhal, Raghav, et al.
Veröffentlicht: (2025)
AutoArabic: A Three-Stage Framework for Localizing Video-Text Retrieval Benchmarks
von: Eltahir, Mohamed, et al.
Veröffentlicht: (2025)
von: Eltahir, Mohamed, et al.
Veröffentlicht: (2025)
Beyond Thumbs Up/Down: Untangling Challenges of Fine-Grained Feedback for Text-to-Image Generation
von: Collins, Katherine M., et al.
Veröffentlicht: (2024)
von: Collins, Katherine M., et al.
Veröffentlicht: (2024)
Image2Text2Image: A Novel Framework for Label-Free Evaluation of Image-to-Text Generation with Text-to-Image Diffusion Models
von: Huang, Jia-Hong, et al.
Veröffentlicht: (2024)
von: Huang, Jia-Hong, et al.
Veröffentlicht: (2024)
Transformer with Controlled Attention for Synchronous Motion Captioning
von: Radouane, Karim, et al.
Veröffentlicht: (2024)
von: Radouane, Karim, et al.
Veröffentlicht: (2024)
Flexible-length Text Infilling for Discrete Diffusion Models
von: Zhang, Andrew, et al.
Veröffentlicht: (2025)
von: Zhang, Andrew, et al.
Veröffentlicht: (2025)
MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
CEIDM: A Controlled Entity and Interaction Diffusion Model for Enhanced Text-to-Image Generation
von: Yang, Mingyue, et al.
Veröffentlicht: (2025)
von: Yang, Mingyue, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
Mostly Text, Smart Visuals: Asymmetric Text-Visual Pruning for Large Vision-Language Models
von: Li, Sijie, et al.
Veröffentlicht: (2026)
von: Li, Sijie, et al.
Veröffentlicht: (2026)
Accelerating Auto-regressive Text-to-Image Generation with Training-free Speculative Jacobi Decoding
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation
von: Ma, Jian, et al.
Veröffentlicht: (2023)
von: Ma, Jian, et al.
Veröffentlicht: (2023)
Evaluating Text-to-Visual Generation with Image-to-Text Generation
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2024)
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2024)
CRCE: Coreference-Retention Concept Erasure in Text-to-Image Diffusion Models
von: Xue, Yuyang, et al.
Veröffentlicht: (2025)
von: Xue, Yuyang, et al.
Veröffentlicht: (2025)
CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation
von: Deng, Dazhen, et al.
Veröffentlicht: (2025)
von: Deng, Dazhen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Human Action Recognition in Still Images Using ConViT
von: Hosseyni, Seyed Rohollah, et al.
Veröffentlicht: (2023) -
Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation
von: Zhang, Yuhui, et al.
Veröffentlicht: (2023) -
MoTe: Learning Motion-Text Diffusion Model for Multiple Generation Tasks
von: Wu, Yiming, et al.
Veröffentlicht: (2024) -
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024) -
AAMDM: Accelerated Auto-regressive Motion Diffusion Model
von: Li, Tianyu, et al.
Veröffentlicht: (2023)