Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Fei, Yuanchen, Zou, Yude, Kang, Zejian, Li, Ming, Zhou, Jiaying, Huang, Xiangru |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
di: Kang, Zejian, et al.
Pubblicazione: (2026)
di: Kang, Zejian, et al.
Pubblicazione: (2026)
KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes
di: Wu, Jingchao, et al.
Pubblicazione: (2025)
di: Wu, Jingchao, et al.
Pubblicazione: (2025)
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
di: Zheng, Kai, et al.
Pubblicazione: (2026)
di: Zheng, Kai, et al.
Pubblicazione: (2026)
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
di: Yang, Wentao, et al.
Pubblicazione: (2026)
di: Yang, Wentao, et al.
Pubblicazione: (2026)
AvatarShield: Visual Reinforcement Learning for Human-Centric Synthetic Video Detection
di: Xu, Zhipei, et al.
Pubblicazione: (2025)
di: Xu, Zhipei, et al.
Pubblicazione: (2025)
SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision
di: Kang, Zejian, et al.
Pubblicazione: (2026)
di: Kang, Zejian, et al.
Pubblicazione: (2026)
InfBaGel: Human-Object-Scene Interaction Generation with Dynamic Perception and Iterative Refinement
di: Zou, Yude, et al.
Pubblicazione: (2026)
di: Zou, Yude, et al.
Pubblicazione: (2026)
Synthetic Human Action Video Data Generation with Pose Transfer
di: Knapp, Vaclav, et al.
Pubblicazione: (2025)
di: Knapp, Vaclav, et al.
Pubblicazione: (2025)
VideoWeave: A Data-Centric Approach for Efficient Video Understanding
di: Durante, Zane, et al.
Pubblicazione: (2026)
di: Durante, Zane, et al.
Pubblicazione: (2026)
HumanVideo-MME: Benchmarking MLLMs for Human-Centric Video Understanding
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions
di: Zhang, Yiyuan, et al.
Pubblicazione: (2024)
di: Zhang, Yiyuan, et al.
Pubblicazione: (2024)
HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks
di: Zhou, Ting, et al.
Pubblicazione: (2024)
di: Zhou, Ting, et al.
Pubblicazione: (2024)
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation
di: Zheng, Sixiao, et al.
Pubblicazione: (2025)
di: Zheng, Sixiao, et al.
Pubblicazione: (2025)
Data Augmentation in Human-Centric Vision
di: Jiang, Wentao, et al.
Pubblicazione: (2024)
di: Jiang, Wentao, et al.
Pubblicazione: (2024)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
Depth Any Video with Scalable Synthetic Data
di: Yang, Honghui, et al.
Pubblicazione: (2024)
di: Yang, Honghui, et al.
Pubblicazione: (2024)
Auto3R: Automated 3D Reconstruction and Scanning via Data-driven Uncertainty Quantification
di: Shen, Chentao, et al.
Pubblicazione: (2025)
di: Shen, Chentao, et al.
Pubblicazione: (2025)
Synthetic FMCW Radar Range Azimuth Maps Augmentation with Generative Diffusion Model
di: Wang, Zhaoze, et al.
Pubblicazione: (2026)
di: Wang, Zhaoze, et al.
Pubblicazione: (2026)
Synthetic Simplicity: Unveiling Bias in Medical Data Augmentation
di: Babu, Krishan Agyakari Raja, et al.
Pubblicazione: (2024)
di: Babu, Krishan Agyakari Raja, et al.
Pubblicazione: (2024)
Non-stationary BERT: Exploring Augmented IMU Data For Robust Human Activity Recognition
di: Sun, Ning, et al.
Pubblicazione: (2024)
di: Sun, Ning, et al.
Pubblicazione: (2024)
SceneRAG: Scene-level Retrieval-Augmented Generation for Video Understanding
di: Zeng, Nianbo, et al.
Pubblicazione: (2025)
di: Zeng, Nianbo, et al.
Pubblicazione: (2025)
MemCam: Memory-Augmented Camera Control for Consistent Video Generation
di: Gao, Xinhang, et al.
Pubblicazione: (2026)
di: Gao, Xinhang, et al.
Pubblicazione: (2026)
Generating Synthetic Data via Augmentations for Improved Facial Resemblance in DreamBooth and InstantID
di: Ulusan, Koray, et al.
Pubblicazione: (2025)
di: Ulusan, Koray, et al.
Pubblicazione: (2025)
Interact3D: Compositional 3D Generation of Interactive Objects
di: Shan, Hui, et al.
Pubblicazione: (2026)
di: Shan, Hui, et al.
Pubblicazione: (2026)
NI-Tex: Non-isometric Image-based Garment Texture Generation
di: Shan, Hui, et al.
Pubblicazione: (2025)
di: Shan, Hui, et al.
Pubblicazione: (2025)
AA-SGAN: Adversarially Augmented Social GAN with Synthetic Data
di: Zaffaroni, Mirko, et al.
Pubblicazione: (2024)
di: Zaffaroni, Mirko, et al.
Pubblicazione: (2024)
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
di: Noghre, Ghazal Alinezhad, et al.
Pubblicazione: (2024)
di: Noghre, Ghazal Alinezhad, et al.
Pubblicazione: (2024)
EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation
di: Vandersanden, Jente, et al.
Pubblicazione: (2026)
di: Vandersanden, Jente, et al.
Pubblicazione: (2026)
Video-RAG: Visually-aligned Retrieval-Augmented Long Video Comprehension
di: Luo, Yongdong, et al.
Pubblicazione: (2024)
di: Luo, Yongdong, et al.
Pubblicazione: (2024)
Are Synthetic Videos Useful? A Benchmark for Retrieval-Centric Evaluation of Synthetic Videos
di: Zhao, Zecheng, et al.
Pubblicazione: (2025)
di: Zhao, Zecheng, et al.
Pubblicazione: (2025)
Rethinking Driving World Model as Synthetic Data Generator for Perception Tasks
di: Zeng, Kai, et al.
Pubblicazione: (2025)
di: Zeng, Kai, et al.
Pubblicazione: (2025)
Boximator: Generating Rich and Controllable Motions for Video Synthesis
di: Wang, Jiawei, et al.
Pubblicazione: (2024)
di: Wang, Jiawei, et al.
Pubblicazione: (2024)
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
di: Noghre, Ghazal Alinezhad, et al.
Pubblicazione: (2024)
di: Noghre, Ghazal Alinezhad, et al.
Pubblicazione: (2024)
Learning to Find Missing Video Frames with Synthetic Data Augmentation: A General Framework and Application in Generating Thermal Images Using RGB Cameras
di: Andersen, Mathias Viborg, et al.
Pubblicazione: (2024)
di: Andersen, Mathias Viborg, et al.
Pubblicazione: (2024)
Concept-as-Tree: A Controllable Synthetic Data Framework Makes Stronger Personalized VLMs
di: An, Ruichuan, et al.
Pubblicazione: (2025)
di: An, Ruichuan, et al.
Pubblicazione: (2025)
Synthetic Data Augmentation for Table Detection: Re-evaluating TableNet's Performance with Automatically Generated Document Images
di: Sahukara, Krishna, et al.
Pubblicazione: (2025)
di: Sahukara, Krishna, et al.
Pubblicazione: (2025)
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos
di: Zhou, Zhiyu, et al.
Pubblicazione: (2026)
di: Zhou, Zhiyu, et al.
Pubblicazione: (2026)
Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks
di: Li, Jie, et al.
Pubblicazione: (2025)
di: Li, Jie, et al.
Pubblicazione: (2025)
FRAG: Frame Selection Augmented Generation for Long Video and Long Document Understanding
di: Huang, De-An, et al.
Pubblicazione: (2025)
di: Huang, De-An, et al.
Pubblicazione: (2025)
AURA: Development and Validation of an Augmented Unplanned Removal Alert System using Synthetic ICU Videos
di: Seo, Junhyuk, et al.
Pubblicazione: (2025)
di: Seo, Junhyuk, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
di: Kang, Zejian, et al.
Pubblicazione: (2026) -
KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes
di: Wu, Jingchao, et al.
Pubblicazione: (2025) -
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
di: Zheng, Kai, et al.
Pubblicazione: (2026) -
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
di: Yang, Wentao, et al.
Pubblicazione: (2026) -
AvatarShield: Visual Reinforcement Learning for Human-Centric Synthetic Video Detection
di: Xu, Zhipei, et al.
Pubblicazione: (2025)