Alignment is All You Need: A Training-free Augmentation Strategy for Pose-guided Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Jin, Xiaoyu, Xu, Zunnan, Ou, Mingwen, Yang, Wenming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Transferable-guided Attention Is All You Need for Video Domain Adaptation
por: Sacilotti, André, et al.
Publicado: (2024)
por: Sacilotti, André, et al.
Publicado: (2024)
Reasoning is All You Need for Video Generalization: A Counterfactual Benchmark with Sub-question Evaluation
por: Zhou, Qiji, et al.
Publicado: (2025)
por: Zhou, Qiji, et al.
Publicado: (2025)
SAM-R1: Leveraging SAM for Reward Feedback in Multimodal Segmentation via Reinforcement Learning
por: Huang, Jiaqi, et al.
Publicado: (2025)
por: Huang, Jiaqi, et al.
Publicado: (2025)
All You Need to Know About Training Image Retrieval Models
por: Berton, Gabriele, et al.
Publicado: (2025)
por: Berton, Gabriele, et al.
Publicado: (2025)
SVAC: Scaling Is All You Need For Referring Video Object Segmentation
por: Zhang, Li, et al.
Publicado: (2025)
por: Zhang, Li, et al.
Publicado: (2025)
Separate to Collaborate: Dual-Stream Diffusion Model for Coordinated Piano Hand Motion Synthesis
por: Liu, Zihao, et al.
Publicado: (2025)
por: Liu, Zihao, et al.
Publicado: (2025)
The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge
por: Wu, Leyi, et al.
Publicado: (2026)
por: Wu, Leyi, et al.
Publicado: (2026)
Pseudo Anomalies Are All You Need: Diffusion-Based Generation for Weakly-Supervised Video Anomaly Detection
por: Hashimoto, Satoshi, et al.
Publicado: (2025)
por: Hashimoto, Satoshi, et al.
Publicado: (2025)
Pairwise Comparisons Are All You Need
por: Chahine, Nicolas, et al.
Publicado: (2024)
por: Chahine, Nicolas, et al.
Publicado: (2024)
PoseAnything: Universal Pose-guided Video Generation with Part-aware Temporal Coherence
por: Wang, Ruiyan, et al.
Publicado: (2025)
por: Wang, Ruiyan, et al.
Publicado: (2025)
Is Discretization Fusion All You Need for Collaborative Perception?
por: Yang, Kang, et al.
Publicado: (2025)
por: Yang, Kang, et al.
Publicado: (2025)
Performance is not All You Need: Sustainability Considerations for Algorithms
por: Li, Xiang, et al.
Publicado: (2025)
por: Li, Xiang, et al.
Publicado: (2025)
Moving Object Segmentation: All You Need Is SAM (and Flow)
por: Xie, Junyu, et al.
Publicado: (2024)
por: Xie, Junyu, et al.
Publicado: (2024)
ParameterNet: Parameters Are All You Need
por: Han, Kai, et al.
Publicado: (2023)
por: Han, Kai, et al.
Publicado: (2023)
Fast Wrong-way Cycling Detection in CCTV Videos: Sparse Sampling is All You Need
por: Xu, Jing, et al.
Publicado: (2024)
por: Xu, Jing, et al.
Publicado: (2024)
KVPO: ODE-Native GRPO for Autoregressive Video Alignment via KV Semantic Exploration
por: Zhang, Ruicheng, et al.
Publicado: (2026)
por: Zhang, Ruicheng, et al.
Publicado: (2026)
One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer
por: Shi, Shijun, et al.
Publicado: (2025)
por: Shi, Shijun, et al.
Publicado: (2025)
VideoMerge: Towards Training-free Long Video Generation
por: Zhang, Siyang, et al.
Publicado: (2025)
por: Zhang, Siyang, et al.
Publicado: (2025)
One Snapshot is All You Need: A Generalized Method for mmWave Signal Generation
por: Huang, Teng, et al.
Publicado: (2025)
por: Huang, Teng, et al.
Publicado: (2025)
Zo3T: Zero-Shot 3D-Aware Trajectory-Guided Image-to-Video Generation via Test-Time Training
por: Zhang, Ruicheng, et al.
Publicado: (2025)
por: Zhang, Ruicheng, et al.
Publicado: (2025)
Inpainting is All You Need: A Diffusion-based Augmentation Method for Semi-supervised Medical Image Segmentation
por: Hu, Xinrong, et al.
Publicado: (2025)
por: Hu, Xinrong, et al.
Publicado: (2025)
[MASK] is All You Need
por: Hu, Vincent Tao, et al.
Publicado: (2024)
por: Hu, Vincent Tao, et al.
Publicado: (2024)
Grounding is All You Need? Dual Temporal Grounding for Video Dialog
por: Qin, You, et al.
Publicado: (2024)
por: Qin, You, et al.
Publicado: (2024)
LLMRA: Multi-modal Large Language Model based Restoration Assistant
por: Jin, Xiaoyu, et al.
Publicado: (2024)
por: Jin, Xiaoyu, et al.
Publicado: (2024)
Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton
por: Kang, Hongbo, et al.
Publicado: (2024)
por: Kang, Hongbo, et al.
Publicado: (2024)
Attention Is All You Need For Mixture-of-Depths Routing
por: Gadhikar, Advait, et al.
Publicado: (2024)
por: Gadhikar, Advait, et al.
Publicado: (2024)
STDAN: Deformable Attention Network for Space-Time Video Super-Resolution
por: Wang, Hai, et al.
Publicado: (2022)
por: Wang, Hai, et al.
Publicado: (2022)
Emu3: Next-Token Prediction is All You Need
por: Wang, Xinlong, et al.
Publicado: (2024)
por: Wang, Xinlong, et al.
Publicado: (2024)
Alignment-free Raw Video Demoireing
por: Xu, Shuning, et al.
Publicado: (2024)
por: Xu, Shuning, et al.
Publicado: (2024)
Distraction is All You Need for Multimodal Large Language Model Jailbreaking
por: Yang, Zuopeng, et al.
Publicado: (2025)
por: Yang, Zuopeng, et al.
Publicado: (2025)
REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment
por: Han, Haonan, et al.
Publicado: (2024)
por: Han, Haonan, et al.
Publicado: (2024)
Generalize Your Face Forgery Detectors: An Insertable Adaptation Module Is All You Need
por: Si, Xiaotian, et al.
Publicado: (2024)
por: Si, Xiaotian, et al.
Publicado: (2024)
CMTA: Cross-Modal Temporal Alignment for Event-guided Video Deblurring
por: Kim, Taewoo, et al.
Publicado: (2024)
por: Kim, Taewoo, et al.
Publicado: (2024)
Pose Magic: Efficient and Temporally Consistent Human Pose Estimation with a Hybrid Mamba-GCN Network
por: Zhang, Xinyi, et al.
Publicado: (2024)
por: Zhang, Xinyi, et al.
Publicado: (2024)
Zoom and Shift are All You Need
por: Qin, Jiahao
Publicado: (2024)
por: Qin, Jiahao
Publicado: (2024)
Unlearnable 3D Point Clouds: Class-wise Transformation Is All You Need
por: Wang, Xianlong, et al.
Publicado: (2024)
por: Wang, Xianlong, et al.
Publicado: (2024)
Unsupervised Real-World Denoising: Sparsity is All You Need
por: Chihaoui, Hamadi, et al.
Publicado: (2025)
por: Chihaoui, Hamadi, et al.
Publicado: (2025)
Search is All You Need for Few-shot Anomaly Detection
por: Wang, Qishan, et al.
Publicado: (2025)
por: Wang, Qishan, et al.
Publicado: (2025)
Exchange Is All You Need for Remote Sensing Change Detection
por: Dong, Sijun, et al.
Publicado: (2026)
por: Dong, Sijun, et al.
Publicado: (2026)
Positive Label Is All You Need for Multi-Label Classification
por: Yuan, Zhixiang, et al.
Publicado: (2023)
por: Yuan, Zhixiang, et al.
Publicado: (2023)
Ejemplares similares
-
Transferable-guided Attention Is All You Need for Video Domain Adaptation
por: Sacilotti, André, et al.
Publicado: (2024) -
Reasoning is All You Need for Video Generalization: A Counterfactual Benchmark with Sub-question Evaluation
por: Zhou, Qiji, et al.
Publicado: (2025) -
SAM-R1: Leveraging SAM for Reward Feedback in Multimodal Segmentation via Reinforcement Learning
por: Huang, Jiaqi, et al.
Publicado: (2025) -
All You Need to Know About Training Image Retrieval Models
por: Berton, Gabriele, et al.
Publicado: (2025) -
SVAC: Scaling Is All You Need For Referring Video Object Segmentation
por: Zhang, Li, et al.
Publicado: (2025)