Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yue, He, Yingqing, Wang, Hongfa, Wang, Andong, Qi, Chenyang, Cai, Chengfei, Li, Xiu, Li, Zhifeng, Shum, Heung-Yeung, Liu, Wei, Chen, Qifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
von: Ma, Yue, et al.
Veröffentlicht: (2024)
von: Ma, Yue, et al.
Veröffentlicht: (2024)
Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling
von: Xue, Jingyun, et al.
Veröffentlicht: (2024)
von: Xue, Jingyun, et al.
Veröffentlicht: (2024)
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
von: Ma, Yue, et al.
Veröffentlicht: (2023)
von: Ma, Yue, et al.
Veröffentlicht: (2023)
Large Investment Model
von: Guo, Jian, et al.
Veröffentlicht: (2024)
von: Guo, Jian, et al.
Veröffentlicht: (2024)
Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation
von: Chen, Qihua, et al.
Veröffentlicht: (2024)
von: Chen, Qihua, et al.
Veröffentlicht: (2024)
Seeing and Hearing: Open-domain Visual-Audio Generation with Diffusion Latent Aligners
von: Xing, Yazhou, et al.
Veröffentlicht: (2024)
von: Xing, Yazhou, et al.
Veröffentlicht: (2024)
Follow-Your-Motion: Video Motion Transfer via Efficient Spatial-Temporal Decoupled Finetuning
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
von: Hu, Jingcheng, et al.
Veröffentlicht: (2025)
von: Hu, Jingcheng, et al.
Veröffentlicht: (2025)
Follow-Your-Color: Multi-Instance Sketch Colorization
von: Zhang, Yinhan, et al.
Veröffentlicht: (2025)
von: Zhang, Yinhan, et al.
Veröffentlicht: (2025)
SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields
von: Kong, Youkang, et al.
Veröffentlicht: (2026)
von: Kong, Youkang, et al.
Veröffentlicht: (2026)
NoiseAR: AutoRegressing Initial Noise Prior for Diffusion Models
von: Li, Zeming, et al.
Veröffentlicht: (2025)
von: Li, Zeming, et al.
Veröffentlicht: (2025)
AnimationBench: Are Video Models Good at Character-Centric Animation?
von: Wu, Leyi, et al.
Veröffentlicht: (2026)
von: Wu, Leyi, et al.
Veröffentlicht: (2026)
MultiBooth: Towards Generating All Your Concepts in an Image from Text
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
Controllable Video Generation: A Survey
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
MagicStick: Controllable Video Editing via Control Handle Transformations
von: Ma, Yue, et al.
Veröffentlicht: (2023)
von: Ma, Yue, et al.
Veröffentlicht: (2023)
Adaptive Domain Learning for Cross-domain Image Denoising
von: Qian, Zian, et al.
Veröffentlicht: (2024)
von: Qian, Zian, et al.
Veröffentlicht: (2024)
Follow-Your-Shape: Shape-Aware Image Editing via Trajectory-Guided Region Control
von: Long, Zeqian, et al.
Veröffentlicht: (2025)
von: Long, Zeqian, et al.
Veröffentlicht: (2025)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
Puppeteer: Rig and Animate Your 3D Models
von: Song, Chaoyue, et al.
Veröffentlicht: (2025)
von: Song, Chaoyue, et al.
Veröffentlicht: (2025)
HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
Alpha-GPT: Human-AI Interactive Alpha Mining for Quantitative Investment
von: Wang, Saizhuo, et al.
Veröffentlicht: (2023)
von: Wang, Saizhuo, et al.
Veröffentlicht: (2023)
Multi-matrix Factorization Attention
von: Hu, Jingcheng, et al.
Veröffentlicht: (2024)
von: Hu, Jingcheng, et al.
Veröffentlicht: (2024)
GaussiAnimate: Reconstruct and Rig Animatable Categories with Level of Dynamics
von: Wang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Wang, Jiaxin, et al.
Veröffentlicht: (2026)
HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation
von: Xu, Zunnan, et al.
Veröffentlicht: (2025)
von: Xu, Zunnan, et al.
Veröffentlicht: (2025)
Instruction-based Image Editing with Planning, Reasoning, and Generation
von: Ji, Liya, et al.
Veröffentlicht: (2026)
von: Ji, Liya, et al.
Veröffentlicht: (2026)
InstanceAnimator: Multi-Instance Sketch Video Colorization
von: Zhang, Yinhan, et al.
Veröffentlicht: (2026)
von: Zhang, Yinhan, et al.
Veröffentlicht: (2026)
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
InstantSwap: Fast Customized Concept Swapping across Sharp Shape Differences
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
Follow-Your-Creation: Empowering 4D Creation through Video Inpainting
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
SPIRE: Semantic Prompt-Driven Image Restoration
von: Qi, Chenyang, et al.
Veröffentlicht: (2023)
von: Qi, Chenyang, et al.
Veröffentlicht: (2023)
ClickAttention: Click Region Similarity Guided Interactive Segmentation
von: Xu, Long, et al.
Veröffentlicht: (2024)
von: Xu, Long, et al.
Veröffentlicht: (2024)
The Vertical Challenge of Low-Altitude Economy: Why We Need a Unified Height System?
von: Yan, Shuaichen, et al.
Veröffentlicht: (2026)
von: Yan, Shuaichen, et al.
Veröffentlicht: (2026)
Follow-Your-Instruction: A Comprehensive MLLM Agent for World Data Synthesis
von: Feng, Kunyu, et al.
Veröffentlicht: (2025)
von: Feng, Kunyu, et al.
Veröffentlicht: (2025)
DiT4Edit: Diffusion Transformer for Image Editing
von: Feng, Kunyu, et al.
Veröffentlicht: (2024)
von: Feng, Kunyu, et al.
Veröffentlicht: (2024)
Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
Open Benchmarking for Click-Through Rate Prediction
von: Zhu, Jieming, et al.
Veröffentlicht: (2020)
von: Zhu, Jieming, et al.
Veröffentlicht: (2020)
Automatic Controllable Colorization via Imagination
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2024)
Deep Pattern Network for Click-Through Rate Prediction
von: Zhang, Hengyu, et al.
Veröffentlicht: (2024)
von: Zhang, Hengyu, et al.
Veröffentlicht: (2024)
HMDN: Hierarchical Multi-Distribution Network for Click-Through Rate Prediction
von: Lou, Xingyu, et al.
Veröffentlicht: (2024)
von: Lou, Xingyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
von: Ma, Yue, et al.
Veröffentlicht: (2024) -
Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
von: Ma, Yue, et al.
Veröffentlicht: (2025) -
Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling
von: Xue, Jingyun, et al.
Veröffentlicht: (2024) -
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
von: Ma, Yue, et al.
Veröffentlicht: (2023) -
Large Investment Model
von: Guo, Jian, et al.
Veröffentlicht: (2024)