Salvato in:
| Autori principali: | Chen, Jiahao, Yuan, Hangjie, Qian, Yichen, Liang, Jingyun, Xing, Jiazheng, Liu, Pengwei, Chen, Weihua, Wang, Fan, Su, Bing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2506.02497 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective
di: Yuan, Hangjie, et al.
Pubblicazione: (2025)
di: Yuan, Hangjie, et al.
Pubblicazione: (2025)
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
di: Liu, Ropeway, et al.
Pubblicazione: (2025)
di: Liu, Ropeway, et al.
Pubblicazione: (2025)
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
di: Liang, Jingyun, et al.
Pubblicazione: (2025)
di: Liang, Jingyun, et al.
Pubblicazione: (2025)
Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization
di: Liang, Jingyun, et al.
Pubblicazione: (2026)
di: Liang, Jingyun, et al.
Pubblicazione: (2026)
MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems
di: Chen, Shuhang, et al.
Pubblicazione: (2025)
di: Chen, Shuhang, et al.
Pubblicazione: (2025)
Knowledge is Power: Advancing Few-shot Action Recognition with Multimodal Semantics from MLLMs
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
di: Xing, Jiazheng, et al.
Pubblicazione: (2026)
SciLT: Long-tailed Image Classification under Scientific Image Domains
di: Chen, Jiahao, et al.
Pubblicazione: (2026)
di: Chen, Jiahao, et al.
Pubblicazione: (2026)
MoVideo: Motion-Aware Video Generation with Diffusion Models
di: Liang, Jingyun, et al.
Pubblicazione: (2023)
di: Liang, Jingyun, et al.
Pubblicazione: (2023)
Towards Reason-Informed Video Editing in Unified Models with Self-Reflective Learning
di: Liu, Xinyu, et al.
Pubblicazione: (2025)
di: Liu, Xinyu, et al.
Pubblicazione: (2025)
LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios
di: Huang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Huang, Zhiyuan, et al.
Pubblicazione: (2025)
SAMora: Enhancing SAM through Hierarchical Self-Supervised Pre-Training for Medical Images
di: Chen, Shuhang, et al.
Pubblicazione: (2025)
di: Chen, Shuhang, et al.
Pubblicazione: (2025)
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
di: Gan, Qijun, et al.
Pubblicazione: (2025)
di: Gan, Qijun, et al.
Pubblicazione: (2025)
Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
Rethinking the Bias of Foundation Model under Long-tailed Distribution
di: Chen, Jiahao, et al.
Pubblicazione: (2025)
di: Chen, Jiahao, et al.
Pubblicazione: (2025)
PanFlow: Decoupled Motion Control for Panoramic Video Generation
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
FlexiFilm: Long Video Generation with Flexible Conditions
di: Ouyang, Yichen, et al.
Pubblicazione: (2024)
di: Ouyang, Yichen, et al.
Pubblicazione: (2024)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
di: Cai, Lingling, et al.
Pubblicazione: (2025)
di: Cai, Lingling, et al.
Pubblicazione: (2025)
Motion Semantics Guided Normalizing Flow for Privacy-Preserving Video Anomaly Detection
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning
di: Wei, Yujie, et al.
Pubblicazione: (2026)
di: Wei, Yujie, et al.
Pubblicazione: (2026)
Domain-adaptive and Subgroup-specific Cascaded Temperature Regression for Out-of-distribution Calibration
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
FlowMotion: Training-Free Flow Guidance for Video Motion Transfer
di: Wang, Zhen, et al.
Pubblicazione: (2026)
di: Wang, Zhen, et al.
Pubblicazione: (2026)
T-CorresNet: Template Guided 3D Point Cloud Completion with Correspondence Pooling Query Generation Strategy
di: Duan, Fan, et al.
Pubblicazione: (2024)
di: Duan, Fan, et al.
Pubblicazione: (2024)
LTCA: Long-range Temporal Context Attention for Referring Video Object Segmentation
di: Yan, Cilin, et al.
Pubblicazione: (2025)
di: Yan, Cilin, et al.
Pubblicazione: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
di: Wang, Yiming, et al.
Pubblicazione: (2026)
di: Wang, Yiming, et al.
Pubblicazione: (2026)
AMG: Avatar Motion Guided Video Generation
di: Yang, Zhangsihao, et al.
Pubblicazione: (2024)
di: Yang, Zhangsihao, et al.
Pubblicazione: (2024)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
di: Xu, Jiahao, et al.
Pubblicazione: (2026)
di: Xu, Jiahao, et al.
Pubblicazione: (2026)
MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation
di: Liu, Yanchen, et al.
Pubblicazione: (2025)
di: Liu, Yanchen, et al.
Pubblicazione: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
di: Wei, Yujie, et al.
Pubblicazione: (2024)
di: Wei, Yujie, et al.
Pubblicazione: (2024)
LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation
di: Gao, Jianxiong, et al.
Pubblicazione: (2025)
di: Gao, Jianxiong, et al.
Pubblicazione: (2025)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
di: Li, Xinyang, et al.
Pubblicazione: (2025)
di: Li, Xinyang, et al.
Pubblicazione: (2025)
MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequence
di: Zhao, Canyu, et al.
Pubblicazione: (2024)
di: Zhao, Canyu, et al.
Pubblicazione: (2024)
Decentralized Gossip Mutual Learning (GML) for brain tumor segmentation on multi-parametric MRI
di: Chen, Jingyun, et al.
Pubblicazione: (2024)
di: Chen, Jingyun, et al.
Pubblicazione: (2024)
Flow-Guided Diffusion for Video Inpainting
di: Gu, Bohai, et al.
Pubblicazione: (2023)
di: Gu, Bohai, et al.
Pubblicazione: (2023)
SAMA: Factorized Semantic Anchoring and Motion Alignment for Instruction-Guided Video Editing
di: Zhang, Xinyao, et al.
Pubblicazione: (2026)
di: Zhang, Xinyao, et al.
Pubblicazione: (2026)
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
di: Jin, Peng, et al.
Pubblicazione: (2024)
di: Jin, Peng, et al.
Pubblicazione: (2024)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
di: Liu, Huijie, et al.
Pubblicazione: (2025)
di: Liu, Huijie, et al.
Pubblicazione: (2025)
VideoMAR: Autoregressive Video Generatio with Continuous Tokens
di: Yu, Hu, et al.
Pubblicazione: (2025)
di: Yu, Hu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
di: Xing, Jiazheng, et al.
Pubblicazione: (2026) -
Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective
di: Yuan, Hangjie, et al.
Pubblicazione: (2025) -
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
di: Liu, Ropeway, et al.
Pubblicazione: (2025) -
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models
di: Xing, Jiazheng, et al.
Pubblicazione: (2026) -
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
di: Liang, Jingyun, et al.
Pubblicazione: (2025)