Guardado en:
| Autores principales: | Chen, Jiahao, Yuan, Hangjie, Qian, Yichen, Liang, Jingyun, Xing, Jiazheng, Liu, Pengwei, Chen, Weihua, Wang, Fan, Su, Bing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2506.02497 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
por: Xing, Jiazheng, et al.
Publicado: (2026)
por: Xing, Jiazheng, et al.
Publicado: (2026)
Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective
por: Yuan, Hangjie, et al.
Publicado: (2025)
por: Yuan, Hangjie, et al.
Publicado: (2025)
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
por: Liu, Ropeway, et al.
Publicado: (2025)
por: Liu, Ropeway, et al.
Publicado: (2025)
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models
por: Xing, Jiazheng, et al.
Publicado: (2026)
por: Xing, Jiazheng, et al.
Publicado: (2026)
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
por: Liang, Jingyun, et al.
Publicado: (2025)
por: Liang, Jingyun, et al.
Publicado: (2025)
Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization
por: Liang, Jingyun, et al.
Publicado: (2026)
por: Liang, Jingyun, et al.
Publicado: (2026)
MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems
por: Chen, Shuhang, et al.
Publicado: (2025)
por: Chen, Shuhang, et al.
Publicado: (2025)
Knowledge is Power: Advancing Few-shot Action Recognition with Multimodal Semantics from MLLMs
por: Xing, Jiazheng, et al.
Publicado: (2026)
por: Xing, Jiazheng, et al.
Publicado: (2026)
SciLT: Long-tailed Image Classification under Scientific Image Domains
por: Chen, Jiahao, et al.
Publicado: (2026)
por: Chen, Jiahao, et al.
Publicado: (2026)
MoVideo: Motion-Aware Video Generation with Diffusion Models
por: Liang, Jingyun, et al.
Publicado: (2023)
por: Liang, Jingyun, et al.
Publicado: (2023)
Towards Reason-Informed Video Editing in Unified Models with Self-Reflective Learning
por: Liu, Xinyu, et al.
Publicado: (2025)
por: Liu, Xinyu, et al.
Publicado: (2025)
LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios
por: Huang, Zhiyuan, et al.
Publicado: (2025)
por: Huang, Zhiyuan, et al.
Publicado: (2025)
SAMora: Enhancing SAM through Hierarchical Self-Supervised Pre-Training for Medical Images
por: Chen, Shuhang, et al.
Publicado: (2025)
por: Chen, Shuhang, et al.
Publicado: (2025)
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
por: Gan, Qijun, et al.
Publicado: (2025)
por: Gan, Qijun, et al.
Publicado: (2025)
Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
por: Cao, Chenjie, et al.
Publicado: (2025)
por: Cao, Chenjie, et al.
Publicado: (2025)
Rethinking the Bias of Foundation Model under Long-tailed Distribution
por: Chen, Jiahao, et al.
Publicado: (2025)
por: Chen, Jiahao, et al.
Publicado: (2025)
PanFlow: Decoupled Motion Control for Panoramic Video Generation
por: Zhang, Cheng, et al.
Publicado: (2025)
por: Zhang, Cheng, et al.
Publicado: (2025)
FlexiFilm: Long Video Generation with Flexible Conditions
por: Ouyang, Yichen, et al.
Publicado: (2024)
por: Ouyang, Yichen, et al.
Publicado: (2024)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
por: Cai, Lingling, et al.
Publicado: (2025)
por: Cai, Lingling, et al.
Publicado: (2025)
Motion Semantics Guided Normalizing Flow for Privacy-Preserving Video Anomaly Detection
por: Liu, Yang, et al.
Publicado: (2026)
por: Liu, Yang, et al.
Publicado: (2026)
DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning
por: Wei, Yujie, et al.
Publicado: (2026)
por: Wei, Yujie, et al.
Publicado: (2026)
Domain-adaptive and Subgroup-specific Cascaded Temperature Regression for Out-of-distribution Calibration
por: Wang, Jiexin, et al.
Publicado: (2024)
por: Wang, Jiexin, et al.
Publicado: (2024)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
por: Yang, Yuxiao, et al.
Publicado: (2025)
por: Yang, Yuxiao, et al.
Publicado: (2025)
FlowMotion: Training-Free Flow Guidance for Video Motion Transfer
por: Wang, Zhen, et al.
Publicado: (2026)
por: Wang, Zhen, et al.
Publicado: (2026)
T-CorresNet: Template Guided 3D Point Cloud Completion with Correspondence Pooling Query Generation Strategy
por: Duan, Fan, et al.
Publicado: (2024)
por: Duan, Fan, et al.
Publicado: (2024)
LTCA: Long-range Temporal Context Attention for Referring Video Object Segmentation
por: Yan, Cilin, et al.
Publicado: (2025)
por: Yan, Cilin, et al.
Publicado: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
por: Wang, Yiming, et al.
Publicado: (2026)
por: Wang, Yiming, et al.
Publicado: (2026)
AMG: Avatar Motion Guided Video Generation
por: Yang, Zhangsihao, et al.
Publicado: (2024)
por: Yang, Zhangsihao, et al.
Publicado: (2024)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
por: Xu, Jiahao, et al.
Publicado: (2026)
por: Xu, Jiahao, et al.
Publicado: (2026)
MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation
por: Liu, Yanchen, et al.
Publicado: (2025)
por: Liu, Yanchen, et al.
Publicado: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
por: Wei, Yujie, et al.
Publicado: (2024)
por: Wei, Yujie, et al.
Publicado: (2024)
LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation
por: Gao, Jianxiong, et al.
Publicado: (2025)
por: Gao, Jianxiong, et al.
Publicado: (2025)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
por: Li, Xinyang, et al.
Publicado: (2025)
por: Li, Xinyang, et al.
Publicado: (2025)
MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequence
por: Zhao, Canyu, et al.
Publicado: (2024)
por: Zhao, Canyu, et al.
Publicado: (2024)
Decentralized Gossip Mutual Learning (GML) for brain tumor segmentation on multi-parametric MRI
por: Chen, Jingyun, et al.
Publicado: (2024)
por: Chen, Jingyun, et al.
Publicado: (2024)
Flow-Guided Diffusion for Video Inpainting
por: Gu, Bohai, et al.
Publicado: (2023)
por: Gu, Bohai, et al.
Publicado: (2023)
SAMA: Factorized Semantic Anchoring and Motion Alignment for Instruction-Guided Video Editing
por: Zhang, Xinyao, et al.
Publicado: (2026)
por: Zhang, Xinyao, et al.
Publicado: (2026)
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
por: Jin, Peng, et al.
Publicado: (2024)
por: Jin, Peng, et al.
Publicado: (2024)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
por: Liu, Huijie, et al.
Publicado: (2025)
por: Liu, Huijie, et al.
Publicado: (2025)
VideoMAR: Autoregressive Video Generatio with Continuous Tokens
por: Yu, Hu, et al.
Publicado: (2025)
por: Yu, Hu, et al.
Publicado: (2025)
Ejemplares similares
-
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
por: Xing, Jiazheng, et al.
Publicado: (2026) -
Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective
por: Yuan, Hangjie, et al.
Publicado: (2025) -
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
por: Liu, Ropeway, et al.
Publicado: (2025) -
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models
por: Xing, Jiazheng, et al.
Publicado: (2026) -
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
por: Liang, Jingyun, et al.
Publicado: (2025)