CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration
Fuente:
arXiv
Guardado en:
| Autores principales: | Meric, Adil, Foo, Lin Geng, Kiray, Mert, Busam, Benjamin, Dabral, Rishabh, Theobalt, Christian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
por: Zhang, Wanyue, et al.
Publicado: (2025)
por: Zhang, Wanyue, et al.
Publicado: (2025)
Dropping the D: RGB-D SLAM Without the Depth Sensor
por: Kiray, Mert, et al.
Publicado: (2025)
por: Kiray, Mert, et al.
Publicado: (2025)
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
por: Ghosh, Anindita, et al.
Publicado: (2023)
por: Ghosh, Anindita, et al.
Publicado: (2023)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
por: Kiray, Mert, et al.
Publicado: (2025)
por: Kiray, Mert, et al.
Publicado: (2025)
PromptVFX: Text-Driven Fields for Open-World 3D Gaussian Animation
por: Kiray, Mert, et al.
Publicado: (2025)
por: Kiray, Mert, et al.
Publicado: (2025)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
por: Gopal, Suhas, et al.
Publicado: (2025)
por: Gopal, Suhas, et al.
Publicado: (2025)
MIBURI: Towards Expressive Interactive Gesture Synthesis
por: Mughal, M. Hamza, et al.
Publicado: (2026)
por: Mughal, M. Hamza, et al.
Publicado: (2026)
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
por: Ghosh, Anindita, et al.
Publicado: (2025)
por: Ghosh, Anindita, et al.
Publicado: (2025)
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
por: Collorone, Luca, et al.
Publicado: (2025)
por: Collorone, Luca, et al.
Publicado: (2025)
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
por: Ghosh, Anindita, et al.
Publicado: (2026)
por: Ghosh, Anindita, et al.
Publicado: (2026)
3D Segmentation Using Viewpoint-Dependent Spatial Relationships
por: Nanri, Ayaka, et al.
Publicado: (2026)
por: Nanri, Ayaka, et al.
Publicado: (2026)
Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions
por: Aytekin, Ayce Idil, et al.
Publicado: (2026)
por: Aytekin, Ayce Idil, et al.
Publicado: (2026)
Real-time Free-view Human Rendering from Sparse-view RGB Videos using Double Unprojected Textures
por: Sun, Guoxing, et al.
Publicado: (2024)
por: Sun, Guoxing, et al.
Publicado: (2024)
MetaCap: Meta-learning Priors from Multi-View Imagery for Sparse-view Human Performance Capture and Rendering
por: Sun, Guoxing, et al.
Publicado: (2024)
por: Sun, Guoxing, et al.
Publicado: (2024)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
UnReflectAnything: RGB-Only Highlight Removal by Rendering Synthetic Specular Supervision
por: Rota, Alberto, et al.
Publicado: (2025)
por: Rota, Alberto, et al.
Publicado: (2025)
PractiLight: Practical Light Control Using Foundational Diffusion Models
por: Erel, Yotam, et al.
Publicado: (2025)
por: Erel, Yotam, et al.
Publicado: (2025)
FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video
por: Camiletto, Andrea Boscolo, et al.
Publicado: (2025)
por: Camiletto, Andrea Boscolo, et al.
Publicado: (2025)
Giant Left Anterior Descending Coronary Artery Aneurysm in a Patient with Behçet’s Disease
por: Mert Meric
Publicado: (2022)
por: Mert Meric
Publicado: (2022)
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
por: Zhang, Wanyue, et al.
Publicado: (2023)
por: Zhang, Wanyue, et al.
Publicado: (2023)
MultiCOIN: Multi-Modal COntrollable Video INbetweening
por: Tanveer, Maham, et al.
Publicado: (2025)
por: Tanveer, Maham, et al.
Publicado: (2025)
BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects
por: Zhang, Wanyue, et al.
Publicado: (2024)
por: Zhang, Wanyue, et al.
Publicado: (2024)
Physical Simulator In-the-Loop Video Generation
por: Foo, Lin Geng, et al.
Publicado: (2026)
por: Foo, Lin Geng, et al.
Publicado: (2026)
Attention (as Discrete-Time Markov) Chains
por: Erel, Yotam, et al.
Publicado: (2025)
por: Erel, Yotam, et al.
Publicado: (2025)
Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis
por: Mughal, M. Hamza, et al.
Publicado: (2024)
por: Mughal, M. Hamza, et al.
Publicado: (2024)
FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
por: Delitzas, Alexandros, et al.
Publicado: (2026)
por: Delitzas, Alexandros, et al.
Publicado: (2026)
Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input
por: Wang, Jian, et al.
Publicado: (2025)
por: Wang, Jian, et al.
Publicado: (2025)
Relightable Holoported Characters: Capturing and Relighting Dynamic Human Performance from Sparse Views
por: Singh, Kunwar Maheep, et al.
Publicado: (2025)
por: Singh, Kunwar Maheep, et al.
Publicado: (2025)
PocoLoco: A Point Cloud Diffusion Model of Human Shape in Loose Clothing
por: Seth, Siddharth, et al.
Publicado: (2024)
por: Seth, Siddharth, et al.
Publicado: (2024)
Physics-based Human Pose Estimation from a Single Moving RGB Camera
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
GENMO: A GENeralist Model for Human MOtion
por: Li, Jiefeng, et al.
Publicado: (2025)
por: Li, Jiefeng, et al.
Publicado: (2025)
Matrix-free Second-order Optimization of Gaussian Splats with Residual Sampling
por: Pehlivan, Hamza, et al.
Publicado: (2025)
por: Pehlivan, Hamza, et al.
Publicado: (2025)
FLex: Joint Pose and Dynamic Radiance Fields Optimization for Stereo Endoscopic Videos
por: Stilz, Florian Philipp, et al.
Publicado: (2024)
por: Stilz, Florian Philipp, et al.
Publicado: (2024)
LLM-based Agentic Reasoning Frameworks: A Survey from Methods to Scenarios
por: Zhao, Bingxi, et al.
Publicado: (2025)
por: Zhao, Bingxi, et al.
Publicado: (2025)
OnlineSplatter: Pose-Free Online 3D Reconstruction for Free-Moving Objects
por: Huang, Mark He, et al.
Publicado: (2025)
por: Huang, Mark He, et al.
Publicado: (2025)
VUGEN: Visual Understanding priors for GENeration
por: Chen, Xiangyi, et al.
Publicado: (2025)
por: Chen, Xiangyi, et al.
Publicado: (2025)
Avatar Concept Slider: Controllable Editing of Concepts in 3D Human Avatars
por: Foo, Lin Geng, et al.
Publicado: (2024)
por: Foo, Lin Geng, et al.
Publicado: (2024)
3HANDS Dataset: Learning from Humans for Generating Naturalistic Handovers with Supernumerary Robotic Limbs
por: Abadian, Artin Saberpour, et al.
Publicado: (2025)
por: Abadian, Artin Saberpour, et al.
Publicado: (2025)
PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes
por: Dabral, Shailendra
Publicado: (2026)
por: Dabral, Shailendra
Publicado: (2026)
Ejemplares similares
-
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
por: Zhang, Wanyue, et al.
Publicado: (2025) -
Dropping the D: RGB-D SLAM Without the Depth Sensor
por: Kiray, Mert, et al.
Publicado: (2025) -
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
por: Ghosh, Anindita, et al.
Publicado: (2023) -
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
por: Kiray, Mert, et al.
Publicado: (2025) -
PromptVFX: Text-Driven Fields for Open-World 3D Gaussian Animation
por: Kiray, Mert, et al.
Publicado: (2025)