SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
Fuente:
arXiv
Guardado en:
| Autores principales: | Ghosh, Anindita, Golyanik, Vladislav, Komura, Taku, Slusallek, Philipp, Theobalt, Christian, Dabral, Rishabh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
por: Ghosh, Anindita, et al.
Publicado: (2023)
por: Ghosh, Anindita, et al.
Publicado: (2023)
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
por: Ghosh, Anindita, et al.
Publicado: (2025)
por: Ghosh, Anindita, et al.
Publicado: (2025)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
por: Gopal, Suhas, et al.
Publicado: (2025)
por: Gopal, Suhas, et al.
Publicado: (2025)
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
por: Zhang, Wanyue, et al.
Publicado: (2023)
por: Zhang, Wanyue, et al.
Publicado: (2023)
PractiLight: Practical Light Control Using Foundational Diffusion Models
por: Erel, Yotam, et al.
Publicado: (2025)
por: Erel, Yotam, et al.
Publicado: (2025)
BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects
por: Zhang, Wanyue, et al.
Publicado: (2024)
por: Zhang, Wanyue, et al.
Publicado: (2024)
Attention (as Discrete-Time Markov) Chains
por: Erel, Yotam, et al.
Publicado: (2025)
por: Erel, Yotam, et al.
Publicado: (2025)
EmbodMocap: In-the-Wild 4D Human-Scene Reconstruction for Embodied Agents
por: Wang, Wenjia, et al.
Publicado: (2026)
por: Wang, Wenjia, et al.
Publicado: (2026)
3D Human Pose Perception from Egocentric Stereo Videos
por: Akada, Hiroyasu, et al.
Publicado: (2023)
por: Akada, Hiroyasu, et al.
Publicado: (2023)
Bring Your Rear Cameras for Egocentric 3D Human Pose Estimation
por: Akada, Hiroyasu, et al.
Publicado: (2025)
por: Akada, Hiroyasu, et al.
Publicado: (2025)
Relightable Holoported Characters: Capturing and Relighting Dynamic Human Performance from Sparse Views
por: Singh, Kunwar Maheep, et al.
Publicado: (2025)
por: Singh, Kunwar Maheep, et al.
Publicado: (2025)
EventEgo3D: 3D Human Motion Capture from Egocentric Event Streams
por: Millerdurai, Christen, et al.
Publicado: (2024)
por: Millerdurai, Christen, et al.
Publicado: (2024)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
por: Zhang, Wanyue, et al.
Publicado: (2025)
por: Zhang, Wanyue, et al.
Publicado: (2025)
Quantum Visual Fields with Neural Amplitude Encoding
por: Wang, Shuteng, et al.
Publicado: (2025)
por: Wang, Shuteng, et al.
Publicado: (2025)
MIBURI: Towards Expressive Interactive Gesture Synthesis
por: Mughal, M. Hamza, et al.
Publicado: (2026)
por: Mughal, M. Hamza, et al.
Publicado: (2026)
E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation
por: Deshmukh, Mayur, et al.
Publicado: (2026)
por: Deshmukh, Mayur, et al.
Publicado: (2026)
Dynamic EventNeRF: Reconstructing General Dynamic Scenes from Multi-view RGB and Event Streams
por: Rudnev, Viktor, et al.
Publicado: (2024)
por: Rudnev, Viktor, et al.
Publicado: (2024)
MetaCap: Meta-learning Priors from Multi-View Imagery for Sparse-view Human Performance Capture and Rendering
por: Sun, Guoxing, et al.
Publicado: (2024)
por: Sun, Guoxing, et al.
Publicado: (2024)
Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input
por: Wang, Jian, et al.
Publicado: (2025)
por: Wang, Jian, et al.
Publicado: (2025)
EventEgo3D++: 3D Human Motion Capture from a Head-Mounted Event Camera
por: Millerdurai, Christen, et al.
Publicado: (2025)
por: Millerdurai, Christen, et al.
Publicado: (2025)
NeuralClothSim: Neural Deformation Fields Meet the Thin Shell Theory
por: Kairanda, Navami, et al.
Publicado: (2023)
por: Kairanda, Navami, et al.
Publicado: (2023)
TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization
por: Pan, Liang, et al.
Publicado: (2025)
por: Pan, Liang, et al.
Publicado: (2025)
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration
por: Meric, Adil, et al.
Publicado: (2026)
por: Meric, Adil, et al.
Publicado: (2026)
EventNeuS: 3D Mesh Reconstruction from a Single Event Camera
por: Sachan, Shreyas, et al.
Publicado: (2026)
por: Sachan, Shreyas, et al.
Publicado: (2026)
Thin-Shell-SfT: Fine-Grained Monocular Non-rigid 3D Surface Tracking with Neural Deformation Fields
por: Kairanda, Navami, et al.
Publicado: (2025)
por: Kairanda, Navami, et al.
Publicado: (2025)
Real-time Free-view Human Rendering from Sparse-view RGB Videos using Double Unprojected Textures
por: Sun, Guoxing, et al.
Publicado: (2024)
por: Sun, Guoxing, et al.
Publicado: (2024)
Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis
por: Mughal, M. Hamza, et al.
Publicado: (2024)
por: Mughal, M. Hamza, et al.
Publicado: (2024)
HumanOLAT: A Large-Scale Dataset for Full-Body Human Relighting and Novel-View Synthesis
por: Teufel, Timo, et al.
Publicado: (2025)
por: Teufel, Timo, et al.
Publicado: (2025)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
Holoported Characters: Real-time Free-viewpoint Rendering of Humans from Sparse RGB Cameras
por: Shetty, Ashwath, et al.
Publicado: (2023)
por: Shetty, Ashwath, et al.
Publicado: (2023)
FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video
por: Camiletto, Andrea Boscolo, et al.
Publicado: (2025)
por: Camiletto, Andrea Boscolo, et al.
Publicado: (2025)
Relightable Neural Actor with Intrinsic Decomposition and Pose Control
por: Luvizon, Diogo, et al.
Publicado: (2023)
por: Luvizon, Diogo, et al.
Publicado: (2023)
TLControl: Trajectory and Language Control for Human Motion Synthesis
por: Wan, Weilin, et al.
Publicado: (2023)
por: Wan, Weilin, et al.
Publicado: (2023)
D-NPC: Dynamic Neural Point Clouds for Non-Rigid View Synthesis from Monocular Video
por: Kappel, Moritz, et al.
Publicado: (2024)
por: Kappel, Moritz, et al.
Publicado: (2024)
D‐NPC: Dynamic Neural Point Clouds for Non‐Rigid View Synthesis from Monocular Video
por: Moritz Kappel, et al.
Publicado: (2025)
por: Moritz Kappel, et al.
Publicado: (2025)
PocoLoco: A Point Cloud Diffusion Model of Human Shape in Loose Clothing
por: Seth, Siddharth, et al.
Publicado: (2024)
por: Seth, Siddharth, et al.
Publicado: (2024)
Physics-based Human Pose Estimation from a Single Moving RGB Camera
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
por: Aytekin, Ayce Idil, et al.
Publicado: (2025)
D-Rex : Diffusion Rendering for Relightable Expressive Avatars
por: Teufel, Timo, et al.
Publicado: (2026)
por: Teufel, Timo, et al.
Publicado: (2026)
Recent Trends in 3D Reconstruction of General Non-Rigid Scenes
por: Yunus, Raza, et al.
Publicado: (2024)
por: Yunus, Raza, et al.
Publicado: (2024)
Ejemplares similares
-
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
por: Ghosh, Anindita, et al.
Publicado: (2023) -
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
por: Ghosh, Anindita, et al.
Publicado: (2025) -
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
por: Gopal, Suhas, et al.
Publicado: (2025) -
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
por: Zhang, Wanyue, et al.
Publicado: (2023) -
PractiLight: Practical Light Control Using Foundational Diffusion Models
por: Erel, Yotam, et al.
Publicado: (2025)