SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Jungbin, Kim, Minsu, Kim, Jisoo, Zheng, Ce, Jeni, Laszlo A., Yang, Ming-Hsuan, Yu, Youngjae, Kim, Seonjoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
by: Kim, Jisoo, et al.
Published: (2024)
by: Kim, Jisoo, et al.
Published: (2024)
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
by: Cho, Jungbin, et al.
Published: (2024)
by: Cho, Jungbin, et al.
Published: (2024)
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025)
by: Kim, Junhyeok, et al.
Published: (2025)
Space-Time Forecasting of Dynamic Scenes with Motion-aware Gaussian Grouping
by: Lee, Junmyeong, et al.
Published: (2026)
by: Lee, Junmyeong, et al.
Published: (2026)
ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
DiSRT-In-Bed: Diffusion-Based Sim-to-Real Transfer Framework for In-Bed Human Mesh Recovery
by: Gao, Jing, et al.
Published: (2025)
by: Gao, Jing, et al.
Published: (2025)
GHOST: Grounded Human Motion Generation with Open Vocabulary Scene-and-Text Contexts
by: Milacski, Zoltán Á., et al.
Published: (2024)
by: Milacski, Zoltán Á., et al.
Published: (2024)
Pri4R: Learning World Dynamics for Vision-Language-Action Models with Privileged 4D Representation
by: Kim, Jisoo, et al.
Published: (2026)
by: Kim, Jisoo, et al.
Published: (2026)
SceneMI: Motion In-betweening for Modeling Human-Scene Interactions
by: Hwang, Inwoo, et al.
Published: (2025)
by: Hwang, Inwoo, et al.
Published: (2025)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024)
by: Kim, Gwanghyun, et al.
Published: (2024)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
Zero-Shot Scene Change Detection
by: Cho, Kyusik, et al.
Published: (2024)
by: Cho, Kyusik, et al.
Published: (2024)
SceneLinker: Compositional 3D Scene Generation via Semantic Scene Graph from RGB Sequences
by: Kim, Seok-Young, et al.
Published: (2026)
by: Kim, Seok-Young, et al.
Published: (2026)
A Unified Diffusion Framework for Scene-aware Human Motion Estimation from Sparse Signals
by: Tang, Jiangnan, et al.
Published: (2024)
by: Tang, Jiangnan, et al.
Published: (2024)
Locality-Aware Zero-Shot Human-Object Interaction Detection
by: Kim, Sanghyun, et al.
Published: (2025)
by: Kim, Sanghyun, et al.
Published: (2025)
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
by: Kim, Seungwook, et al.
Published: (2025)
by: Kim, Seungwook, et al.
Published: (2025)
Towards Holistic Surgical Scene Graph
by: Shin, Jongmin, et al.
Published: (2025)
by: Shin, Jongmin, et al.
Published: (2025)
Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis
by: Gong, Jingyu, et al.
Published: (2024)
by: Gong, Jingyu, et al.
Published: (2024)
4Real: Towards Photorealistic 4D Scene Generation via Video Diffusion Models
by: Yu, Heng, et al.
Published: (2024)
by: Yu, Heng, et al.
Published: (2024)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
by: Zhao, Yiwen, et al.
Published: (2026)
by: Zhao, Yiwen, et al.
Published: (2026)
RefFusion: Reference Adapted Diffusion Models for 3D Scene Inpainting
by: Mirzaei, Ashkan, et al.
Published: (2024)
by: Mirzaei, Ashkan, et al.
Published: (2024)
Object-aware Sound Source Localization via Audio-Visual Scene Understanding
by: Um, Sung Jin, et al.
Published: (2025)
by: Um, Sung Jin, et al.
Published: (2025)
Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation
by: Chen, Mu, et al.
Published: (2023)
by: Chen, Mu, et al.
Published: (2023)
Int3DNet: Scene-Motion Cross Attention Network for 3D Intention Prediction in Mixed Reality
by: Ha, Taewook, et al.
Published: (2026)
by: Ha, Taewook, et al.
Published: (2026)
Through the Curved Cover: Synthesizing Cover Aberrated Scenes with Refractive Field
by: Xie, Liuyue, et al.
Published: (2024)
by: Xie, Liuyue, et al.
Published: (2024)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
by: Kim, Seungwook, et al.
Published: (2026)
by: Kim, Seungwook, et al.
Published: (2026)
Scene-aware Human Motion Forecasting via Mutual Distance Prediction
by: Xing, Chaoyue, et al.
Published: (2023)
by: Xing, Chaoyue, et al.
Published: (2023)
LaserHuman: Language-guided Scene-aware Human Motion Generation in Free Environment
by: Cong, Peishan, et al.
Published: (2024)
by: Cong, Peishan, et al.
Published: (2024)
Tex4D: Zero-shot 4D Scene Texturing with Video Diffusion Models
by: Bao, Jingzhi, et al.
Published: (2024)
by: Bao, Jingzhi, et al.
Published: (2024)
HUMOF: Human Motion Forecasting in Interactive Social Scenes
by: Sun, Caiyi, et al.
Published: (2025)
by: Sun, Caiyi, et al.
Published: (2025)
Tri-Prompting: Video Diffusion with Unified Control over Scene, Subject, and Motion
by: Zhou, Zhenghong, et al.
Published: (2026)
by: Zhou, Zhenghong, et al.
Published: (2026)
Motion Cues from Image-based Point Tracking for LiDAR Scene Flow Estimation
by: Jang, Youngdong, et al.
Published: (2026)
by: Jang, Youngdong, et al.
Published: (2026)
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
by: Xing, Chaoyue, et al.
Published: (2026)
by: Xing, Chaoyue, et al.
Published: (2026)
Towards Generalizable Scene Change Detection
by: Kim, Jaewoo, et al.
Published: (2024)
by: Kim, Jaewoo, et al.
Published: (2024)
Multi-Condition Latent Diffusion Network for Scene-Aware Neural Human Motion Prediction
by: Gao, Xuehao, et al.
Published: (2024)
by: Gao, Xuehao, et al.
Published: (2024)
Pyramid Diffusion for Fine 3D Large Scene Generation
by: Liu, Yuheng, et al.
Published: (2023)
by: Liu, Yuheng, et al.
Published: (2023)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
by: Kim, Youngmin, et al.
Published: (2025)
by: Kim, Youngmin, et al.
Published: (2025)
Generic Event Boundary Detection via Denoising Diffusion
by: Hwang, Jaejun, et al.
Published: (2025)
by: Hwang, Jaejun, et al.
Published: (2025)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Similar Items
-
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
by: Kim, Jisoo, et al.
Published: (2024) -
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
by: Cho, Jungbin, et al.
Published: (2024) -
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025) -
Space-Time Forecasting of Dynamic Scenes with Motion-aware Gaussian Grouping
by: Lee, Junmyeong, et al.
Published: (2026) -
ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors
by: Kim, Minsu, et al.
Published: (2025)