Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Jingyu, Zhang, Chong, Liu, Fengqi, Fan, Ke, Zhou, Qianyu, Tan, Xin, Zhang, Zhizhong, Xie, Yuan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
by: Jingyu, Gong, et al.
Published: (2025)
by: Jingyu, Gong, et al.
Published: (2025)
DEMOS: Dynamic Environment Motion Synthesis in 3D Scenes via Local Spherical-BEV Perception
by: Gong, Jingyu, et al.
Published: (2024)
by: Gong, Jingyu, et al.
Published: (2024)
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
by: Zhang, Renhe, et al.
Published: (2026)
by: Zhang, Renhe, et al.
Published: (2026)
GSCompleter: A Distillation-Free Plugin for Metric-Aware 3D Gaussian Splatting Completion in Seconds
by: Gao, Ao, et al.
Published: (2026)
by: Gao, Ao, et al.
Published: (2026)
Continuous Piecewise-Affine Based Motion Model for Image Animation
by: Wang, Hexiang, et al.
Published: (2024)
by: Wang, Hexiang, et al.
Published: (2024)
PFDepth: Heterogeneous Pinhole-Fisheye Joint Depth Estimation via Distortion-aware Gaussian-Splatted Volumetric Fusion
by: Zhang, Zhiwei, et al.
Published: (2025)
by: Zhang, Zhiwei, et al.
Published: (2025)
World2Minecraft: Occupancy-Driven Simulated Scenes Construction
by: Zhang, Lechao, et al.
Published: (2026)
by: Zhang, Lechao, et al.
Published: (2026)
YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos
by: Chen, Haoming, et al.
Published: (2025)
by: Chen, Haoming, et al.
Published: (2025)
GEOcc: Geometrically Enhanced 3D Occupancy Network with Implicit-Explicit Depth Fusion and Contextual Self-Supervision
by: Tan, Xin, et al.
Published: (2024)
by: Tan, Xin, et al.
Published: (2024)
UniForward: Unified 3D Scene and Semantic Field Reconstruction via Feed-Forward Gaussian Splatting from Only Sparse-View Images
by: Tian, Qijian, et al.
Published: (2025)
by: Tian, Qijian, et al.
Published: (2025)
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
by: Fan, Ke, et al.
Published: (2024)
by: Fan, Ke, et al.
Published: (2024)
Vision-language models lag human performance on physical dynamics and intent reasoning
by: Gu, Tianjun, et al.
Published: (2026)
by: Gu, Tianjun, et al.
Published: (2026)
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
by: Fan, Ke, et al.
Published: (2024)
by: Fan, Ke, et al.
Published: (2024)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
by: Liu, Fengqi, et al.
Published: (2024)
by: Liu, Fengqi, et al.
Published: (2024)
Learning Implicit Neural Degradation Representation for Unpaired Image Dehazing
by: Fan, Shuaibin, et al.
Published: (2025)
by: Fan, Shuaibin, et al.
Published: (2025)
Beyond the Label Itself: Latent Labels Enhance Semi-supervised Point Cloud Panoptic Segmentation
by: Chen, Yujun, et al.
Published: (2023)
by: Chen, Yujun, et al.
Published: (2023)
Exploring the Untouched Sweeps for Conflict-Aware 3D Segmentation Pretraining
by: Sun, Tianfang, et al.
Published: (2024)
by: Sun, Tianfang, et al.
Published: (2024)
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
by: Cho, Jungbin, et al.
Published: (2025)
by: Cho, Jungbin, et al.
Published: (2025)
Decoupling Multi-Contrast Super-Resolution: Self-Supervised Implicit Re-Representation for Unpaired Cross-Modal Synthesis
by: Wu, Yinzhe, et al.
Published: (2025)
by: Wu, Yinzhe, et al.
Published: (2025)
Multi-modal In-Context Learning Makes an Ego-evolving Scene Text Recognizer
by: Zhao, Zhen, et al.
Published: (2023)
by: Zhao, Zhen, et al.
Published: (2023)
Building a Strong Pre-Training Baseline for Universal 3D Large-Scale Perception
by: Chen, Haoming, et al.
Published: (2024)
by: Chen, Haoming, et al.
Published: (2024)
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
by: Xing, Chaoyue, et al.
Published: (2026)
by: Xing, Chaoyue, et al.
Published: (2026)
COTR: Compact Occupancy TRansformer for Vision-based 3D Occupancy Prediction
by: Ma, Qihang, et al.
Published: (2023)
by: Ma, Qihang, et al.
Published: (2023)
Let Your Image Move with Your Motion! -- Implicit Multi-Object Multi-Motion Transfer
by: Li, Yuze, et al.
Published: (2026)
by: Li, Yuze, et al.
Published: (2026)
PathDiff: Histopathology Image Synthesis with Unpaired Text and Mask Conditions
by: Bhosale, Mahesh, et al.
Published: (2025)
by: Bhosale, Mahesh, et al.
Published: (2025)
Multi-Object Sketch Animation by Scene Decomposition and Motion Planning
by: Liu, Jingyu, et al.
Published: (2025)
by: Liu, Jingyu, et al.
Published: (2025)
PDF: Point Diffusion Implicit Function for Large-scale Scene Neural Representation
by: Ding, Yuhan, et al.
Published: (2023)
by: Ding, Yuhan, et al.
Published: (2023)
Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis
by: Phan, Vu Minh Hieu, et al.
Published: (2024)
by: Phan, Vu Minh Hieu, et al.
Published: (2024)
Mutual Information Guided Optimal Transport for Unsupervised Visible-Infrared Person Re-identification
by: Zhang, Zhizhong, et al.
Published: (2024)
by: Zhang, Zhizhong, et al.
Published: (2024)
IRIS: Intersection-aware Ray-based Implicit Editable Scenes
by: Wilczyński, Grzegorz, et al.
Published: (2026)
by: Wilczyński, Grzegorz, et al.
Published: (2026)
One-for-More: Continual Diffusion Model for Anomaly Detection
by: Li, Xiaofan, et al.
Published: (2025)
by: Li, Xiaofan, et al.
Published: (2025)
Generative Motion In-betweening by Diffusion over Continuous Implicit Representations
by: Fan, Shiyu, et al.
Published: (2026)
by: Fan, Shiyu, et al.
Published: (2026)
A Unified Diffusion Framework for Scene-aware Human Motion Estimation from Sparse Signals
by: Tang, Jiangnan, et al.
Published: (2024)
by: Tang, Jiangnan, et al.
Published: (2024)
LookCloser: Frequency-aware Radiance Field for Tiny-Detail Scene
by: Zhang, Xiaoyu, et al.
Published: (2025)
by: Zhang, Xiaoyu, et al.
Published: (2025)
PromptAD: Learning Prompts with only Normal Samples for Few-Shot Anomaly Detection
by: Li, Xiaofan, et al.
Published: (2024)
by: Li, Xiaofan, et al.
Published: (2024)
FastLGS: Speeding up Language Embedded Gaussians with Feature Grid Mapping
by: Ji, Yuzhou, et al.
Published: (2024)
by: Ji, Yuzhou, et al.
Published: (2024)
Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction
by: Guo, Zhiyang, et al.
Published: (2024)
by: Guo, Zhiyang, et al.
Published: (2024)
Learning Neural Implicit through Volume Rendering with Attentive Depth Fusion Priors
by: Hu, Pengchong, et al.
Published: (2023)
by: Hu, Pengchong, et al.
Published: (2023)
Exploiting Diffusion Prior for Real-World Image Dehazing with Unpaired Training
by: Lan, Yunwei, et al.
Published: (2025)
by: Lan, Yunwei, et al.
Published: (2025)
Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback
by: Li, Zongjian, et al.
Published: (2025)
by: Li, Zongjian, et al.
Published: (2025)
Similar Items
-
Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
by: Jingyu, Gong, et al.
Published: (2025) -
DEMOS: Dynamic Environment Motion Synthesis in 3D Scenes via Local Spherical-BEV Perception
by: Gong, Jingyu, et al.
Published: (2024) -
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
by: Zhang, Renhe, et al.
Published: (2026) -
GSCompleter: A Distillation-Free Plugin for Metric-Aware 3D Gaussian Splatting Completion in Seconds
by: Gao, Ao, et al.
Published: (2026) -
Continuous Piecewise-Affine Based Motion Model for Image Animation
by: Wang, Hexiang, et al.
Published: (2024)