THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Qianyang, Shi, Ye, Huang, Xiaoshui, Yu, Jingyi, Xu, Lan, Wang, Jingya |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
di: Deng, Zekai, et al.
Pubblicazione: (2025)
di: Deng, Zekai, et al.
Pubblicazione: (2025)
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
di: Zhang, Juze, et al.
Pubblicazione: (2024)
di: Zhang, Juze, et al.
Pubblicazione: (2024)
StackFLOW: Monocular Human-Object Reconstruction by Stacked Normalizing Flow with Offset
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
A Unified Diffusion Framework for Scene-aware Human Motion Estimation from Sparse Signals
di: Tang, Jiangnan, et al.
Pubblicazione: (2024)
di: Tang, Jiangnan, et al.
Pubblicazione: (2024)
HandDiffuse: Generative Controllers for Two-Hand Interactions via Diffusion Models
di: Lin, Pei, et al.
Pubblicazione: (2023)
di: Lin, Pei, et al.
Pubblicazione: (2023)
Gaze-guided Hand-Object Interaction Synthesis: Dataset and Method
di: Tian, Jie, et al.
Pubblicazione: (2024)
di: Tian, Jie, et al.
Pubblicazione: (2024)
I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions
di: Zhao, Chengfeng, et al.
Pubblicazione: (2023)
di: Zhao, Chengfeng, et al.
Pubblicazione: (2023)
Monocular Human-Object Reconstruction in the Wild
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
Towards Immersive Human-X Interaction: A Real-Time Framework for Physically Plausible Motion Synthesis
di: Ji, Kaiyang, et al.
Pubblicazione: (2025)
di: Ji, Kaiyang, et al.
Pubblicazione: (2025)
InterAgent: Physics-based Multi-agent Command Execution via Diffusion on Interaction Graphs
di: Li, Bin, et al.
Pubblicazione: (2025)
di: Li, Bin, et al.
Pubblicazione: (2025)
Diffusion Bridge or Flow Matching? A Unifying Framework and Comparative Analysis
di: Zhu, Kaizhen, et al.
Pubblicazione: (2025)
di: Zhu, Kaizhen, et al.
Pubblicazione: (2025)
BOTH2Hands: Inferring 3D Hands from Both Text Prompts and Body Dynamics
di: Zhang, Wenqian, et al.
Pubblicazione: (2023)
di: Zhang, Wenqian, et al.
Pubblicazione: (2023)
Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion
di: Qu, Wentao, et al.
Pubblicazione: (2025)
di: Qu, Wentao, et al.
Pubblicazione: (2025)
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
di: Wu, Lin, et al.
Pubblicazione: (2025)
di: Wu, Lin, et al.
Pubblicazione: (2025)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
LiveHPS: LiDAR-based Scene-level Human Pose and Shape Estimation in Free Environment
di: Ren, Yiming, et al.
Pubblicazione: (2024)
di: Ren, Yiming, et al.
Pubblicazione: (2024)
THOR: Thermal-guided Hand-Object Reasoning via Adaptive Vision Sampling
di: Shahi, Soroush, et al.
Pubblicazione: (2025)
di: Shahi, Soroush, et al.
Pubblicazione: (2025)
Unsupervised Cross-Domain Image Retrieval via Prototypical Optimal Transport
di: Li, Bin, et al.
Pubblicazione: (2024)
di: Li, Bin, et al.
Pubblicazione: (2024)
Contextually Affinitive Neighborhood Refinery for Deep Clustering
di: Yu, Chunlin, et al.
Pubblicazione: (2023)
di: Yu, Chunlin, et al.
Pubblicazione: (2023)
A Self-Conditioned Representation Guided Diffusion Model for Realistic Text-to-LiDAR Scene Generation
di: Qu, Wentao, et al.
Pubblicazione: (2025)
di: Qu, Wentao, et al.
Pubblicazione: (2025)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
di: Li, Liulei, et al.
Pubblicazione: (2024)
di: Li, Liulei, et al.
Pubblicazione: (2024)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
di: Zhang, Cheng, et al.
Pubblicazione: (2024)
di: Zhang, Cheng, et al.
Pubblicazione: (2024)
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2024)
di: Ma, Junyi, et al.
Pubblicazione: (2024)
InterGen: Diffusion-based Multi-human Motion Generation under Complex Interactions
di: Liang, Han, et al.
Pubblicazione: (2023)
di: Liang, Han, et al.
Pubblicazione: (2023)
A Unified and Fast-Sampling Diffusion Bridge Framework via Stochastic Optimal Control
di: Pan, Mokai, et al.
Pubblicazione: (2025)
di: Pan, Mokai, et al.
Pubblicazione: (2025)
UniDB: A Unified Diffusion Bridge Framework via Stochastic Optimal Control
di: Zhu, Kaizhen, et al.
Pubblicazione: (2025)
di: Zhu, Kaizhen, et al.
Pubblicazione: (2025)
Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
di: Wei, Yana, et al.
Pubblicazione: (2025)
di: Wei, Yana, et al.
Pubblicazione: (2025)
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model
di: Yu, Chunlin, et al.
Pubblicazione: (2024)
di: Yu, Chunlin, et al.
Pubblicazione: (2024)
Multimodal Graph Network Modeling for Human-Object Interaction Detection with PDE Graph Diffusion
di: Ji, Wenxuan, et al.
Pubblicazione: (2025)
di: Ji, Wenxuan, et al.
Pubblicazione: (2025)
Guiding Human-Object Interactions with Rich Geometry and Relations
di: Xue, Mengqing, et al.
Pubblicazione: (2025)
di: Xue, Mengqing, et al.
Pubblicazione: (2025)
InterFusion: Text-Driven Generation of 3D Human-Object Interaction
di: Dai, Sisi, et al.
Pubblicazione: (2024)
di: Dai, Sisi, et al.
Pubblicazione: (2024)
THOR2: Topological Analysis for 3D Shape and Color-Based Human-Inspired Object Recognition in Unseen Environments
di: Samani, Ekta U., et al.
Pubblicazione: (2024)
di: Samani, Ekta U., et al.
Pubblicazione: (2024)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
di: Peng, Xiaogang, et al.
Pubblicazione: (2023)
di: Peng, Xiaogang, et al.
Pubblicazione: (2023)
SMGDiff: Soccer Motion Generation using diffusion probabilistic models
di: Yang, Hongdi, et al.
Pubblicazione: (2024)
di: Yang, Hongdi, et al.
Pubblicazione: (2024)
COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
di: Sun, Shaorong, et al.
Pubblicazione: (2024)
di: Sun, Shaorong, et al.
Pubblicazione: (2024)
DiscoForcing: A Unified Framework for Real-Time Audio-Driven Character Control with Diffusion Forcing
di: Ji, Kaiyang, et al.
Pubblicazione: (2026)
di: Ji, Kaiyang, et al.
Pubblicazione: (2026)
InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing
di: Yu, Haoran, et al.
Pubblicazione: (2025)
di: Yu, Haoran, et al.
Pubblicazione: (2025)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
di: Zhang, Jinlu, et al.
Pubblicazione: (2025)
di: Zhang, Jinlu, et al.
Pubblicazione: (2025)
InterDreamer: Zero-Shot Text to 3D Dynamic Human-Object Interaction
di: Xu, Sirui, et al.
Pubblicazione: (2024)
di: Xu, Sirui, et al.
Pubblicazione: (2024)
HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid Explorations
di: Dong, Yilan, et al.
Pubblicazione: (2023)
di: Dong, Yilan, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
di: Deng, Zekai, et al.
Pubblicazione: (2025) -
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
di: Zhang, Juze, et al.
Pubblicazione: (2024) -
StackFLOW: Monocular Human-Object Reconstruction by Stacked Normalizing Flow with Offset
di: Huo, Chaofan, et al.
Pubblicazione: (2024) -
A Unified Diffusion Framework for Scene-aware Human Motion Estimation from Sparse Signals
di: Tang, Jiangnan, et al.
Pubblicazione: (2024) -
HandDiffuse: Generative Controllers for Two-Hand Interactions via Diffusion Models
di: Lin, Pei, et al.
Pubblicazione: (2023)