Multimodal Priors-Augmented Text-Driven 3D Human-Object Interaction Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yin, Zhang, Ziyao, Leng, Zhiying, Liu, Haitian, Li, Frederick W. B., Li, Mu, Liang, Xiaohui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fg-T2M++: LLMs-Augmented Fine-Grained Text Driven Human Motion Generation
by: Wang, Yin, et al.
Published: (2025)
by: Wang, Yin, et al.
Published: (2025)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
MOST: Motion Diffusion Model for Rare Text via Temporal Clip Banzhaf Interaction
by: Wang, Yin, et al.
Published: (2025)
by: Wang, Yin, et al.
Published: (2025)
Cross-Temporal 3D Gaussian Splatting for Sparse-View Guided Scene Update
by: An, Zeyuan, et al.
Published: (2025)
by: An, Zeyuan, et al.
Published: (2025)
Fine-grained text-driven dual-human motion generation via dynamic hierarchical interaction
by: Li, Mu, et al.
Published: (2025)
by: Li, Mu, et al.
Published: (2025)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors
by: Zhu, Thomas Hanwen, et al.
Published: (2024)
by: Zhu, Thomas Hanwen, et al.
Published: (2024)
InterFusion: Text-Driven Generation of 3D Human-Object Interaction
by: Dai, Sisi, et al.
Published: (2024)
by: Dai, Sisi, et al.
Published: (2024)
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
by: Huang, Ziyao, et al.
Published: (2025)
by: Huang, Ziyao, et al.
Published: (2025)
Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars
by: Zhang, Youliang, et al.
Published: (2026)
by: Zhang, Youliang, et al.
Published: (2026)
Uncertainty-aware Probabilistic 3D Human Motion Forecasting via Invertible Networks
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
OOD-HOI: Text-Driven 3D Whole-Body Human-Object Interactions Generation Beyond Training Domains
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
From Category to Scenery: An End-to-End Framework for Multi-Person Human-Object Interaction Recognition in Videos
by: Qiao, Tanqiu, et al.
Published: (2024)
by: Qiao, Tanqiu, et al.
Published: (2024)
InteractMove: Text-Controlled Human-Object Interaction Generation in 3D Scenes with Movable Objects
by: Cai, Xinhao, et al.
Published: (2025)
by: Cai, Xinhao, et al.
Published: (2025)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
by: Lou, Yuke, et al.
Published: (2025)
by: Lou, Yuke, et al.
Published: (2025)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
by: Peng, Xiaogang, et al.
Published: (2023)
by: Peng, Xiaogang, et al.
Published: (2023)
GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects
by: Li, Shujia, et al.
Published: (2025)
by: Li, Shujia, et al.
Published: (2025)
InterDreamer: Zero-Shot Text to 3D Dynamic Human-Object Interaction
by: Xu, Sirui, et al.
Published: (2024)
by: Xu, Sirui, et al.
Published: (2024)
SldprtNet: A Large-Scale Multimodal Dataset for CAD Generation in Language-Driven 3D Design
by: Li, Ruogu, et al.
Published: (2026)
by: Li, Ruogu, et al.
Published: (2026)
Geometric Visual Fusion Graph Neural Networks for Multi-Person Human-Object Interaction Recognition in Videos
by: Qiao, Tanqiu, et al.
Published: (2025)
by: Qiao, Tanqiu, et al.
Published: (2025)
Learning Human-Object Interaction for 3D Human Pose Estimation from LiDAR Point Clouds
by: Jung, Daniel Sungho, et al.
Published: (2026)
by: Jung, Daniel Sungho, et al.
Published: (2026)
TeHOR: Text-Guided 3D Human and Object Reconstruction with Textures
by: Nam, Hyeongjin, et al.
Published: (2026)
by: Nam, Hyeongjin, et al.
Published: (2026)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
by: Li, Jincheng, et al.
Published: (2025)
by: Li, Jincheng, et al.
Published: (2025)
PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects
by: Cao, Ziang, et al.
Published: (2026)
by: Cao, Ziang, et al.
Published: (2026)
InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation
by: Xu, Sirui, et al.
Published: (2025)
by: Xu, Sirui, et al.
Published: (2025)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
by: Liu, Zexiang, et al.
Published: (2023)
by: Liu, Zexiang, et al.
Published: (2023)
InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
by: Lin, Chenguo, et al.
Published: (2024)
by: Lin, Chenguo, et al.
Published: (2024)
ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors
by: Huang, Zihao, et al.
Published: (2026)
by: Huang, Zihao, et al.
Published: (2026)
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
by: Zhou, Donghao, et al.
Published: (2026)
by: Zhou, Donghao, et al.
Published: (2026)
HumanGaussian: Text-Driven 3D Human Generation with Gaussian Splatting
by: Liu, Xian, et al.
Published: (2023)
by: Liu, Xian, et al.
Published: (2023)
HOIMotion: Forecasting Human Motion During Human-Object Interactions Using Egocentric 3D Object Bounding Boxes
by: Hu, Zhiming, et al.
Published: (2024)
by: Hu, Zhiming, et al.
Published: (2024)
FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors
by: Li, Chenxi, et al.
Published: (2025)
by: Li, Chenxi, et al.
Published: (2025)
Text2Avatar: Text to 3D Human Avatar Generation with Codebook-Driven Body Controllable Attribute
by: Gong, Chaoqun, et al.
Published: (2024)
by: Gong, Chaoqun, et al.
Published: (2024)
AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation
by: Xu, Ziyi, et al.
Published: (2024)
by: Xu, Ziyi, et al.
Published: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
by: Cha, Junuk, et al.
Published: (2024)
by: Cha, Junuk, et al.
Published: (2024)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
by: Cai, Songjin, et al.
Published: (2026)
by: Cai, Songjin, et al.
Published: (2026)
AnchorHOI: Zero-shot Generation of 4D Human-Object Interaction via Anchor-based Prior Distillation
by: Dai, Sisi, et al.
Published: (2025)
by: Dai, Sisi, et al.
Published: (2025)
Two-Person Interaction Augmentation with Skeleton Priors
by: Li, Baiyi, et al.
Published: (2024)
by: Li, Baiyi, et al.
Published: (2024)
Interact3D: Compositional 3D Generation of Interactive Objects
by: Shan, Hui, et al.
Published: (2026)
by: Shan, Hui, et al.
Published: (2026)
TPG-INR: Target Prior-Guided Implicit 3D CT Reconstruction for Enhanced Sparse-view Imaging
by: Cao, Qinglei, et al.
Published: (2025)
by: Cao, Qinglei, et al.
Published: (2025)
Similar Items
-
Fg-T2M++: LLMs-Augmented Fine-Grained Text Driven Human Motion Generation
by: Wang, Yin, et al.
Published: (2025) -
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026) -
MOST: Motion Diffusion Model for Rare Text via Temporal Clip Banzhaf Interaction
by: Wang, Yin, et al.
Published: (2025) -
Cross-Temporal 3D Gaussian Splatting for Sparse-View Guided Scene Update
by: An, Zeyuan, et al.
Published: (2025) -
Fine-grained text-driven dual-human motion generation via dynamic hierarchical interaction
by: Li, Mu, et al.
Published: (2025)