HOComp: Interaction-Aware Human-Object Composition
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Dong, Jia, Jinyuan, Liu, Yuhao, Lau, Rynson W. H. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
World-Shaper: A Unified Framework for 360° Panoramic Editing
by: Liang, Dong, et al.
Published: (2026)
by: Liang, Dong, et al.
Published: (2026)
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
by: Yang, Zaiquan, et al.
Published: (2025)
by: Yang, Zaiquan, et al.
Published: (2025)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
by: Liu, Yuhao, et al.
Published: (2025)
by: Liu, Yuhao, et al.
Published: (2025)
Diff-Plugin: Revitalizing Details for Diffusion-based Low-level Tasks
by: Liu, Yuhao, et al.
Published: (2024)
by: Liu, Yuhao, et al.
Published: (2024)
Inverse Rendering of Glossy Objects via the Neural Plenoptic Function and Radiance Fields
by: Wang, Haoyuan, et al.
Published: (2024)
by: Wang, Haoyuan, et al.
Published: (2024)
Boosting Weakly-Supervised Referring Image Segmentation via Progressive Comprehension
by: Yang, Zaiquan, et al.
Published: (2024)
by: Yang, Zaiquan, et al.
Published: (2024)
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
Recasting Regional Lighting for Shadow Removal
by: Liu, Yuhao, et al.
Published: (2024)
by: Liu, Yuhao, et al.
Published: (2024)
Do MLLMs Exhibit Human-like Perceptual Behaviors? HVSBench: A Benchmark for MLLM Alignment with Human Perceptual Behavior
by: Lin, Jiaying, et al.
Published: (2024)
by: Lin, Jiaying, et al.
Published: (2024)
Delving into Dark Regions for Robust Shadow Detection
by: Guan, Huankang, et al.
Published: (2024)
by: Guan, Huankang, et al.
Published: (2024)
MirrorMamba: Towards Scalable and Robust Mirror Detection in Videos
by: Song, Rui, et al.
Published: (2025)
by: Song, Rui, et al.
Published: (2025)
Structure-Informed Shadow Removal Networks
by: Liu, Yuhao, et al.
Published: (2023)
by: Liu, Yuhao, et al.
Published: (2023)
Color Shift Estimation-and-Correction for Image Enhancement
by: Li, Yiyu, et al.
Published: (2024)
by: Li, Yiyu, et al.
Published: (2024)
Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface Detection
by: Lin, Jiaying, et al.
Published: (2022)
by: Lin, Jiaying, et al.
Published: (2022)
LuSh-NeRF: Lighting up and Sharpening NeRFs for Low-light Scenes
by: Qu, Zefan, et al.
Published: (2024)
by: Qu, Zefan, et al.
Published: (2024)
OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding
by: Zhao, Youjun, et al.
Published: (2024)
by: Zhao, Youjun, et al.
Published: (2024)
Revisiting the Integration of Convolution and Attention for Vision Backbone
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
by: Qu, Zefan, et al.
Published: (2025)
by: Qu, Zefan, et al.
Published: (2025)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
by: Zhang, Jiajun, et al.
Published: (2023)
by: Zhang, Jiajun, et al.
Published: (2023)
Locality-Aware Zero-Shot Human-Object Interaction Detection
by: Kim, Sanghyun, et al.
Published: (2025)
by: Kim, Sanghyun, et al.
Published: (2025)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
by: Huang, Tianyu, et al.
Published: (2025)
by: Huang, Tianyu, et al.
Published: (2025)
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
by: Huang, Tianyu, et al.
Published: (2024)
by: Huang, Tianyu, et al.
Published: (2024)
MDeRainNet: An Efficient Macro-pixel Image Rain Removal Network
by: Yan, Tao, et al.
Published: (2024)
by: Yan, Tao, et al.
Published: (2024)
PA-HOI: A Physics-Aware Human and Object Interaction Dataset
by: Wang, Ruiyan, et al.
Published: (2025)
by: Wang, Ruiyan, et al.
Published: (2025)
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
by: Wen, Boran, et al.
Published: (2025)
by: Wen, Boran, et al.
Published: (2025)
Multimodal Priors-Augmented Text-Driven 3D Human-Object Interaction Generation
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
Controllable Human-Object Interaction Synthesis
by: Li, Jiaman, et al.
Published: (2023)
by: Li, Jiaman, et al.
Published: (2023)
Glass Surface Detection: Leveraging Reflection Dynamics in Flash/No-flash Imagery
by: Yan, Tao, et al.
Published: (2025)
by: Yan, Tao, et al.
Published: (2025)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
Orchestrating the Symphony of Prompt Distribution Learning for Human-Object Interaction Detection
by: Jia, Mingda, et al.
Published: (2024)
by: Jia, Mingda, et al.
Published: (2024)
ContextHOI: Spatial Context Learning for Human-Object Interaction Detection
by: Jia, Mingda, et al.
Published: (2024)
by: Jia, Mingda, et al.
Published: (2024)
RefSTAR: Blind Facial Image Restoration with Reference Selection, Transfer, and Reconstruction
by: Yin, Zhicun, et al.
Published: (2025)
by: Yin, Zhicun, et al.
Published: (2025)
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
by: Liu, Yangyang, et al.
Published: (2025)
by: Liu, Yangyang, et al.
Published: (2025)
RHOBIN Challenge: Reconstruction of Human Object Interaction
by: Xie, Xianghui, et al.
Published: (2024)
by: Xie, Xianghui, et al.
Published: (2024)
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation
by: Cao, Yukang, et al.
Published: (2024)
by: Cao, Yukang, et al.
Published: (2024)
Geometric Features Enhanced Human-Object Interaction Detection
by: Zhu, Manli, et al.
Published: (2024)
by: Zhu, Manli, et al.
Published: (2024)
Similar Items
-
World-Shaper: A Unified Framework for 360° Panoramic Editing
by: Liang, Dong, et al.
Published: (2026) -
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
by: Yang, Zaiquan, et al.
Published: (2025) -
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025) -
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
by: Liu, Yuhao, et al.
Published: (2025) -
Diff-Plugin: Revitalizing Details for Diffusion-based Low-level Tasks
by: Liu, Yuhao, et al.
Published: (2024)