An Image-like Diffusion Method for Human-Object Interaction Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Hui, Xiaofei, Qu, Haoxuan, Rahmani, Hossein, Liu, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Visual Privacy Protection Meets Multimodal Large Language Models
by: Hui, Xiaofei, et al.
Published: (2026)
by: Hui, Xiaofei, et al.
Published: (2026)
MicroscopyMatching: Towards a Ready-to-use Framework for Microscopy Image Analysis in Diverse Conditions
by: Hui, Xiaofei, et al.
Published: (2026)
by: Hui, Xiaofei, et al.
Published: (2026)
A Mixed-Primitive-based Gaussian Splatting Method for Surface Reconstruction
by: Qu, Haoxuan, et al.
Published: (2025)
by: Qu, Haoxuan, et al.
Published: (2025)
Action Detection via an Image Diffusion Process
by: Foo, Lin Geng, et al.
Published: (2024)
by: Foo, Lin Geng, et al.
Published: (2024)
DisC-GS: Discontinuity-aware Gaussian Splatting
by: Qu, Haoxuan, et al.
Published: (2024)
by: Qu, Haoxuan, et al.
Published: (2024)
Recent Advances of Continual Learning in Computer Vision: An Overview
by: Qu, Haoxuan, et al.
Published: (2021)
by: Qu, Haoxuan, et al.
Published: (2021)
ToolFG: Towards Well-Grounded Fine-Grained Image Classification
by: Xue, Yu, et al.
Published: (2026)
by: Xue, Yu, et al.
Published: (2026)
Translating Signals to Languages for sEMG-Based Activity Recognition
by: Wang, Ming, et al.
Published: (2026)
by: Wang, Ming, et al.
Published: (2026)
GPT-Connect: Interaction between Text-Driven Human Motion Generator and 3D Scenes in a Training-free Manner
by: Qu, Haoxuan, et al.
Published: (2024)
by: Qu, Haoxuan, et al.
Published: (2024)
6D-Diff: A Keypoint Diffusion Framework for 6D Object Pose Estimation
by: Xu, Li, et al.
Published: (2023)
by: Xu, Li, et al.
Published: (2023)
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
by: Zhang, Zhengbo, et al.
Published: (2024)
by: Zhang, Zhengbo, et al.
Published: (2024)
Learning to Generate Cross-Task Unexploitable Examples
by: Qu, Haoxuan, et al.
Published: (2025)
by: Qu, Haoxuan, et al.
Published: (2025)
TSTMotion: Training-free Scene-aware Text-to-motion Generation
by: Guo, Ziyan, et al.
Published: (2025)
by: Guo, Ziyan, et al.
Published: (2025)
Exploring Self- and Cross-Triplet Correlations for Human-Object Interaction Detection
by: Jiang, Weibo, et al.
Published: (2024)
by: Jiang, Weibo, et al.
Published: (2024)
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
HyLiFormer: Hyperbolic Linear Attention for Skeleton-based Human Action Recognition
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
AI-Generated Content (AIGC) for Various Data Modalities: A Survey
by: Foo, Lin Geng, et al.
Published: (2023)
by: Foo, Lin Geng, et al.
Published: (2023)
LLMs are Good Action Recognizers
by: Qu, Haoxuan, et al.
Published: (2024)
by: Qu, Haoxuan, et al.
Published: (2024)
MINet: Multi-scale Interactive Network for Real-time Salient Object Detection of Strip Steel Surface Defects
by: Shen, Kunye, et al.
Published: (2024)
by: Shen, Kunye, et al.
Published: (2024)
Egocentric Human-Object Interaction Detection: A New Benchmark and Method
by: Deng, Kunyuan, et al.
Published: (2025)
by: Deng, Kunyuan, et al.
Published: (2025)
Off-the-shelf ChatGPT is a Good Few-shot Human Motion Predictor
by: Qu, Haoxuan, et al.
Published: (2024)
by: Qu, Haoxuan, et al.
Published: (2024)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
by: Li, Liulei, et al.
Published: (2024)
by: Li, Liulei, et al.
Published: (2024)
LongDiff: Training-Free Long Video Generation in One Go
by: Li, Zhuoling, et al.
Published: (2025)
by: Li, Zhuoling, et al.
Published: (2025)
Hoi2Threat: An Interpretable Threat Detection Method for Human Violence Scenarios Guided by Human-Object Interaction
by: Wang, Yuhan, et al.
Published: (2025)
by: Wang, Yuhan, et al.
Published: (2025)
Multimodal Graph Network Modeling for Human-Object Interaction Detection with PDE Graph Diffusion
by: Ji, Wenxuan, et al.
Published: (2025)
by: Ji, Wenxuan, et al.
Published: (2025)
TSkel-Mamba: Temporal Dynamic Modeling via State Space Model for Human Skeleton-based Action Recognition
by: Liu, Yanan, et al.
Published: (2025)
by: Liu, Yanan, et al.
Published: (2025)
Avatar Concept Slider: Controllable Editing of Concepts in 3D Human Avatars
by: Foo, Lin Geng, et al.
Published: (2024)
by: Foo, Lin Geng, et al.
Published: (2024)
Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning
by: Zhang, Hang, et al.
Published: (2024)
by: Zhang, Hang, et al.
Published: (2024)
A Plug-and-Play Method for Rare Human-Object Interactions Detection by Bridging Domain Gap
by: Zhang, Lijun, et al.
Published: (2024)
by: Zhang, Lijun, et al.
Published: (2024)
Deep Learning-Based Object Pose Estimation: A Comprehensive Survey
by: Liu, Jian, et al.
Published: (2024)
by: Liu, Jian, et al.
Published: (2024)
Geometric Features Enhanced Human-Object Interaction Detection
by: Zhu, Manli, et al.
Published: (2024)
by: Zhu, Manli, et al.
Published: (2024)
Streamlined Open-Vocabulary Human-Object Interaction Detection
by: Sun, Chang, et al.
Published: (2026)
by: Sun, Chang, et al.
Published: (2026)
Disentangled Pre-training for Human-Object Interaction Detection
by: Li, Zhuolong, et al.
Published: (2024)
by: Li, Zhuolong, et al.
Published: (2024)
DA-HFNet: Progressive Fine-Grained Forgery Image Detection and Localization Based on Dual Attention
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
MTADiffusion: Mask Text Alignment Diffusion Model for Object Inpainting
by: Huang, Jun, et al.
Published: (2025)
by: Huang, Jun, et al.
Published: (2025)
Answering from Sure to Uncertain: Uncertainty-Aware Curriculum Learning for Video Question Answering
by: Li, Haopeng, et al.
Published: (2024)
by: Li, Haopeng, et al.
Published: (2024)
Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
by: Wei, Yana, et al.
Published: (2025)
by: Wei, Yana, et al.
Published: (2025)
Interact-Custom: Customized Human Object Interaction Image Generation
by: Xu, Zhu, et al.
Published: (2025)
by: Xu, Zhu, et al.
Published: (2025)
A Review of Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
Similar Items
-
When Visual Privacy Protection Meets Multimodal Large Language Models
by: Hui, Xiaofei, et al.
Published: (2026) -
MicroscopyMatching: Towards a Ready-to-use Framework for Microscopy Image Analysis in Diverse Conditions
by: Hui, Xiaofei, et al.
Published: (2026) -
A Mixed-Primitive-based Gaussian Splatting Method for Surface Reconstruction
by: Qu, Haoxuan, et al.
Published: (2025) -
Action Detection via an Image Diffusion Process
by: Foo, Lin Geng, et al.
Published: (2024) -
DisC-GS: Discontinuity-aware Gaussian Splatting
by: Qu, Haoxuan, et al.
Published: (2024)