Multimodal Graph Network Modeling for Human-Object Interaction Detection with PDE Graph Diffusion
Fuente:
arXiv
Guardado en:
| Autores principales: | Ji, Wenxuan, Shi, Haichao, Zhang, Xiao-Yu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Graph-Based Uncertainty Modeling and Multimodal Fusion for Salient Object Detection
por: Xiong, Yuqi, et al.
Publicado: (2025)
por: Xiong, Yuqi, et al.
Publicado: (2025)
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
por: Feng, Chen-Bin, et al.
Publicado: (2026)
por: Feng, Chen-Bin, et al.
Publicado: (2026)
Hierarchical Graph Interaction Transformer with Dynamic Token Clustering for Camouflaged Object Detection
por: Yao, Siyuan, et al.
Publicado: (2024)
por: Yao, Siyuan, et al.
Publicado: (2024)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
por: Li, Liulei, et al.
Publicado: (2024)
por: Li, Liulei, et al.
Publicado: (2024)
THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
por: Wu, Qianyang, et al.
Publicado: (2024)
por: Wu, Qianyang, et al.
Publicado: (2024)
An Image-like Diffusion Method for Human-Object Interaction Detection
por: Hui, Xiaofei, et al.
Publicado: (2025)
por: Hui, Xiaofei, et al.
Publicado: (2025)
HOIverse: A Synthetic Scene Graph Dataset With Human Object Interactions
por: Phatak, Mrunmai Vivek, et al.
Publicado: (2025)
por: Phatak, Mrunmai Vivek, et al.
Publicado: (2025)
Geometric Visual Fusion Graph Neural Networks for Multi-Person Human-Object Interaction Recognition in Videos
por: Qiao, Tanqiu, et al.
Publicado: (2025)
por: Qiao, Tanqiu, et al.
Publicado: (2025)
HOID-R1: Reinforcement Learning for Open-World Human-Object Interaction Detection Reasoning with Multimodal Large Language Model
por: Zhang, Zhenhao, et al.
Publicado: (2025)
por: Zhang, Zhenhao, et al.
Publicado: (2025)
HHMR: Holistic Hand Mesh Recovery by Enhancing the Multimodal Controllability of Graph Diffusion Models
por: Li, Mengcheng, et al.
Publicado: (2024)
por: Li, Mengcheng, et al.
Publicado: (2024)
No More Sibling Rivalry: Debiasing Human-Object Interaction Detection
por: Yang, Bin, et al.
Publicado: (2025)
por: Yang, Bin, et al.
Publicado: (2025)
TGBFormer: Transformer-GraphFormer Blender Network for Video Object Detection
por: Qi, Qiang, et al.
Publicado: (2025)
por: Qi, Qiang, et al.
Publicado: (2025)
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
por: Xing, Hao, et al.
Publicado: (2024)
por: Xing, Hao, et al.
Publicado: (2024)
ClickRemoval: An Interactive Open-Source Tool for Object Removal in Diffusion Models
por: Zhang, Ledun, et al.
Publicado: (2026)
por: Zhang, Ledun, et al.
Publicado: (2026)
Exploiting Multimodal Synthetic Data for Egocentric Human-Object Interaction Detection in an Industrial Scenario
por: Leonardi, Rosario, et al.
Publicado: (2023)
por: Leonardi, Rosario, et al.
Publicado: (2023)
Parse Graph-Based Visual-Language Interaction for Human Pose Estimation
por: Liu, Shibang, et al.
Publicado: (2025)
por: Liu, Shibang, et al.
Publicado: (2025)
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
por: Wang, Qishun, et al.
Publicado: (2025)
por: Wang, Qishun, et al.
Publicado: (2025)
Prototype Embedding Optimization for Human-Object Interaction Detection in Livestreaming
por: Zhang, Menghui, et al.
Publicado: (2025)
por: Zhang, Menghui, et al.
Publicado: (2025)
Explicit Motion Handling and Interactive Prompting for Video Camouflaged Object Detection
por: Zhang, Xin, et al.
Publicado: (2024)
por: Zhang, Xin, et al.
Publicado: (2024)
Interactive Masked Image Modeling for Multimodal Object Detection in Remote Sensing
por: Vu, Minh-Duc, et al.
Publicado: (2024)
por: Vu, Minh-Duc, et al.
Publicado: (2024)
Straighter Flow Matching via a Diffusion-Based Coupling Prior
por: Xing, Siyu, et al.
Publicado: (2023)
por: Xing, Siyu, et al.
Publicado: (2023)
Multimodal Hate Detection Using Dual-Stream Graph Neural Networks
por: Yue, Jiangbei, et al.
Publicado: (2025)
por: Yue, Jiangbei, et al.
Publicado: (2025)
Graph Query Networks for Object Detection with Automotive Radar
por: Saini, Loveneet, et al.
Publicado: (2025)
por: Saini, Loveneet, et al.
Publicado: (2025)
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
por: Deng, Zekai, et al.
Publicado: (2025)
por: Deng, Zekai, et al.
Publicado: (2025)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
por: Li, Jincheng, et al.
Publicado: (2025)
por: Li, Jincheng, et al.
Publicado: (2025)
GraphMMP: A Graph Neural Network Model with Mutual Information and Global Fusion for Multimodal Medical Prognosis
por: Shan, Xuhao, et al.
Publicado: (2025)
por: Shan, Xuhao, et al.
Publicado: (2025)
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
por: Zhang, Juze, et al.
Publicado: (2024)
por: Zhang, Juze, et al.
Publicado: (2024)
Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
por: Wei, Yana, et al.
Publicado: (2025)
por: Wei, Yana, et al.
Publicado: (2025)
GraphRelate3D: Context-Dependent 3D Object Detection with Inter-Object Relationship Graphs
por: Liu, Mingyu, et al.
Publicado: (2024)
por: Liu, Mingyu, et al.
Publicado: (2024)
Taking A Closer Look at Interacting Objects: Interaction-Aware Open Vocabulary Scene Graph Generation
por: Li, Lin, et al.
Publicado: (2025)
por: Li, Lin, et al.
Publicado: (2025)
Geometric Features Enhanced Human-Object Interaction Detection
por: Zhu, Manli, et al.
Publicado: (2024)
por: Zhu, Manli, et al.
Publicado: (2024)
Streamlined Open-Vocabulary Human-Object Interaction Detection
por: Sun, Chang, et al.
Publicado: (2026)
por: Sun, Chang, et al.
Publicado: (2026)
Disentangled Pre-training for Human-Object Interaction Detection
por: Li, Zhuolong, et al.
Publicado: (2024)
por: Li, Zhuolong, et al.
Publicado: (2024)
A Review of Human-Object Interaction Detection
por: Wang, Yuxiao, et al.
Publicado: (2024)
por: Wang, Yuxiao, et al.
Publicado: (2024)
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
por: Huang, Ziyao, et al.
Publicado: (2025)
por: Huang, Ziyao, et al.
Publicado: (2025)
ProGraph: Temporally-alignable Probability Guided Graph Topological Modeling for 3D Human Reconstruction
por: Wang, Hongsheng, et al.
Publicado: (2024)
por: Wang, Hongsheng, et al.
Publicado: (2024)
Graph Integrated Multimodal Concept Bottleneck Model
por: Lin, Jiakai, et al.
Publicado: (2025)
por: Lin, Jiakai, et al.
Publicado: (2025)
GraphVLM: Benchmarking Vision Language Models for Multimodal Graph Learning
por: Liu, Jiajin, et al.
Publicado: (2026)
por: Liu, Jiajin, et al.
Publicado: (2026)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
por: Zhang, Zhenhao, et al.
Publicado: (2025)
por: Zhang, Zhenhao, et al.
Publicado: (2025)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
por: Yang, Jie, et al.
Publicado: (2024)
por: Yang, Jie, et al.
Publicado: (2024)
Ejemplares similares
-
Graph-Based Uncertainty Modeling and Multimodal Fusion for Salient Object Detection
por: Xiong, Yuqi, et al.
Publicado: (2025) -
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
por: Feng, Chen-Bin, et al.
Publicado: (2026) -
Hierarchical Graph Interaction Transformer with Dynamic Token Clustering for Camouflaged Object Detection
por: Yao, Siyuan, et al.
Publicado: (2024) -
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
por: Li, Liulei, et al.
Publicado: (2024) -
THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
por: Wu, Qianyang, et al.
Publicado: (2024)