From Category to Scenery: An End-to-End Framework for Multi-Person Human-Object Interaction Recognition in Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiao, Tanqiu, Li, Ruochen, Li, Frederick W. B., Shum, Hubert P. H. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Geometric Visual Fusion Graph Neural Networks for Multi-Person Human-Object Interaction Recognition in Videos
von: Qiao, Tanqiu, et al.
Veröffentlicht: (2025)
von: Qiao, Tanqiu, et al.
Veröffentlicht: (2025)
Unified Spatial-Temporal Edge-Enhanced Graph Networks for Pedestrian Trajectory Prediction
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
TraIL-Det: Transformation-Invariant Local Feature Networks for 3D LiDAR Object Detection with Unsupervised Pre-Training
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
An End-to-End Framework for Video Multi-Person Pose Estimation
von: Wei, Zhihong
Veröffentlicht: (2025)
von: Wei, Zhihong
Veröffentlicht: (2025)
Two-Person Interaction Augmentation with Skeleton Priors
von: Li, Baiyi, et al.
Veröffentlicht: (2024)
von: Li, Baiyi, et al.
Veröffentlicht: (2024)
M$^3$-VOS: Multi-Phase, Multi-Transition, and Multi-Scenery Video Object Segmentation
von: Chen, Zixuan, et al.
Veröffentlicht: (2024)
von: Chen, Zixuan, et al.
Veröffentlicht: (2024)
Geometric Features Enhanced Human-Object Interaction Detection
von: Zhu, Manli, et al.
Veröffentlicht: (2024)
von: Zhu, Manli, et al.
Veröffentlicht: (2024)
STORM: End-to-End Referring Multi-Object Tracking in Videos
von: Lu, Zijia, et al.
Veröffentlicht: (2026)
von: Lu, Zijia, et al.
Veröffentlicht: (2026)
InterMesh: Explicit Interaction-Aware End-to-End Multi-Person Human Mesh Recovery
von: Zheng, Kaili, et al.
Veröffentlicht: (2026)
von: Zheng, Kaili, et al.
Veröffentlicht: (2026)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
von: Zhong, Pengzhi, et al.
Veröffentlicht: (2025)
von: Zhong, Pengzhi, et al.
Veröffentlicht: (2025)
ViTE: Virtual Graph Trajectory Expert Router for Pedestrian Trajectory Prediction
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
End-to-End Multi-Person Pose Estimation with Pose-Aware Video Transformer
von: Yu, Yonghui, et al.
Veröffentlicht: (2025)
von: Yu, Yonghui, et al.
Veröffentlicht: (2025)
BP-SGCN: Behavioral Pseudo-Label Informed Sparse Graph Convolution Network for Pedestrian and Heterogeneous Trajectory Prediction
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
von: Li, Ruochen, et al.
Veröffentlicht: (2025)
FusionTrack: End-to-End Multi-Object Tracking in Arbitrary Multi-View Environment
von: Li, Xiaohe, et al.
Veröffentlicht: (2025)
von: Li, Xiaohe, et al.
Veröffentlicht: (2025)
ST-SACLF: Style Transfer Informed Self-Attention Classifier for Bias-Aware Painting Classification
von: Vijendran, Mridula, et al.
Veröffentlicht: (2024)
von: Vijendran, Mridula, et al.
Veröffentlicht: (2024)
Tracking by Detection and Query: An Efficient End-to-End Framework for Multi-Object Tracking
von: Jia, Shukun, et al.
Veröffentlicht: (2024)
von: Jia, Shukun, et al.
Veröffentlicht: (2024)
ScrewSplat: An End-to-End Method for Articulated Object Recognition
von: Kim, Seungyeon, et al.
Veröffentlicht: (2025)
von: Kim, Seungyeon, et al.
Veröffentlicht: (2025)
LPSNet: End-to-End Human Pose and Shape Estimation with Lensless Imaging
von: Ge, Haoyang, et al.
Veröffentlicht: (2024)
von: Ge, Haoyang, et al.
Veröffentlicht: (2024)
PHI: Bridging Domain Shift in Long-Term Action Quality Assessment via Progressive Hierarchical Instruction
von: Zhou, Kanglei, et al.
Veröffentlicht: (2025)
von: Zhou, Kanglei, et al.
Veröffentlicht: (2025)
USAD: End-to-End Human Activity Recognition via Diffusion Model with Spatiotemporal Attention
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
Large-Scale Multi-Character Interaction Synthesis
von: Chang, Ziyi, et al.
Veröffentlicht: (2025)
von: Chang, Ziyi, et al.
Veröffentlicht: (2025)
End-to-End Chess Recognition
von: Masouris, Athanasios, et al.
Veröffentlicht: (2023)
von: Masouris, Athanasios, et al.
Veröffentlicht: (2023)
Exploring Disentangled and Controllable Human Image Synthesis: From End-to-End to Stage-by-Stage
von: Sun, Zhengwentai, et al.
Veröffentlicht: (2025)
von: Sun, Zhengwentai, et al.
Veröffentlicht: (2025)
StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipation
von: Ragusa, Francesco, et al.
Veröffentlicht: (2023)
von: Ragusa, Francesco, et al.
Veröffentlicht: (2023)
LMVC: An End-to-End Learned Multiview Video Coding Framework
von: Sheng, Xihua, et al.
Veröffentlicht: (2025)
von: Sheng, Xihua, et al.
Veröffentlicht: (2025)
Two-Stage Human Verification using HandCAPTCHA and Anti-Spoofed Finger Biometrics with Feature Selection
von: Bera, Asish, et al.
Veröffentlicht: (2024)
von: Bera, Asish, et al.
Veröffentlicht: (2024)
OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer
von: Li, Jinyang, et al.
Veröffentlicht: (2025)
von: Li, Jinyang, et al.
Veröffentlicht: (2025)
DMAT: An End-to-End Framework for Joint Atmospheric Turbulence Mitigation and Object Detection
von: Hill, Paul, et al.
Veröffentlicht: (2025)
von: Hill, Paul, et al.
Veröffentlicht: (2025)
VSD-MOT: End-to-End Multi-Object Tracking in Low-Quality Video Scenes Guided by Visual Semantic Distillation
von: Du, Jun
Veröffentlicht: (2026)
von: Du, Jun
Veröffentlicht: (2026)
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
End-to-End Streaming Video Temporal Action Segmentation with Reinforce Learning
von: Zhang, Jinrong, et al.
Veröffentlicht: (2023)
von: Zhang, Jinrong, et al.
Veröffentlicht: (2023)
Action Images: End-to-End Policy Learning via Multiview Video Generation
von: Zhen, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2026)
VividAnimator: An End-to-End Audio and Pose-driven Half-Body Human Animation Framework
von: Huang, Donglin, et al.
Veröffentlicht: (2025)
von: Huang, Donglin, et al.
Veröffentlicht: (2025)
ART: Adaptive Relational Transformer for Pedestrian Trajectory Prediction with Temporal-Aware Relations
von: Li, Ruochen, et al.
Veröffentlicht: (2026)
von: Li, Ruochen, et al.
Veröffentlicht: (2026)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
MAGR: Manifold-Aligned Graph Regularization for Continual Action Quality Assessment
von: Zhou, Kanglei, et al.
Veröffentlicht: (2024)
von: Zhou, Kanglei, et al.
Veröffentlicht: (2024)
Motion In-Betweening for Densely Interacting Characters
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaotang, et al.
Veröffentlicht: (2025)
Towards Fully Decoupled End-to-End Person Search
von: Zhang, Pengcheng, et al.
Veröffentlicht: (2023)
von: Zhang, Pengcheng, et al.
Veröffentlicht: (2023)
Beyond Hungarian: Match-Free Supervision for End-to-End Object Detection
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
End-To-End Underwater Video Enhancement: Dataset and Model
von: Du, Dazhao, et al.
Veröffentlicht: (2024)
von: Du, Dazhao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Geometric Visual Fusion Graph Neural Networks for Multi-Person Human-Object Interaction Recognition in Videos
von: Qiao, Tanqiu, et al.
Veröffentlicht: (2025) -
Unified Spatial-Temporal Edge-Enhanced Graph Networks for Pedestrian Trajectory Prediction
von: Li, Ruochen, et al.
Veröffentlicht: (2025) -
TraIL-Det: Transformation-Invariant Local Feature Networks for 3D LiDAR Object Detection with Unsupervised Pre-Training
von: Li, Li, et al.
Veröffentlicht: (2024) -
An End-to-End Framework for Video Multi-Person Pose Estimation
von: Wei, Zhihong
Veröffentlicht: (2025) -
Two-Person Interaction Augmentation with Skeleton Priors
von: Li, Baiyi, et al.
Veröffentlicht: (2024)