Simultaneous Detection and Interaction Reasoning for Object-Centric Action Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Xunsong, Sun, Pengzhan, Liu, Yangcen, Duan, Lixin, Li, Wen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer
por: Liu, Ziyi, et al.
Publicado: (2025)
por: Liu, Ziyi, et al.
Publicado: (2025)
STAT: Towards Generalizable Temporal Action Localization
por: Liu, Yangcen, et al.
Publicado: (2024)
por: Liu, Yangcen, et al.
Publicado: (2024)
Instance-Free Domain Adaptive Object Detection
por: Yu, Hengfu, et al.
Publicado: (2026)
por: Yu, Hengfu, et al.
Publicado: (2026)
Towards Unsupervised Model Selection for Domain Adaptive Object Detection
por: Yu, Hengfu, et al.
Publicado: (2024)
por: Yu, Hengfu, et al.
Publicado: (2024)
Structure-Aware Human Body Reshaping with Adaptive Affinity-Graph Network
por: Deng, Qiwen, et al.
Publicado: (2024)
por: Deng, Qiwen, et al.
Publicado: (2024)
Beyond Viewpoint: Robust 3D Object Recognition under Arbitrary Views through Joint Multi-Part Representation
por: Fan, Linlong, et al.
Publicado: (2024)
por: Fan, Linlong, et al.
Publicado: (2024)
Video Spatial Reasoning with Object-Centric 3D Rollout
por: Tang, Haoran, et al.
Publicado: (2025)
por: Tang, Haoran, et al.
Publicado: (2025)
Reasoning-Enhanced Object-Centric Learning for Videos
por: Li, Jian, et al.
Publicado: (2024)
por: Li, Jian, et al.
Publicado: (2024)
Improving Skeleton-based Action Recognition with Interactive Object Information
por: Wen, Hao, et al.
Publicado: (2025)
por: Wen, Hao, et al.
Publicado: (2025)
Balanced Sharpness-Aware Minimization for Imbalanced Regression
por: Liu, Yahao, et al.
Publicado: (2025)
por: Liu, Yahao, et al.
Publicado: (2025)
Human-Centric Transformer for Domain Adaptive Action Recognition
por: Lin, Kun-Yu, et al.
Publicado: (2024)
por: Lin, Kun-Yu, et al.
Publicado: (2024)
BikeActions: An Open Platform and Benchmark for Cyclist-Centric VRU Action Recognition
por: Buettner, Max A., et al.
Publicado: (2026)
por: Buettner, Max A., et al.
Publicado: (2026)
Representation-Centric Survey of Supervised Skeletal Action Recognition and the New Benchmark
por: Liu, Yang, et al.
Publicado: (2022)
por: Liu, Yang, et al.
Publicado: (2022)
SkeletonAgent: An Agentic Interaction Framework for Skeleton-based Action Recognition
por: Liu, Hongda, et al.
Publicado: (2025)
por: Liu, Hongda, et al.
Publicado: (2025)
OccludeNet: A Causal Journey into Mixed-View Actor-Centric Video Action Recognition under Occlusions
por: Zhou, Guanyu, et al.
Publicado: (2024)
por: Zhou, Guanyu, et al.
Publicado: (2024)
Object-Centric Latent Action Learning
por: Klepach, Albina, et al.
Publicado: (2025)
por: Klepach, Albina, et al.
Publicado: (2025)
RayFormer: Improving Query-Based Multi-Camera 3D Object Detection via Ray-Centric Strategies
por: Chu, Xiaomeng, et al.
Publicado: (2024)
por: Chu, Xiaomeng, et al.
Publicado: (2024)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
por: Kundu, Sanjoy, et al.
Publicado: (2024)
por: Kundu, Sanjoy, et al.
Publicado: (2024)
ResCLIP: Residual Attention for Training-free Dense Vision-language Inference
por: Yang, Yuhang, et al.
Publicado: (2024)
por: Yang, Yuhang, et al.
Publicado: (2024)
SSR: SAM is a Strong Regularizer for domain adaptive semantic segmentation
por: Ge, Yanqi, et al.
Publicado: (2024)
por: Ge, Yanqi, et al.
Publicado: (2024)
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
por: Xu, Guowei, et al.
Publicado: (2024)
por: Xu, Guowei, et al.
Publicado: (2024)
Object-Centric Data Synthesis for Category-level Object Detection
por: Agarwal, Vikhyat, et al.
Publicado: (2025)
por: Agarwal, Vikhyat, et al.
Publicado: (2025)
Object-Centric Framework for Video Moment Retrieval
por: Li, Zongyao, et al.
Publicado: (2025)
por: Li, Zongyao, et al.
Publicado: (2025)
OpenSlot: Mixed Open-Set Recognition with Object-Centric Learning
por: Yin, Xu, et al.
Publicado: (2024)
por: Yin, Xu, et al.
Publicado: (2024)
EigenActor: Variant Body-Object Interaction Generation Evolved from Invariant Action Basis Reasoning
por: Gao, Xuehao, et al.
Publicado: (2025)
por: Gao, Xuehao, et al.
Publicado: (2025)
ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming Videos
por: Kang, Hyolim, et al.
Publicado: (2024)
por: Kang, Hyolim, et al.
Publicado: (2024)
Patch as Node: Human-Centric Graph Representation Learning for Multimodal Action Recognition
por: Liang, Zeyu, et al.
Publicado: (2025)
por: Liang, Zeyu, et al.
Publicado: (2025)
DroneMOT: Drone-based Multi-Object Tracking Considering Detection Difficulties and Simultaneous Moving of Drones and Objects
por: Wang, Peng, et al.
Publicado: (2024)
por: Wang, Peng, et al.
Publicado: (2024)
Learning Action and Reasoning-Centric Image Editing from Videos and Simulations
por: Krojer, Benno, et al.
Publicado: (2024)
por: Krojer, Benno, et al.
Publicado: (2024)
Generative Human-Object Interaction Detection via Differentiable Cognitive Steering of Multi-modal LLMs
por: Cai, Zhaolin, et al.
Publicado: (2025)
por: Cai, Zhaolin, et al.
Publicado: (2025)
Object-Centric Instruction Augmentation for Robotic Manipulation
por: Wen, Junjie, et al.
Publicado: (2024)
por: Wen, Junjie, et al.
Publicado: (2024)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
por: Li, Congliang, et al.
Publicado: (2022)
por: Li, Congliang, et al.
Publicado: (2022)
Disentangled Pre-training for Human-Object Interaction Detection
por: Li, Zhuolong, et al.
Publicado: (2024)
por: Li, Zhuolong, et al.
Publicado: (2024)
WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes
por: Hu, Jichen, et al.
Publicado: (2026)
por: Hu, Jichen, et al.
Publicado: (2026)
VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?
por: Liu, Yuanxin, et al.
Publicado: (2025)
por: Liu, Yuanxin, et al.
Publicado: (2025)
Interaction-Consistent Object Removal via MLLM-Based Reasoning
por: Huang, Ching-Kai, et al.
Publicado: (2026)
por: Huang, Ching-Kai, et al.
Publicado: (2026)
Improvement of Human-Object Interaction Action Recognition Using Scene Information and Multi-Task Learning Approach
por: Shehata, Hesham M., et al.
Publicado: (2025)
por: Shehata, Hesham M., et al.
Publicado: (2025)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
por: Liu, Xin, et al.
Publicado: (2024)
por: Liu, Xin, et al.
Publicado: (2024)
Hand-Centric Motion Refinement for 3D Hand-Object Interaction via Hierarchical Spatial-Temporal Modeling
por: Hao, Yuze, et al.
Publicado: (2024)
por: Hao, Yuze, et al.
Publicado: (2024)
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
por: Mirjalili, Vahid, et al.
Publicado: (2025)
por: Mirjalili, Vahid, et al.
Publicado: (2025)
Ejemplares similares
-
Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer
por: Liu, Ziyi, et al.
Publicado: (2025) -
STAT: Towards Generalizable Temporal Action Localization
por: Liu, Yangcen, et al.
Publicado: (2024) -
Instance-Free Domain Adaptive Object Detection
por: Yu, Hengfu, et al.
Publicado: (2026) -
Towards Unsupervised Model Selection for Domain Adaptive Object Detection
por: Yu, Hengfu, et al.
Publicado: (2024) -
Structure-Aware Human Body Reshaping with Adaptive Affinity-Graph Network
por: Deng, Qiwen, et al.
Publicado: (2024)