Exploring Interactive Semantic Alignment for Efficient HOI Detection with Vision-language Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Dong, Jihao, Pan, Renjie, Yang, Hua |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
di: Lei, Ting, et al.
Pubblicazione: (2024)
di: Lei, Ting, et al.
Pubblicazione: (2024)
UAHOI: Uncertainty-aware Robust Interaction Learning for HOI Detection
di: Chen, Mu, et al.
Pubblicazione: (2024)
di: Chen, Mu, et al.
Pubblicazione: (2024)
CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific Methods
di: Lei, Qinqian, et al.
Pubblicazione: (2025)
di: Lei, Qinqian, et al.
Pubblicazione: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
di: Bansal, Siddhant, et al.
Pubblicazione: (2024)
di: Bansal, Siddhant, et al.
Pubblicazione: (2024)
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
di: Lei, Ting, et al.
Pubblicazione: (2024)
di: Lei, Ting, et al.
Pubblicazione: (2024)
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
di: Lei, Ting, et al.
Pubblicazione: (2025)
di: Lei, Ting, et al.
Pubblicazione: (2025)
Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models
di: Zhan, Yu-Wei, et al.
Pubblicazione: (2023)
di: Zhan, Yu-Wei, et al.
Pubblicazione: (2023)
Enhancing Vision-Language Model with Unmasked Token Alignment
di: Liu, Jihao, et al.
Pubblicazione: (2024)
di: Liu, Jihao, et al.
Pubblicazione: (2024)
F-HOI: Toward Fine-grained Semantic-Aligned 3D Human-Object Interactions
di: Yang, Jie, et al.
Pubblicazione: (2024)
di: Yang, Jie, et al.
Pubblicazione: (2024)
Funnel-HOI: Top-Down Perception for Zero-Shot HOI Detection
di: Sarma, Sandipan, et al.
Pubblicazione: (2025)
di: Sarma, Sandipan, et al.
Pubblicazione: (2025)
Efficient Explicit Joint-level Interaction Modeling with Mamba for Text-guided HOI Generation
di: Huang, Guohong, et al.
Pubblicazione: (2025)
di: Huang, Guohong, et al.
Pubblicazione: (2025)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
di: Chen, Junwen, et al.
Pubblicazione: (2025)
di: Chen, Junwen, et al.
Pubblicazione: (2025)
ContextHOI: Spatial Context Learning for Human-Object Interaction Detection
di: Jia, Mingda, et al.
Pubblicazione: (2024)
di: Jia, Mingda, et al.
Pubblicazione: (2024)
VLM-HOI: Vision Language Models for Interpretable Human-Object Interaction Analysis
di: Kang, Donggoo, et al.
Pubblicazione: (2024)
di: Kang, Donggoo, et al.
Pubblicazione: (2024)
CycleHOI: Improving Human-Object Interaction Detection with Cycle Consistency of Detection and Generation
di: Wang, Yisen, et al.
Pubblicazione: (2024)
di: Wang, Yisen, et al.
Pubblicazione: (2024)
Zero-shot HOI Detection with MLLM-based Detector-agnostic Interaction Recognition
di: Xuan, Shiyu, et al.
Pubblicazione: (2026)
di: Xuan, Shiyu, et al.
Pubblicazione: (2026)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
di: Nguyen, Minh Anh, et al.
Pubblicazione: (2026)
di: Nguyen, Minh Anh, et al.
Pubblicazione: (2026)
Asymmetric Visual Semantic Embedding Framework for Efficient Vision-Language Alignment
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection
di: Lei, Qinqian, et al.
Pubblicazione: (2024)
di: Lei, Qinqian, et al.
Pubblicazione: (2024)
CL-HOI: Cross-Level Human-Object Interaction Distillation from Vision Large Language Models
di: Gao, Jianjun, et al.
Pubblicazione: (2024)
di: Gao, Jianjun, et al.
Pubblicazione: (2024)
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
di: Noack, Maja, et al.
Pubblicazione: (2026)
di: Noack, Maja, et al.
Pubblicazione: (2026)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
RD-ViT: Recurrent-Depth Vision Transformer for Semantic Segmentation with Reduced Data Dependence Extending the Recurrent-Depth Transformer Architecture to Dense Prediction
di: He, Renjie
Pubblicazione: (2026)
di: He, Renjie
Pubblicazione: (2026)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
di: Zhao, Yubo, et al.
Pubblicazione: (2026)
di: Zhao, Yubo, et al.
Pubblicazione: (2026)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
di: Cai, Songjin, et al.
Pubblicazione: (2026)
di: Cai, Songjin, et al.
Pubblicazione: (2026)
MF2Summ: Multimodal Fusion for Video Summarization with Temporal Alignment
di: wang, Shuo, et al.
Pubblicazione: (2025)
di: wang, Shuo, et al.
Pubblicazione: (2025)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
di: Zeng, Ling-An, et al.
Pubblicazione: (2025)
di: Zeng, Ling-An, et al.
Pubblicazione: (2025)
GlovEgo-HOI: Bridging the Synthetic-to-Real Gap for Industrial Egocentric Human-Object Interaction Detection
di: Spoto, Alfio, et al.
Pubblicazione: (2026)
di: Spoto, Alfio, et al.
Pubblicazione: (2026)
Visual Diversity and Region-aware Prompt Learning for Zero-shot HOI Detection
di: Yang, Chanhyeong, et al.
Pubblicazione: (2025)
di: Yang, Chanhyeong, et al.
Pubblicazione: (2025)
SGC-Net: Stratified Granular Comparison Network for Open-Vocabulary HOI Detection
di: Lin, Xin, et al.
Pubblicazione: (2025)
di: Lin, Xin, et al.
Pubblicazione: (2025)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
di: Xue, Zihui, et al.
Pubblicazione: (2024)
di: Xue, Zihui, et al.
Pubblicazione: (2024)
Dex2HOI: Dexterous Bimanual Two-Object Interaction Generation
di: Pratikaki, Chrysa, et al.
Pubblicazione: (2026)
di: Pratikaki, Chrysa, et al.
Pubblicazione: (2026)
PA-HOI: A Physics-Aware Human and Object Interaction Dataset
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
di: Wu, Lin, et al.
Pubblicazione: (2025)
di: Wu, Lin, et al.
Pubblicazione: (2025)
Controllable Hand Grasp Generation for HOI and Efficient Evaluation Methods
di: Ishant, et al.
Pubblicazione: (2025)
di: Ishant, et al.
Pubblicazione: (2025)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
di: Zeng, Huajian, et al.
Pubblicazione: (2026)
di: Zeng, Huajian, et al.
Pubblicazione: (2026)
OneHOI: Unifying Human-Object Interaction Generation and Editing
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2026)
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2026)
PersonaHOI: Effortlessly Improving Personalized Face with Human-Object Interaction Generation
di: Hu, Xinting, et al.
Pubblicazione: (2025)
di: Hu, Xinting, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
di: Lei, Ting, et al.
Pubblicazione: (2024) -
UAHOI: Uncertainty-aware Robust Interaction Learning for HOI Detection
di: Chen, Mu, et al.
Pubblicazione: (2024) -
CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific Methods
di: Lei, Qinqian, et al.
Pubblicazione: (2025) -
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
di: Bansal, Siddhant, et al.
Pubblicazione: (2024) -
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
di: Lei, Ting, et al.
Pubblicazione: (2024)