CL-HOI: Cross-Level Human-Object Interaction Distillation from Vision Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Gao, Jianjun, Cai, Chen, Wang, Ruoyu, Liu, Wenyang, Yap, Kim-Hui, Garg, Kratika, Han, Boon-Siew |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OccluTrack: Rethinking Awareness of Occlusion for Enhancing Multiple Pedestrian Tracking
di: Gao, Jianjun, et al.
Pubblicazione: (2023)
di: Gao, Jianjun, et al.
Pubblicazione: (2023)
CM2-Net: Continual Cross-Modal Mapping Network for Driver Action Recognition
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
SSH-Net: A Self-Supervised and Hybrid Network for Noisy Image Watermark Removal
di: Liu, Wenyang, et al.
Pubblicazione: (2025)
di: Liu, Wenyang, et al.
Pubblicazione: (2025)
Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting
di: Cai, Chen, et al.
Pubblicazione: (2024)
di: Cai, Chen, et al.
Pubblicazione: (2024)
OneHOI: Unifying Human-Object Interaction Generation and Editing
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2026)
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2026)
VLM-HOI: Vision Language Models for Interpretable Human-Object Interaction Analysis
di: Kang, Donggoo, et al.
Pubblicazione: (2024)
di: Kang, Donggoo, et al.
Pubblicazione: (2024)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
di: Cai, Songjin, et al.
Pubblicazione: (2026)
di: Cai, Songjin, et al.
Pubblicazione: (2026)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
di: Wen, Di, et al.
Pubblicazione: (2025)
di: Wen, Di, et al.
Pubblicazione: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
di: Bansal, Siddhant, et al.
Pubblicazione: (2024)
di: Bansal, Siddhant, et al.
Pubblicazione: (2024)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
PromptSR: Cascade Prompting for Lightweight Image Super-Resolution
di: Liu, Wenyang, et al.
Pubblicazione: (2025)
di: Liu, Wenyang, et al.
Pubblicazione: (2025)
CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific Methods
di: Lei, Qinqian, et al.
Pubblicazione: (2025)
di: Lei, Qinqian, et al.
Pubblicazione: (2025)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
di: Chen, Junwen, et al.
Pubblicazione: (2025)
di: Chen, Junwen, et al.
Pubblicazione: (2025)
From Semantics, Scene to Instance-awareness: Distilling Foundation Model for Grounded Open-vocabulary Situation Recognition
di: Cai, Chen, et al.
Pubblicazione: (2025)
di: Cai, Chen, et al.
Pubblicazione: (2025)
HOI4D: A 4D Egocentric Dataset for Category-Level Human-Object Interaction
di: Liu, Yunze, et al.
Pubblicazione: (2022)
di: Liu, Yunze, et al.
Pubblicazione: (2022)
PA-HOI: A Physics-Aware Human and Object Interaction Dataset
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
ContextHOI: Spatial Context Learning for Human-Object Interaction Detection
di: Jia, Mingda, et al.
Pubblicazione: (2024)
di: Jia, Mingda, et al.
Pubblicazione: (2024)
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
di: Wu, Lin, et al.
Pubblicazione: (2025)
di: Wu, Lin, et al.
Pubblicazione: (2025)
AnchorHOI: Zero-shot Generation of 4D Human-Object Interaction via Anchor-based Prior Distillation
di: Dai, Sisi, et al.
Pubblicazione: (2025)
di: Dai, Sisi, et al.
Pubblicazione: (2025)
MultiFuser: Multimodal Fusion Transformer for Enhanced Driver Action Recognition
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
A Structure-aware and Motion-adaptive Framework for 3D Human Pose Estimation with Mamba
di: Lu, Ye, et al.
Pubblicazione: (2025)
di: Lu, Ye, et al.
Pubblicazione: (2025)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
di: Zhang, Zhenhao, et al.
Pubblicazione: (2025)
PersonaHOI: Effortlessly Improving Personalized Face with Human-Object Interaction Generation
di: Hu, Xinting, et al.
Pubblicazione: (2025)
di: Hu, Xinting, et al.
Pubblicazione: (2025)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
di: Ji, Yihong, et al.
Pubblicazione: (2025)
di: Ji, Yihong, et al.
Pubblicazione: (2025)
RoT: Enhancing Large Language Models with Reflection on Search Trees
di: Hui, Wenyang, et al.
Pubblicazione: (2024)
di: Hui, Wenyang, et al.
Pubblicazione: (2024)
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
di: Zhang, Juze, et al.
Pubblicazione: (2024)
di: Zhang, Juze, et al.
Pubblicazione: (2024)
CycleHOI: Improving Human-Object Interaction Detection with Cycle Consistency of Detection and Generation
di: Wang, Yisen, et al.
Pubblicazione: (2024)
di: Wang, Yisen, et al.
Pubblicazione: (2024)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
di: Zeng, Ling-An, et al.
Pubblicazione: (2025)
di: Zeng, Ling-An, et al.
Pubblicazione: (2025)
Video sentence grounding with temporally global textual knowledge
di: Chen, Cai, et al.
Pubblicazione: (2024)
di: Chen, Cai, et al.
Pubblicazione: (2024)
OOD-HOI: Text-Driven 3D Whole-Body Human-Object Interactions Generation Beyond Training Domains
di: Zhang, Yixuan, et al.
Pubblicazione: (2024)
di: Zhang, Yixuan, et al.
Pubblicazione: (2024)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
di: Xue, Zihui, et al.
Pubblicazione: (2024)
di: Xue, Zihui, et al.
Pubblicazione: (2024)
HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance
di: Li, Lei, et al.
Pubblicazione: (2025)
di: Li, Lei, et al.
Pubblicazione: (2025)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs
di: Mascaro, Esteve Valls, et al.
Pubblicazione: (2023)
di: Mascaro, Esteve Valls, et al.
Pubblicazione: (2023)
GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects
di: Li, Shujia, et al.
Pubblicazione: (2025)
di: Li, Shujia, et al.
Pubblicazione: (2025)
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
di: Yang, Panqi, et al.
Pubblicazione: (2025)
di: Yang, Panqi, et al.
Pubblicazione: (2025)
I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions
di: Zhao, Chengfeng, et al.
Pubblicazione: (2023)
di: Zhao, Chengfeng, et al.
Pubblicazione: (2023)
ScoreHOI: Physically Plausible Reconstruction of Human-Object Interaction via Score-Guided Diffusion
di: Li, Ao, et al.
Pubblicazione: (2025)
di: Li, Ao, et al.
Pubblicazione: (2025)
Uni-HOI:A Unified framework for Learning the Joint distribution of Text and Human-Object Interaction
di: Zhang, Mengfei, et al.
Pubblicazione: (2026)
di: Zhang, Mengfei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
OccluTrack: Rethinking Awareness of Occlusion for Enhancing Multiple Pedestrian Tracking
di: Gao, Jianjun, et al.
Pubblicazione: (2023) -
CM2-Net: Continual Cross-Modal Mapping Network for Driver Action Recognition
di: Wang, Ruoyu, et al.
Pubblicazione: (2024) -
SSH-Net: A Self-Supervised and Hybrid Network for Noisy Image Watermark Removal
di: Liu, Wenyang, et al.
Pubblicazione: (2025) -
Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting
di: Cai, Chen, et al.
Pubblicazione: (2024) -
OneHOI: Unifying Human-Object Interaction Generation and Editing
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2026)