A Plug-and-Play Method for Rare Human-Object Interactions Detection by Bridging Domain Gap
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Lijun, Suo, Wei, Wang, Peng, Zhang, Yanning |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
C3L: Content Correlated Vision-Language Instruction Tuning Data Generation via Contrastive Learning
por: Ma, Ji, et al.
Publicado: (2024)
por: Ma, Ji, et al.
Publicado: (2024)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
por: Ma, Ji, et al.
Publicado: (2025)
por: Ma, Ji, et al.
Publicado: (2025)
Understanding and Mitigating Hallucinations in Multimodal Chain-of-Thought Models
por: Ma, Ji, et al.
Publicado: (2026)
por: Ma, Ji, et al.
Publicado: (2026)
Hallucination-aware intermediate representation edit in large vision-language models
por: Suo, Wei, et al.
Publicado: (2026)
por: Suo, Wei, et al.
Publicado: (2026)
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
por: Suo, Wei, et al.
Publicado: (2025)
por: Suo, Wei, et al.
Publicado: (2025)
Plug and Play Active Learning for Object Detection
por: Yang, Chenhongyi, et al.
Publicado: (2022)
por: Yang, Chenhongyi, et al.
Publicado: (2022)
MQADet: A Plug-and-Play Paradigm for Enhancing Open-Vocabulary Object Detection via Multimodal Question Answering
por: Li, Caixiong, et al.
Publicado: (2025)
por: Li, Caixiong, et al.
Publicado: (2025)
FreDFT: Frequency Domain Fusion Transformer for Visible-Infrared Object Detection
por: Wu, Wencong, et al.
Publicado: (2025)
por: Wu, Wencong, et al.
Publicado: (2025)
PalmBridge: A Plug-and-Play Feature Alignment Framework for Open-Set Palmprint Verification
por: Zhang, Chenke, et al.
Publicado: (2026)
por: Zhang, Chenke, et al.
Publicado: (2026)
GlovEgo-HOI: Bridging the Synthetic-to-Real Gap for Industrial Egocentric Human-Object Interaction Detection
por: Spoto, Alfio, et al.
Publicado: (2026)
por: Spoto, Alfio, et al.
Publicado: (2026)
Large Self-Supervised Models Bridge the Gap in Domain Adaptive Object Detection
por: Lavoie, Marc-Antoine, et al.
Publicado: (2025)
por: Lavoie, Marc-Antoine, et al.
Publicado: (2025)
Towards Accurate Camouflaged Object Detection with Mixture Convolution and Interactive Fusion
por: Chen, Geng, et al.
Publicado: (2021)
por: Chen, Geng, et al.
Publicado: (2021)
CBNet: A Plug-and-Play Network for Segmentation-Based Scene Text Detection
por: Zhao, Xi, et al.
Publicado: (2022)
por: Zhao, Xi, et al.
Publicado: (2022)
From Camera to World: A Plug-and-Play Module for Human Mesh Transformation
por: Ma, Changhai, et al.
Publicado: (2025)
por: Ma, Changhai, et al.
Publicado: (2025)
Investigating Domain Gaps for Indoor 3D Object Detection
por: Zhao, Zijing, et al.
Publicado: (2025)
por: Zhao, Zijing, et al.
Publicado: (2025)
Plug-and-Play Diffusion Distillation
por: Hsiao, Yi-Ting, et al.
Publicado: (2024)
por: Hsiao, Yi-Ting, et al.
Publicado: (2024)
Egocentric Human-Object Interaction Detection: A New Benchmark and Method
por: Deng, Kunyuan, et al.
Publicado: (2025)
por: Deng, Kunyuan, et al.
Publicado: (2025)
Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models
por: Suo, Wei, et al.
Publicado: (2024)
por: Suo, Wei, et al.
Publicado: (2024)
Mitigating Information Loss under High Pruning Rates for Efficient Large Vision Language Models
por: Fu, Mingyu, et al.
Publicado: (2025)
por: Fu, Mingyu, et al.
Publicado: (2025)
Visual Prompt Selection for In-Context Learning Segmentation
por: Suo, Wei, et al.
Publicado: (2024)
por: Suo, Wei, et al.
Publicado: (2024)
PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments
por: Peng, Kebin, et al.
Publicado: (2024)
por: Peng, Kebin, et al.
Publicado: (2024)
Single-Shot Plug-and-Play Methods for Inverse Problems
por: Cheng, Yanqi, et al.
Publicado: (2023)
por: Cheng, Yanqi, et al.
Publicado: (2023)
OOD-HOI: Text-Driven 3D Whole-Body Human-Object Interactions Generation Beyond Training Domains
por: Zhang, Yixuan, et al.
Publicado: (2024)
por: Zhang, Yixuan, et al.
Publicado: (2024)
SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene Flow
por: Wang, Qingyuan, et al.
Publicado: (2025)
por: Wang, Qingyuan, et al.
Publicado: (2025)
FEDEXCHANGE: Bridging the Domain Gap in Federated Object Detection for Free
por: Yuan, Haolin, et al.
Publicado: (2025)
por: Yuan, Haolin, et al.
Publicado: (2025)
Hoi2Threat: An Interpretable Threat Detection Method for Human Violence Scenarios Guided by Human-Object Interaction
por: Wang, Yuhan, et al.
Publicado: (2025)
por: Wang, Yuhan, et al.
Publicado: (2025)
Vote&Mix: Plug-and-Play Token Reduction for Efficient Vision Transformer
por: Peng, Shuai, et al.
Publicado: (2024)
por: Peng, Shuai, et al.
Publicado: (2024)
A Unified Plug-and-Play Algorithm with Projected Landweber Operator for Split Convex Feasibility Problems
por: Zhang, Shuchang, et al.
Publicado: (2024)
por: Zhang, Shuchang, et al.
Publicado: (2024)
An Image-like Diffusion Method for Human-Object Interaction Detection
por: Hui, Xiaofei, et al.
Publicado: (2025)
por: Hui, Xiaofei, et al.
Publicado: (2025)
Prototype Embedding Optimization for Human-Object Interaction Detection in Livestreaming
por: Zhang, Menghui, et al.
Publicado: (2025)
por: Zhang, Menghui, et al.
Publicado: (2025)
LAB-Det: Language as a Domain-Invariant Bridge for Training-Free One-Shot Domain Generalization in Object Detection
por: Zhang, Xu, et al.
Publicado: (2026)
por: Zhang, Xu, et al.
Publicado: (2026)
GS-Net: Generalizable Plug-and-Play 3D Gaussian Splatting Module
por: Zhang, Yichen, et al.
Publicado: (2024)
por: Zhang, Yichen, et al.
Publicado: (2024)
ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop
por: Li, Shuangzhi, et al.
Publicado: (2026)
por: Li, Shuangzhi, et al.
Publicado: (2026)
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
por: Xue, Bowen, et al.
Publicado: (2025)
por: Xue, Bowen, et al.
Publicado: (2025)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
por: Lee, Saehyung, et al.
Publicado: (2024)
por: Lee, Saehyung, et al.
Publicado: (2024)
Finding Dino: A Plug-and-Play Framework for Zero-Shot Detection of Out-of-Distribution Objects Using Prototypes
por: Sinhamahapatra, Poulami, et al.
Publicado: (2024)
por: Sinhamahapatra, Poulami, et al.
Publicado: (2024)
Meta-Exploiting Frequency Prior for Cross-Domain Few-Shot Learning
por: Zhou, Fei, et al.
Publicado: (2024)
por: Zhou, Fei, et al.
Publicado: (2024)
Knowing the Unknown: Interpretable Open-World Object Detection via Concept Decomposition Model
por: Lv, Xueqiang, et al.
Publicado: (2026)
por: Lv, Xueqiang, et al.
Publicado: (2026)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
por: Bian, Yuxuan, et al.
Publicado: (2024)
por: Bian, Yuxuan, et al.
Publicado: (2024)
Bridging Annotation Gaps: Transferring Labels to Align Object Detection Datasets
por: Kennerley, Mikhail, et al.
Publicado: (2025)
por: Kennerley, Mikhail, et al.
Publicado: (2025)
Ejemplares similares
-
C3L: Content Correlated Vision-Language Instruction Tuning Data Generation via Contrastive Learning
por: Ma, Ji, et al.
Publicado: (2024) -
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
por: Ma, Ji, et al.
Publicado: (2025) -
Understanding and Mitigating Hallucinations in Multimodal Chain-of-Thought Models
por: Ma, Ji, et al.
Publicado: (2026) -
Hallucination-aware intermediate representation edit in large vision-language models
por: Suo, Wei, et al.
Publicado: (2026) -
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
por: Suo, Wei, et al.
Publicado: (2025)