Mining Instance-Centric Vision-Language Contexts for Human-Object Interaction Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Seo, Soo Won, Lee, KyungChae, Cho, Hyungchan, Son, Taein, Cho, Nam Ik, Choi, Jun Won |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts
por: Son, Taein, et al.
Publicado: (2024)
por: Son, Taein, et al.
Publicado: (2024)
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
por: Lee, Seok Hwan, et al.
Publicado: (2024)
por: Lee, Seok Hwan, et al.
Publicado: (2024)
Hermit Kingdom Through the Lens of Multiple Perspectives: A Case Study of LLM Hallucination on North Korea
por: Cho, Eunjung, et al.
Publicado: (2025)
por: Cho, Eunjung, et al.
Publicado: (2025)
Behavior Modeling for Training-free Building of Private Domain Multi Agent System
por: Cho, Won Ik, et al.
Publicado: (2025)
por: Cho, Won Ik, et al.
Publicado: (2025)
Evaluating Span Extraction in Generative Paradigm: A Reflection on Aspect-Based Sentiment Analysis
por: Yang, Soyoung, et al.
Publicado: (2024)
por: Yang, Soyoung, et al.
Publicado: (2024)
Three Disclaimers for Safe Disclosure: A Cardwriter for Reporting the Use of Generative AI in Writing Process
por: Cho, Won Ik, et al.
Publicado: (2024)
por: Cho, Won Ik, et al.
Publicado: (2024)
RICoTA: Red-teaming of In-the-wild Conversation with Test Attempts
por: Choi, Eujeong, et al.
Publicado: (2025)
por: Choi, Eujeong, et al.
Publicado: (2025)
Semantic Watermarking Reinvented: Enhancing Robustness and Generation Quality with Fourier Integrity
por: Lee, Sung Ju, et al.
Publicado: (2025)
por: Lee, Sung Ju, et al.
Publicado: (2025)
PhaseMark: A Post-hoc, Optimization-Free Watermarking of AI-generated Images in the Latent Frequency Domain
por: Lee, Sung Ju, et al.
Publicado: (2026)
por: Lee, Sung Ju, et al.
Publicado: (2026)
Leveraging Positional Encoding for Robust Multi-Reference-Based Object 6D Pose Estimation
por: Park, Jaewoo, et al.
Publicado: (2024)
por: Park, Jaewoo, et al.
Publicado: (2024)
RS-Net: Context-Aware Relation Scoring for Dynamic Scene Graph Generation
por: Jo, Hae-Won, et al.
Publicado: (2025)
por: Jo, Hae-Won, et al.
Publicado: (2025)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
por: Yang, Soyoung, et al.
Publicado: (2024)
por: Yang, Soyoung, et al.
Publicado: (2024)
RefPose: Leveraging Reference Geometric Correspondences for Accurate 6D Pose Estimation of Unseen Objects
por: Kim, Jaeguk, et al.
Publicado: (2025)
por: Kim, Jaeguk, et al.
Publicado: (2025)
Position: Adopting AI in Practice Does Not Guarantee the Productivity Boost
por: Cho, Won Ik, et al.
Publicado: (2026)
por: Cho, Won Ik, et al.
Publicado: (2026)
Enhancing Multi-Exposure High Dynamic Range Imaging with Overlapped Codebook for Improved Representation Learning
por: Lee, Keuntek, et al.
Publicado: (2025)
por: Lee, Keuntek, et al.
Publicado: (2025)
Lightweight and Fast Real-time Image Enhancement via Decomposition of the Spatial-aware Lookup Tables
por: Kim, Wontae, et al.
Publicado: (2025)
por: Kim, Wontae, et al.
Publicado: (2025)
Continuity of solutions to complex Monge-Ampère equations on compact Kähler spaces
por: Cho, Ye-Won Luke, et al.
Publicado: (2024)
por: Cho, Ye-Won Luke, et al.
Publicado: (2024)
From National Curricula to Cultural Awareness: Constructing Open-Ended Culture-Specific Question Answering Dataset
por: Yoo, Haneul, et al.
Publicado: (2026)
por: Yoo, Haneul, et al.
Publicado: (2026)
AGB and Post-AGB Stars versus Planetary Nebulae and Young Stellar Objects: Properties in Visual and IR Bands
por: Suh, Kyung-Won
Publicado: (2024)
por: Suh, Kyung-Won
Publicado: (2024)
Cellular architecture in metastasis leads to an aggressive cancer environment by spatial transcriptomics
por: Cho, Jae Won
Publicado: (2025)
por: Cho, Jae Won
Publicado: (2025)
STONE Dataset: A Scalable Multi-Modal Surround-View 3D Traversability Dataset for Off-Road Robot Navigation
por: Park, Konyul, et al.
Publicado: (2026)
por: Park, Konyul, et al.
Publicado: (2026)
MATT-GS: Masked Attention-based 3DGS for Robot Perception and Object Detection
por: Lee, Jee Won, et al.
Publicado: (2025)
por: Lee, Jee Won, et al.
Publicado: (2025)
A Sharp Lower Bound for the Spectrum of the Hodge Laplacian on Kähler Hyperbolic Manifolds and its Applications
por: Cho, Ye-Won Luke, et al.
Publicado: (2026)
por: Cho, Ye-Won Luke, et al.
Publicado: (2026)
CSF-Net: Context-Semantic Fusion Network for Large Mask Inpainting
por: Heo, Chae-Yeon, et al.
Publicado: (2025)
por: Heo, Chae-Yeon, et al.
Publicado: (2025)
Locality-Aware Zero-Shot Human-Object Interaction Detection
por: Kim, Sanghyun, et al.
Publicado: (2025)
por: Kim, Sanghyun, et al.
Publicado: (2025)
Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion
por: Jang, Oh-Tae, et al.
Publicado: (2025)
por: Jang, Oh-Tae, et al.
Publicado: (2025)
TM-BSN: Triangular-Masked Blind-Spot Network for Real-World Self-Supervised Image Denoising
por: Park, Junyoung, et al.
Publicado: (2026)
por: Park, Junyoung, et al.
Publicado: (2026)
DINOLight: Robust Ambient Light Normalization with Self-supervised Visual Prior Integration
por: Oh, Youngjin, et al.
Publicado: (2026)
por: Oh, Youngjin, et al.
Publicado: (2026)
Opioid‐Induced Myoclonus During Premedication in Two Dogs With Hydrocephalus Proactively Managed With Antiepileptic Medications
por: Changhoon Nam, et al.
Publicado: (2025)
por: Changhoon Nam, et al.
Publicado: (2025)
Dataset Distillation for Super-Resolution without Class Labels and Pre-trained Models
por: Cho, Sunwoo, et al.
Publicado: (2025)
por: Cho, Sunwoo, et al.
Publicado: (2025)
Functional Regression with Nonstationarity and Error Contamination: Application to the Economic Impact of Climate Change
por: Nam, Kyungsik, et al.
Publicado: (2025)
por: Nam, Kyungsik, et al.
Publicado: (2025)
Nonlinear Temperature Sensitivity of Residential Electricity Demand: Evidence from a Distributional Regression Approach
por: Nam, Kyungsik, et al.
Publicado: (2025)
por: Nam, Kyungsik, et al.
Publicado: (2025)
Towards Controllable Real Image Denoising with Camera Parameters
por: Oh, Youngjin, et al.
Publicado: (2025)
por: Oh, Youngjin, et al.
Publicado: (2025)
Pixel2Catch: Multi-Agent Sim-to-Real Transfer for Agile Manipulation with a Single RGB Camera
por: Kim, Seongyong, et al.
Publicado: (2026)
por: Kim, Seongyong, et al.
Publicado: (2026)
MoLT: Mixture of Layer-Wise Tokens for Efficient Audio-Visual Learning
por: Rho, Kyeongha, et al.
Publicado: (2025)
por: Rho, Kyeongha, et al.
Publicado: (2025)
Asymmetrical Functionalization of Polarizable Interface Restructuring Molecules for Rapid and Longer Operative Lithium Metal Batteries
por: Chae Yeong Son, et al.
Publicado: (2024)
por: Chae Yeong Son, et al.
Publicado: (2024)
Asymmetrical Functionalization of Polarizable Interface Restructuring Molecules for Rapid and Longer Operative Lithium Metal Batteries (Small 51/2024)
por: Chae Yeong Son, et al.
Publicado: (2024)
por: Chae Yeong Son, et al.
Publicado: (2024)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
por: Seo, Hoigi, et al.
Publicado: (2025)
por: Seo, Hoigi, et al.
Publicado: (2025)
Room‐Temperature, High‐Resolution Soft Anisotropic Conductive Film for Electrical Interfacing in Stretchable Electronics
por: Jun Choi, et al.
Publicado: (2025)
por: Jun Choi, et al.
Publicado: (2025)
CRT-Fusion: Camera, Radar, Temporal Fusion Using Motion Information for 3D Object Detection
por: Kim, Jisong, et al.
Publicado: (2024)
por: Kim, Jisong, et al.
Publicado: (2024)
Ejemplares similares
-
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts
por: Son, Taein, et al.
Publicado: (2024) -
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
por: Lee, Seok Hwan, et al.
Publicado: (2024) -
Hermit Kingdom Through the Lens of Multiple Perspectives: A Case Study of LLM Hallucination on North Korea
por: Cho, Eunjung, et al.
Publicado: (2025) -
Behavior Modeling for Training-free Building of Private Domain Multi Agent System
por: Cho, Won Ik, et al.
Publicado: (2025) -
Evaluating Span Extraction in Generative Paradigm: A Reflection on Aspect-Based Sentiment Analysis
por: Yang, Soyoung, et al.
Publicado: (2024)