QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yuxiao, Liang, Wolin, Lei, Yu, Xue, Weiying, Zhuang, Nan, Liu, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FreeA: Human-object Interaction Detection using Free Annotation Labels
by: Liu, Qi, et al.
Published: (2024)
by: Liu, Qi, et al.
Published: (2024)
Changing the Paradigm from Dynamic Queries to LLM-generated SQL Queries with Human Intervention
by: Assor, Ambre, et al.
Published: (2025)
by: Assor, Ambre, et al.
Published: (2025)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025)
by: Duan, Junwen, et al.
Published: (2025)
QueryGenie: Making LLM-Based Database Querying Transparent and Controllable
by: Chen, Longfei, et al.
Published: (2025)
by: Chen, Longfei, et al.
Published: (2025)
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
What-Meets-Where: Unified Learning of Action and Contact Localization in Images
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
Do Object Detection Localization Errors Affect Human Performance and Trust?
by: de Witte, Sven, et al.
Published: (2024)
by: de Witte, Sven, et al.
Published: (2024)
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
by: Liu, Can, et al.
Published: (2025)
by: Liu, Can, et al.
Published: (2025)
Memory Remedy: An AI-Enhanced Interactive Story Exploring Human-Robot Interaction and Companionship
by: Han, Lei, et al.
Published: (2024)
by: Han, Lei, et al.
Published: (2024)
Towards Directive Explanations: Crafting Explainable AI Systems for Actionable Human-AI Interactions
by: Bhattacharya, Aditya
Published: (2023)
by: Bhattacharya, Aditya
Published: (2023)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
FontCraft: Multimodal Font Design Using Interactive Bayesian Optimization
by: Tatsukawa, Yuki, et al.
Published: (2025)
by: Tatsukawa, Yuki, et al.
Published: (2025)
SpriteHand: Real-Time Versatile Hand-Object Interaction with Autoregressive Video Generation
by: Li, Zisu, et al.
Published: (2025)
by: Li, Zisu, et al.
Published: (2025)
GentleHumanoid: Learning Upper-body Compliance for Contact-rich Human and Object Interaction
by: Lu, Qingzhou, et al.
Published: (2025)
by: Lu, Qingzhou, et al.
Published: (2025)
A Taxonomy for Human-LLM Interaction Modes: An Initial Exploration
by: Gao, Jie, et al.
Published: (2024)
by: Gao, Jie, et al.
Published: (2024)
MagicCraft: Natural Language-Driven Generation of Dynamic and Interactive 3D Objects for Commercial Metaverse Platforms
by: Kurai, Ryutaro, et al.
Published: (2025)
by: Kurai, Ryutaro, et al.
Published: (2025)
Crepe: A Mobile Screen Data Collector Using Graph Query
by: Lu, Yuwen, et al.
Published: (2024)
by: Lu, Yuwen, et al.
Published: (2024)
A Review of Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
Relation-driven Query of Multiple Time Series
by: Liu, Shuhan, et al.
Published: (2023)
by: Liu, Shuhan, et al.
Published: (2023)
Envisage: Towards Expressive Visual Graph Querying
by: Wen, Xiaolin, et al.
Published: (2025)
by: Wen, Xiaolin, et al.
Published: (2025)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
VerSe: Integrating Multiple Queries as Prompts for Versatile Cardiac MRI Segmentation
by: Guo, Bangwei, et al.
Published: (2024)
by: Guo, Bangwei, et al.
Published: (2024)
Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
Fuzzy Ontology Embeddings and Visual Query Building for Ontology Exploration
by: Zhurov, Vladimir, et al.
Published: (2025)
by: Zhurov, Vladimir, et al.
Published: (2025)
GenQuery: Supporting Expressive Visual Search with Generative Models
by: Son, Kihoon, et al.
Published: (2023)
by: Son, Kihoon, et al.
Published: (2023)
ViT-Explainer: An Interactive Walkthrough of the Vision Transformer Pipeline
by: Hernandez, Juan Manuel, et al.
Published: (2026)
by: Hernandez, Juan Manuel, et al.
Published: (2026)
G-VOILA: Gaze-Facilitated Information Querying in Daily Scenarios
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
LacAIDes: Generative AI-Supported Creative Interactive Circuits Crafting to Enliven Traditional Lacquerware
by: Li, Yaning, et al.
Published: (2025)
by: Li, Yaning, et al.
Published: (2025)
DeepSORT-Driven Visual Tracking Approach for Gesture Recognition in Interactive Systems
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
Multi-Masked Querying Network for Robust Emotion Recognition from Incomplete Multi-Modal Physiological Signals
by: Xu, Geng-Xin, et al.
Published: (2025)
by: Xu, Geng-Xin, et al.
Published: (2025)
Adaptive Human-Computer Interaction Strategies Through Reinforcement Learning in Complex
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
Leveraging Foundation Models for Crafting Narrative Visualization: A Survey
by: He, Yi, et al.
Published: (2024)
by: He, Yi, et al.
Published: (2024)
MAPWise: Evaluating Vision-Language Models for Advanced Map Queries
by: Mukhopadhyay, Srija, et al.
Published: (2024)
by: Mukhopadhyay, Srija, et al.
Published: (2024)
Real-Time Prediction for Athletes' Psychological States Using BERT-XGBoost: Enhancing Human-Computer Interaction
by: Duan, Chenming, et al.
Published: (2024)
by: Duan, Chenming, et al.
Published: (2024)
PoseAugment: Generative Human Pose Data Augmentation with Physical Plausibility for IMU-based Motion Capture
by: Li, Zhuojun, et al.
Published: (2024)
by: Li, Zhuojun, et al.
Published: (2024)
Smoothing Grounding and Reasoning for MLLM-Powered GUI Agents with Query-Oriented Pivot Tasks
by: Wu, Zongru, et al.
Published: (2025)
by: Wu, Zongru, et al.
Published: (2025)
Algorithmic Ways of Seeing: Using Object Detection to Facilitate Art Exploration
by: Meyer, Louie Søs, et al.
Published: (2024)
by: Meyer, Louie Søs, et al.
Published: (2024)
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026)
by: Shaw, Richard, et al.
Published: (2026)
Similar Items
-
FreeA: Human-object Interaction Detection using Free Annotation Labels
by: Liu, Qi, et al.
Published: (2024) -
Changing the Paradigm from Dynamic Queries to LLM-generated SQL Queries with Human Intervention
by: Assor, Ambre, et al.
Published: (2025) -
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025) -
QueryGenie: Making LLM-Based Database Querying Transparent and Controllable
by: Chen, Longfei, et al.
Published: (2025) -
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
by: Tian, Yuan, et al.
Published: (2024)