Saved in:
| Main Authors: | de Witte, Sven, Strafforello, Ombretta, van Gemert, Jan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2401.17821 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aligning Object Detector Bounding Boxes with Human Preference
by: Strafforello, Ombretta, et al.
Published: (2024)
by: Strafforello, Ombretta, et al.
Published: (2024)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
by: Kim, Namhee, et al.
Published: (2025)
by: Kim, Namhee, et al.
Published: (2025)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025)
by: Duan, Junwen, et al.
Published: (2025)
Algorithmic Ways of Seeing: Using Object Detection to Facilitate Art Exploration
by: Meyer, Louie Søs, et al.
Published: (2024)
by: Meyer, Louie Søs, et al.
Published: (2024)
A Comparison of Human and Machine Learning Errors in Face Recognition
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
by: Nowicki, Filip, et al.
Published: (2026)
by: Nowicki, Filip, et al.
Published: (2026)
Adaptive Modality Balanced Online Knowledge Distillation for Brain-Eye-Computer based Dim Object Detection
by: Li, Zixing, et al.
Published: (2024)
by: Li, Zixing, et al.
Published: (2024)
GentleHumanoid: Learning Upper-body Compliance for Contact-rich Human and Object Interaction
by: Lu, Qingzhou, et al.
Published: (2025)
by: Lu, Qingzhou, et al.
Published: (2025)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
Multimodal Attention-Aware Fusion for Diagnosing Distal Myopathy: Evaluating Model Interpretability and Clinician Trust
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
Do Vision Language Models Understand Human Engagement in Games?
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
SVFAP: Self-supervised Video Facial Affect Perceiver
by: Sun, Licai, et al.
Published: (2023)
by: Sun, Licai, et al.
Published: (2023)
"It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language Models
by: Garg, Kapil, et al.
Published: (2025)
by: Garg, Kapil, et al.
Published: (2025)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
AccessLens: Auto-detecting Inaccessibility of Everyday Objects
by: Kwon, Nahyun, et al.
Published: (2024)
by: Kwon, Nahyun, et al.
Published: (2024)
Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision
by: Li, Chentao, et al.
Published: (2026)
by: Li, Chentao, et al.
Published: (2026)
Computer Vision for Objects used in Group Work: Challenges and Opportunities
by: Jung, Changsoo, et al.
Published: (2025)
by: Jung, Changsoo, et al.
Published: (2025)
Ninja Codes: Neurally Generated Fiducial Markers for Stealthy 6-DoF Tracking
by: Takeuchi, Yuichiro, et al.
Published: (2025)
by: Takeuchi, Yuichiro, et al.
Published: (2025)
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
by: Guo, Hao, et al.
Published: (2025)
by: Guo, Hao, et al.
Published: (2025)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
by: Islam, Md Touhidul, et al.
Published: (2024)
by: Islam, Md Touhidul, et al.
Published: (2024)
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
Deep Learning-based Lightweight RGB Object Tracking for Augmented Reality Devices
by: Smith, Alice, et al.
Published: (2025)
by: Smith, Alice, et al.
Published: (2025)
ConceptFactory: Facilitate 3D Object Knowledge Annotation with Object Conceptualization
by: Sun, Jianhua, et al.
Published: (2024)
by: Sun, Jianhua, et al.
Published: (2024)
SpriteHand: Real-Time Versatile Hand-Object Interaction with Autoregressive Video Generation
by: Li, Zisu, et al.
Published: (2025)
by: Li, Zisu, et al.
Published: (2025)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
Towards an End-to-End System for 3D Tracking of Physical Objects in Virtual Immersive Environments
by: Knapiński, Stanisław, et al.
Published: (2026)
by: Knapiński, Stanisław, et al.
Published: (2026)
VFA: Vision Frequency Analysis of Foundation Models and Human
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
Extracting Human Attention through Crowdsourced Patch Labeling
by: Chang, Minsuk, et al.
Published: (2024)
by: Chang, Minsuk, et al.
Published: (2024)
OccRobNet : Occlusion Robust Network for Accurate 3D Interacting Hand-Object Pose Estimation
by: Garg, Mallika, et al.
Published: (2025)
by: Garg, Mallika, et al.
Published: (2025)
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
by: Song, Sicheng, et al.
Published: (2025)
by: Song, Sicheng, et al.
Published: (2025)
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026)
by: Shaw, Richard, et al.
Published: (2026)
MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
by: He, Yuchen, et al.
Published: (2025)
by: He, Yuchen, et al.
Published: (2025)
Analyzing Swimming Performance Using Drone Captured Aerial Videos
by: Tran, Thu, et al.
Published: (2025)
by: Tran, Thu, et al.
Published: (2025)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
LiDAR-based Human Activity Recognition through Laplacian Spectral Analysis
by: Sharifipour, Sasan, et al.
Published: (2025)
by: Sharifipour, Sasan, et al.
Published: (2025)
Similar Items
-
Aligning Object Detector Bounding Boxes with Human Preference
by: Strafforello, Ombretta, et al.
Published: (2024) -
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025) -
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
by: Kim, Namhee, et al.
Published: (2025) -
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025) -
Algorithmic Ways of Seeing: Using Object Detection to Facilitate Art Exploration
by: Meyer, Louie Søs, et al.
Published: (2024)