VILOD: A Visual Interactive Labeling Tool for Object Detection
Fuente:
arXiv
Saved in:
| Main Author: | Holm, Isac |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Object Recognition in Human Computer Interaction:- A Comparative Analysis
by: Ranade, Kaushik, et al.
Published: (2024)
by: Ranade, Kaushik, et al.
Published: (2024)
Monocular 3D Object Position Estimation with VLMs for Human-Robot Interaction
by: Wahl, Ari, et al.
Published: (2026)
by: Wahl, Ari, et al.
Published: (2026)
Interaction as Explanation: A User Interaction-based Method for Explaining Image Classification Models
by: Yun, Hyeonggeun
Published: (2024)
by: Yun, Hyeonggeun
Published: (2024)
ArtWhisperer: A Dataset for Characterizing Human-AI Interactions in Artistic Creations
by: Vodrahalli, Kailas, et al.
Published: (2023)
by: Vodrahalli, Kailas, et al.
Published: (2023)
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
Agile Deliberation: Concept Deliberation for Subjective Visual Classification
by: Wang, Leijie, et al.
Published: (2025)
by: Wang, Leijie, et al.
Published: (2025)
Beyond One-Size-Fits-All: A Survey of Personalized Affective Computing in Human-Agent Interaction
by: Li, Jialin, et al.
Published: (2023)
by: Li, Jialin, et al.
Published: (2023)
RadioActive: 3D Radiological Interactive Segmentation Benchmark
by: Ulrich, Constantin, et al.
Published: (2024)
by: Ulrich, Constantin, et al.
Published: (2024)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
by: Chaudhari, Shravan, et al.
Published: (2025)
by: Chaudhari, Shravan, et al.
Published: (2025)
Improving Prototypical Visual Explanations with Reward Reweighing, Reselection, and Retraining
by: Li, Aaron J., et al.
Published: (2023)
by: Li, Aaron J., et al.
Published: (2023)
Looking for a better fit? An Incremental Learning Multimodal Object Referencing Framework adapting to Individual Drivers
by: Gomaa, Amr, et al.
Published: (2024)
by: Gomaa, Amr, et al.
Published: (2024)
Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
InterVLS: Interactive Model Understanding and Improvement with Vision-Language Surrogates
by: Huang, Jinbin, et al.
Published: (2023)
by: Huang, Jinbin, et al.
Published: (2023)
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
by: Ye, Yilin, et al.
Published: (2025)
by: Ye, Yilin, et al.
Published: (2025)
MoRE-Brain: Routed Mixture of Experts for Interpretable and Generalizable Cross-Subject fMRI Visual Decoding
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
OpenDriver: An Open-Road Driver State Detection Dataset
by: Liu, Delong, et al.
Published: (2023)
by: Liu, Delong, et al.
Published: (2023)
IVISIT: An Interactive Visual Simulation Tool for system simulation, visualization, optimization, and parameter management
by: Knoblauch, Andreas
Published: (2024)
by: Knoblauch, Andreas
Published: (2024)
HOSt3R: Keypoint-free Hand-Object 3D Reconstruction from RGB images
by: Swamy, Anilkumar, et al.
Published: (2025)
by: Swamy, Anilkumar, et al.
Published: (2025)
BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis
by: Rondelli, Massimo, et al.
Published: (2026)
by: Rondelli, Massimo, et al.
Published: (2026)
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
by: Ki, Taekyung, et al.
Published: (2026)
by: Ki, Taekyung, et al.
Published: (2026)
RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
by: Garg, Sourav, et al.
Published: (2024)
by: Garg, Sourav, et al.
Published: (2024)
HIVE: Harnessing Human Feedback for Instructional Visual Editing
by: Zhang, Shu, et al.
Published: (2023)
by: Zhang, Shu, et al.
Published: (2023)
Lost in Edits? A $λ$-Compass for AIGC Provenance
by: You, Wenhao, et al.
Published: (2025)
by: You, Wenhao, et al.
Published: (2025)
A Foundational Generative Model for Breast Ultrasound Image Analysis
by: Yu, Haojun, et al.
Published: (2025)
by: Yu, Haojun, et al.
Published: (2025)
UISim: An Interactive Image-Based UI Simulator for Dynamic Mobile Environments
by: Xiang, Jiannan, et al.
Published: (2025)
by: Xiang, Jiannan, et al.
Published: (2025)
On the Interpretability of Part-Prototype Based Classifiers: A Human Centric Analysis
by: Davoodi, Omid, et al.
Published: (2023)
by: Davoodi, Omid, et al.
Published: (2023)
A Study of Acquisition Functions for Medical Imaging Deep Active Learning
by: Dossou, Bonaventure F. P.
Published: (2024)
by: Dossou, Bonaventure F. P.
Published: (2024)
Transformers Utilization in Chart Understanding: A Review of Recent Advances & Future Trends
by: Al-Shetairy, Mirna, et al.
Published: (2024)
by: Al-Shetairy, Mirna, et al.
Published: (2024)
Visual Evaluative AI: A Hypothesis-Driven Tool with Concept-Based Explanations and Weight of Evidence
by: Le, Thao, et al.
Published: (2024)
by: Le, Thao, et al.
Published: (2024)
ThermoHands: A Benchmark for 3D Hand Pose Estimation from Egocentric Thermal Images
by: Ding, Fangqiang, et al.
Published: (2024)
by: Ding, Fangqiang, et al.
Published: (2024)
How Good is ChatGPT at Audiovisual Deepfake Detection: A Comparative Study of ChatGPT, AI Models and Human Perception
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
MNIST-Gen: A Modular MNIST-Style Dataset Generation Using Hierarchical Semantics, Reinforcement Learning, and Category Theory
by: Shaeri, Pouya, et al.
Published: (2025)
by: Shaeri, Pouya, et al.
Published: (2025)
VocalEyes: Enhancing Environmental Perception for the Visually Impaired through Vision-Language Models and Distance-Aware Object Detection
by: Chavan, Kunal, et al.
Published: (2025)
by: Chavan, Kunal, et al.
Published: (2025)
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
Screen2AX: Vision-Based Approach for Automatic macOS Accessibility Generation
by: Muryn, Viktor, et al.
Published: (2025)
by: Muryn, Viktor, et al.
Published: (2025)
Human-like visual computing advances explainability and few-shot learning in deep neural networks for complex physiological data
by: Alahmadi, Alaa, et al.
Published: (2025)
by: Alahmadi, Alaa, et al.
Published: (2025)
Facial Analysis Systems and Down Syndrome
by: Rondina, Marco, et al.
Published: (2025)
by: Rondina, Marco, et al.
Published: (2025)
Learning with Category-Equivariant Representations for Human Activity Recognition
by: Maruyama, Yoshihiro
Published: (2025)
by: Maruyama, Yoshihiro
Published: (2025)
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
Similar Items
-
Object Recognition in Human Computer Interaction:- A Comparative Analysis
by: Ranade, Kaushik, et al.
Published: (2024) -
Monocular 3D Object Position Estimation with VLMs for Human-Robot Interaction
by: Wahl, Ari, et al.
Published: (2026) -
Interaction as Explanation: A User Interaction-based Method for Explaining Image Classification Models
by: Yun, Hyeonggeun
Published: (2024) -
ArtWhisperer: A Dataset for Characterizing Human-AI Interactions in Artistic Creations
by: Vodrahalli, Kailas, et al.
Published: (2023) -
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
by: Nguyen, Tuong Vy, et al.
Published: (2024)