Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, He, Fu, Xinyi, Carroll, John M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-shot Emotion Annotation in Facial Images Using Large Multimodal Models: Benchmarking and Prospects for Multi-Class, Multi-Frame Approaches
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
Generative AI Framework for 3D Object Generation in Augmented Reality
by: Behravan, Majid
Published: (2025)
by: Behravan, Majid
Published: (2025)
Human-in-the-Loop Annotation for Image-Based Engagement Estimation: Assessing the Impact of Model Reliability on Annotation Accuracy
by: Subramanya, Sahana Yadnakudige, et al.
Published: (2025)
by: Subramanya, Sahana Yadnakudige, et al.
Published: (2025)
K-Sort Arena: Efficient and Reliable Benchmarking for Generative Models via K-wise Human Preferences
by: Li, Zhikai, et al.
Published: (2024)
by: Li, Zhikai, et al.
Published: (2024)
Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
by: Shen, Junxiao, et al.
Published: (2023)
by: Shen, Junxiao, et al.
Published: (2023)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026)
by: Baek, Injun, et al.
Published: (2026)
Generative Augmented Reality: Paradigms, Technologies, and Future Applications
by: Liang, Chen, et al.
Published: (2025)
by: Liang, Chen, et al.
Published: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
by: Song, Jookyung, et al.
Published: (2025)
by: Song, Jookyung, et al.
Published: (2025)
Radar-APLANC: Unsupervised Radar-based Heartbeat Sensing via Augmented Pseudo-Label and Noise Contrast
by: Wang, Ying, et al.
Published: (2025)
by: Wang, Ying, et al.
Published: (2025)
Effective Guidance for Model Attention with Simple Yes-no Annotations
by: Lee, Seongmin, et al.
Published: (2024)
by: Lee, Seongmin, et al.
Published: (2024)
Towards a Humanized Social-Media Ecosystem: AI-Augmented HCI Design Patterns for Safety, Agency & Well-Being
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
VILOD: A Visual Interactive Labeling Tool for Object Detection
by: Holm, Isac
Published: (2025)
by: Holm, Isac
Published: (2025)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Gaze Detection and Analysis for Initiating Joint Activity in Industrial Human-Robot Collaboration
by: Prajod, Pooja, et al.
Published: (2023)
by: Prajod, Pooja, et al.
Published: (2023)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
by: Yakolli, Nivedan, et al.
Published: (2025)
by: Yakolli, Nivedan, et al.
Published: (2025)
Coral Model Generation from Single Images for Virtual Reality Applications
by: Fu, Jie, et al.
Published: (2024)
by: Fu, Jie, et al.
Published: (2024)
LCE: A Framework for Explainability of DNNs for Ultrasound Image Based on Concept Discovery
by: Kong, Weiji, et al.
Published: (2024)
by: Kong, Weiji, et al.
Published: (2024)
Using Text-to-Image Generation for Architectural Design Ideation
by: Paananen, Ville, et al.
Published: (2023)
by: Paananen, Ville, et al.
Published: (2023)
Characterizing Photorealism and Artifacts in Diffusion Model-Generated Images
by: Kamali, Negar, et al.
Published: (2025)
by: Kamali, Negar, et al.
Published: (2025)
Expansion Quantization Network: An Efficient Micro-emotion Annotation and Detection Framework
by: Zhou, Jingyi, et al.
Published: (2024)
by: Zhou, Jingyi, et al.
Published: (2024)
How to Distinguish AI-Generated Images from Authentic Photographs
by: Kamali, Negar, et al.
Published: (2024)
by: Kamali, Negar, et al.
Published: (2024)
PixelWeb: The First Web GUI Dataset with Pixel-Wise Labels
by: Yang, Qi, et al.
Published: (2025)
by: Yang, Qi, et al.
Published: (2025)
Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators
by: Ramaswamy, Aditi, et al.
Published: (2024)
by: Ramaswamy, Aditi, et al.
Published: (2024)
A Picture is Worth a Thousand Prompts? Efficacy of Iterative Human-Driven Prompt Refinement in Image Regeneration Tasks
by: Trinh, Khoi, et al.
Published: (2025)
by: Trinh, Khoi, et al.
Published: (2025)
LEyes: A Lightweight Framework for Deep Learning-Based Eye Tracking using Synthetic Eye Images
by: Byrne, Sean Anthony, et al.
Published: (2023)
by: Byrne, Sean Anthony, et al.
Published: (2023)
SigmaCollab: An Application-Driven Dataset for Physically Situated Collaboration
by: Bohus, Dan, et al.
Published: (2025)
by: Bohus, Dan, et al.
Published: (2025)
VRMN-bD: A Multi-modal Natural Behavior Dataset of Immersive Human Fear Responses in VR Stand-up Interactive Games
by: Zhang, He, et al.
Published: (2024)
by: Zhang, He, et al.
Published: (2024)
Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation
by: De Simone, Zoe, et al.
Published: (2026)
by: De Simone, Zoe, et al.
Published: (2026)
SynthoGestures: A Novel Framework for Synthetic Dynamic Hand Gesture Generation for Driving Scenarios
by: Gomaa, Amr, et al.
Published: (2023)
by: Gomaa, Amr, et al.
Published: (2023)
Regressor-Guided Generative Image Editing Balances User Emotions to Reduce Time Spent Online
by: Gebhardt, Christoph, et al.
Published: (2025)
by: Gebhardt, Christoph, et al.
Published: (2025)
AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People
by: Lee, Seonghee, et al.
Published: (2024)
by: Lee, Seonghee, et al.
Published: (2024)
Towards Human-AI Collaboration System for the Detection of Invasive Ductal Carcinoma in Histopathology Images
by: Han, Shuo, et al.
Published: (2025)
by: Han, Shuo, et al.
Published: (2025)
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
by: Wang, Zaitian, et al.
Published: (2025)
by: Wang, Zaitian, et al.
Published: (2025)
Yume: An Interactive World Generation Model
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
SASG-DA: Sparse-Aware Semantic-Guided Diffusion Augmentation For Myoelectric Gesture Recognition
by: Liu, Chen, et al.
Published: (2025)
by: Liu, Chen, et al.
Published: (2025)
Enabling Collaborative Clinical Diagnosis of Infectious Keratitis by Integrating Expert Knowledge and Interpretable Data-driven Intelligence
by: Fang, Zhengqing, et al.
Published: (2024)
by: Fang, Zhengqing, et al.
Published: (2024)
HuLP: Human-in-the-Loop for Prognosis
by: Ridzuan, Muhammad, et al.
Published: (2024)
by: Ridzuan, Muhammad, et al.
Published: (2024)
Object Recognition in Human Computer Interaction:- A Comparative Analysis
by: Ranade, Kaushik, et al.
Published: (2024)
by: Ranade, Kaushik, et al.
Published: (2024)
Similar Items
-
Zero-shot Emotion Annotation in Facial Images Using Large Multimodal Models: Benchmarking and Prospects for Multi-Class, Multi-Frame Approaches
by: Zhang, He, et al.
Published: (2025) -
Generative AI Framework for 3D Object Generation in Augmented Reality
by: Behravan, Majid
Published: (2025) -
Human-in-the-Loop Annotation for Image-Based Engagement Estimation: Assessing the Impact of Model Reliability on Annotation Accuracy
by: Subramanya, Sahana Yadnakudige, et al.
Published: (2025) -
K-Sort Arena: Efficient and Reliable Benchmarking for Generative Models via K-wise Human Preferences
by: Li, Zhikai, et al.
Published: (2024) -
Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
by: Shen, Junxiao, et al.
Published: (2023)