Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
Fuente:
arXiv
Guardado en:
| Autores principales: | Zaher, Kawtar, Buisson, Olivier, Joly, Alexis |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Positive-First Most Ambiguous: A Simple Active Learning Criterion for Interactive Retrieval of Rare Categories
por: Zaher, Kawtar, et al.
Publicado: (2026)
por: Zaher, Kawtar, et al.
Publicado: (2026)
Self-Supervised Learning as Discrete Communication
por: Zaher, Kawtar, et al.
Publicado: (2026)
por: Zaher, Kawtar, et al.
Publicado: (2026)
Image Hashing via Cross-View Code Alignment in the Age of Foundation Models
por: Moummad, Ilyass, et al.
Publicado: (2025)
por: Moummad, Ilyass, et al.
Publicado: (2025)
ERPA: Efficient RPA Model Integrating OCR and LLMs for Intelligent Document Processing
por: Abdellaif, Osama, et al.
Publicado: (2024)
por: Abdellaif, Osama, et al.
Publicado: (2024)
Spacewalker: Traversing Representation Spaces for Fast Interactive Exploration and Annotation of Unstructured Data
por: Heine, Lukas, et al.
Publicado: (2024)
por: Heine, Lukas, et al.
Publicado: (2024)
CLAS: A Machine Learning Enhanced Framework for Exploring Large 3D Design Datasets
por: Zhang, XiuYu, et al.
Publicado: (2024)
por: Zhang, XiuYu, et al.
Publicado: (2024)
A Versatile Dataset of Mouse and Eye Movements on Search Engine Results Pages
por: Latifzadeh, Kayhan, et al.
Publicado: (2025)
por: Latifzadeh, Kayhan, et al.
Publicado: (2025)
Compact Hypercube Embeddings for Fast Text-based Wildlife Observation Retrieval
por: Moummad, Ilyass, et al.
Publicado: (2026)
por: Moummad, Ilyass, et al.
Publicado: (2026)
ArtCognition: A Multimodal AI Framework for Affective State Sensing from Visual and Kinematic Drawing Cues
por: Binaei-Haghighi, Behrad, et al.
Publicado: (2026)
por: Binaei-Haghighi, Behrad, et al.
Publicado: (2026)
Civiverse: A Dataset for Analyzing User Engagement with Open-Source Text-to-Image Models
por: Palmini, Maria-Teresa De Rosa, et al.
Publicado: (2024)
por: Palmini, Maria-Teresa De Rosa, et al.
Publicado: (2024)
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
por: Ye, Yilin, et al.
Publicado: (2024)
por: Ye, Yilin, et al.
Publicado: (2024)
Generative AI for Video Trailer Synthesis: From Extractive Heuristics to Autoregressive Creativity
por: Dharmaratnakar, Abhishek, et al.
Publicado: (2026)
por: Dharmaratnakar, Abhishek, et al.
Publicado: (2026)
Echoes in the Loop: Diagnosing Risks in LLM-Powered Recommender Systems under Feedback Loops
por: Park, Donguk, et al.
Publicado: (2026)
por: Park, Donguk, et al.
Publicado: (2026)
Evaluating VisualRAG: Quantifying Cross-Modal Performance in Enterprise Document Understanding
por: Mannam, Varun, et al.
Publicado: (2025)
por: Mannam, Varun, et al.
Publicado: (2025)
Chaining text-to-image and large language model: A novel approach for generating personalized e-commerce banners
por: Vashishtha, Shanu, et al.
Publicado: (2024)
por: Vashishtha, Shanu, et al.
Publicado: (2024)
Seeing Faces in Things: A Model and Dataset for Pareidolia
por: Hamilton, Mark, et al.
Publicado: (2024)
por: Hamilton, Mark, et al.
Publicado: (2024)
Human-Computer Interaction as a basis for assessing Geographic Information Retrieval Systems.
por: Manuel Enrique Puebla Martínez
Publicado: (2018)
por: Manuel Enrique Puebla Martínez
Publicado: (2018)
Emancipatory Information Retrieval
por: Mitra, Bhaskar
Publicado: (2025)
por: Mitra, Bhaskar
Publicado: (2025)
LoopLens: Supporting Search as Creation in Loop-Based Music Composition
por: Long, Sheng, et al.
Publicado: (2026)
por: Long, Sheng, et al.
Publicado: (2026)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
por: Wang, Yuxiao, et al.
Publicado: (2025)
por: Wang, Yuxiao, et al.
Publicado: (2025)
UCDR-Adapter: Exploring Adaptation of Pre-Trained Vision-Language Models for Universal Cross-Domain Retrieval
por: Jiang, Haoyu, et al.
Publicado: (2024)
por: Jiang, Haoyu, et al.
Publicado: (2024)
Open-Source Retrieval Augmented Generation Framework for Retrieving Accurate Medication Insights from Formularies for African Healthcare Workers
por: AI, Axum, et al.
Publicado: (2025)
por: AI, Axum, et al.
Publicado: (2025)
The Klarna Product Page Dataset: Web Element Nomination with Graph Neural Networks and Large Language Models
por: Hotti, Alexandra, et al.
Publicado: (2021)
por: Hotti, Alexandra, et al.
Publicado: (2021)
Can Instructed Retrieval Models Really Support Exploration?
por: Maheshwari, Piyush, et al.
Publicado: (2026)
por: Maheshwari, Piyush, et al.
Publicado: (2026)
OwlerLite: Scope- and Freshness-Aware Web Retrieval for LLM Assistants
por: Zerhoudi, Saber, et al.
Publicado: (2026)
por: Zerhoudi, Saber, et al.
Publicado: (2026)
From Query to Conscience: The Importance of Information Retrieval in Empowering Socially Responsible Consumerism
por: van der Sluis, Frans, et al.
Publicado: (2026)
por: van der Sluis, Frans, et al.
Publicado: (2026)
DataScout: Automatic Data Fact Retrieval for Statement Augmentation with an LLM-Based Agent
por: Chen, Chuer, et al.
Publicado: (2025)
por: Chen, Chuer, et al.
Publicado: (2025)
Cross-Format Retrieval-Augmented Generation in XR with LLMs for Context-Aware Maintenance Assistance
por: Nagy, Akos, et al.
Publicado: (2025)
por: Nagy, Akos, et al.
Publicado: (2025)
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices
por: Serajeh, Neda Taghizadeh, et al.
Publicado: (2024)
por: Serajeh, Neda Taghizadeh, et al.
Publicado: (2024)
Doc To The Future: Infomorphs for Interactive, Multimodal Document Transformation and Generation
por: Kumaravel, Balasaravanan Thoravi
Publicado: (2025)
por: Kumaravel, Balasaravanan Thoravi
Publicado: (2025)
PoultryTalk: A Multi-modal Retrieval-Augmented Generation (RAG) System for Intelligent Poultry Management and Decision Support
por: Khanal, Kapalik, et al.
Publicado: (2025)
por: Khanal, Kapalik, et al.
Publicado: (2025)
Computer Vision for Objects used in Group Work: Challenges and Opportunities
por: Jung, Changsoo, et al.
Publicado: (2025)
por: Jung, Changsoo, et al.
Publicado: (2025)
Co-Matching: Towards Human-Machine Collaborative Legal Case Matching
por: Huang, Chen, et al.
Publicado: (2024)
por: Huang, Chen, et al.
Publicado: (2024)
SymbioticRAG: Enhancing Document Intelligence Through Human-LLM Symbiotic Collaboration
por: Sun, Qiang, et al.
Publicado: (2025)
por: Sun, Qiang, et al.
Publicado: (2025)
Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications
por: Tavakoli, Leila, et al.
Publicado: (2025)
por: Tavakoli, Leila, et al.
Publicado: (2025)
Task Supportive and Personalized Human-Large Language Model Interaction: A User Study
por: Wang, Ben, et al.
Publicado: (2024)
por: Wang, Ben, et al.
Publicado: (2024)
"Can You Tell Me?": Designing Copilots to Support Human Judgement in Online Information Seeking
por: Bink, Markus, et al.
Publicado: (2026)
por: Bink, Markus, et al.
Publicado: (2026)
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
por: Pan, Bo, et al.
Publicado: (2026)
por: Pan, Bo, et al.
Publicado: (2026)
ViT-Explainer: An Interactive Walkthrough of the Vision Transformer Pipeline
por: Hernandez, Juan Manuel, et al.
Publicado: (2026)
por: Hernandez, Juan Manuel, et al.
Publicado: (2026)
VFA: Vision Frequency Analysis of Foundation Models and Human
por: Darvishi-Bayazi, Mohammad-Javad, et al.
Publicado: (2024)
por: Darvishi-Bayazi, Mohammad-Javad, et al.
Publicado: (2024)
Ejemplares similares
-
Positive-First Most Ambiguous: A Simple Active Learning Criterion for Interactive Retrieval of Rare Categories
por: Zaher, Kawtar, et al.
Publicado: (2026) -
Self-Supervised Learning as Discrete Communication
por: Zaher, Kawtar, et al.
Publicado: (2026) -
Image Hashing via Cross-View Code Alignment in the Age of Foundation Models
por: Moummad, Ilyass, et al.
Publicado: (2025) -
ERPA: Efficient RPA Model Integrating OCR and LLMs for Intelligent Document Processing
por: Abdellaif, Osama, et al.
Publicado: (2024) -
Spacewalker: Traversing Representation Spaces for Fast Interactive Exploration and Annotation of Unstructured Data
por: Heine, Lukas, et al.
Publicado: (2024)