Leveraging Lightweight Entity Extraction for Scalable Event-Based Image Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Minh, Dao Sy Duy, Kiet, Huynh Trung, Quy, Nguyen Lam Phu, Pham, Phu-Hoa, Nguyen, Tran Chi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
Navigating Simply, Aligning Deeply: Winning Solutions for Mouse vs. AI 2025
by: Pham, Phu-Hoa, et al.
Published: (2026)
by: Pham, Phu-Hoa, et al.
Published: (2026)
EvoQRE: Modeling Bounded Rationality in Safety-Critical Traffic Simulation via Evolutionary Quantal Response Equilibrium
by: Pham, Phu-Hoa, et al.
Published: (2026)
by: Pham, Phu-Hoa, et al.
Published: (2026)
MEMRES: A Memory-Augmented Resolver with Confidence Cascade for Agentic Python Dependency Resolution
by: Minh, Dao Sy Duy, et al.
Published: (2026)
by: Minh, Dao Sy Duy, et al.
Published: (2026)
Unlocking Compositional Generalization in Continual Few-Shot Learning
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
MIST: Reliable Streaming Decision Trees for Online Class-Incremental Learning via McDiarmid Bound
by: Pham, Phu-Hoa, et al.
Published: (2026)
by: Pham, Phu-Hoa, et al.
Published: (2026)
Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery
by: Tran, Chi-Nguyen, et al.
Published: (2026)
by: Tran, Chi-Nguyen, et al.
Published: (2026)
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement
by: Kiet, Huynh Trung, et al.
Published: (2026)
by: Kiet, Huynh Trung, et al.
Published: (2026)
Early-Stage Prediction of Review Effort in AI-Generated Pull Requests
by: Minh, Dao Sy Duy, et al.
Published: (2026)
by: Minh, Dao Sy Duy, et al.
Published: (2026)
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
by: Zhang, Bingqing, et al.
Published: (2026)
by: Zhang, Bingqing, et al.
Published: (2026)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
by: Hu, Ruofan, et al.
Published: (2026)
by: Hu, Ruofan, et al.
Published: (2026)
Quantifying and Narrowing the Unknown: Interactive Text-to-Video Retrieval via Uncertainty Minimization
by: Zhang, Bingqing, et al.
Published: (2025)
by: Zhang, Bingqing, et al.
Published: (2025)
Loom: Hybrid Retrieval-Scoring Outfit Recommendation with Semantic Material Compatibility and Occasion-Aware Embedding Priors
by: Berlia, Anushree
Published: (2026)
by: Berlia, Anushree
Published: (2026)
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation
by: Kiet, Huynh Trung, et al.
Published: (2026)
by: Kiet, Huynh Trung, et al.
Published: (2026)
CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA
by: Kovalev, Vsevolod, et al.
Published: (2025)
by: Kovalev, Vsevolod, et al.
Published: (2025)
A Scalable Pipeline Combining Procedural 3D Graphics and Guided Diffusion for Photorealistic Synthetic Training Data Generation in White Button Mushroom Segmentation
by: Károly, Artúr I., et al.
Published: (2025)
by: Károly, Artúr I., et al.
Published: (2025)
Spatially-Grounded Document Retrieval via Patch-to-Region Relevance Propagation
by: Georgiou, Athos
Published: (2025)
by: Georgiou, Athos
Published: (2025)
Leveraging OpenFlamingo for Multimodal Embedding Analysis of C2C Car Parts Data
by: Rashid, Maisha Binte, et al.
Published: (2025)
by: Rashid, Maisha Binte, et al.
Published: (2025)
Meta-Monomorphizing Specializations
by: Bruzzone, Federico, et al.
Published: (2026)
by: Bruzzone, Federico, et al.
Published: (2026)
Robust Test-time Video-Text Retrieval: Benchmarking and Adapting for Query Shifts
by: Zhang, Bingqing, et al.
Published: (2026)
by: Zhang, Bingqing, et al.
Published: (2026)
Experimenting active and sequential learning in a medieval music manuscript
by: Sharma, Sachin, et al.
Published: (2025)
by: Sharma, Sachin, et al.
Published: (2025)
Large Language Model for Qualitative Research -- A Systematic Mapping Study
by: Barros, Cauã Ferreira, et al.
Published: (2024)
by: Barros, Cauã Ferreira, et al.
Published: (2024)
GeoSearch: Augmenting Worldwide Geolocalization with Web-Scale Reverse Image Search and Image Matching
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
Expertized Caption Auto-Enhancement for Video-Text Retrieval
by: Yang, Baoyao, et al.
Published: (2025)
by: Yang, Baoyao, et al.
Published: (2025)
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
by: Nguyen, Andy, et al.
Published: (2026)
by: Nguyen, Andy, et al.
Published: (2026)
Clinical Knowledge Graph Construction and Evaluation with Multi-LLMs via Retrieval-Augmented Generation
by: Das, Udiptaman, et al.
Published: (2026)
by: Das, Udiptaman, et al.
Published: (2026)
Integrating Corpus Analysis and ChatGPT in Teaching English Collocations: A Hybrid Approach
by: Quy Huynh Phu Pham
Published: (2025)
by: Quy Huynh Phu Pham
Published: (2025)
Structured Interfaces for Automated Reasoning with 3D Scene Graphs
by: Ray, Aaron, et al.
Published: (2025)
by: Ray, Aaron, et al.
Published: (2025)
Hierarchical Multi-Positive Contrastive Learning for Patent Image Retrieval
by: Kavimandan, Kshitij, et al.
Published: (2025)
by: Kavimandan, Kshitij, et al.
Published: (2025)
Visualization Retrieval for Data Literacy: Position Paper
by: Nguyen, Huyen N., et al.
Published: (2026)
by: Nguyen, Huyen N., et al.
Published: (2026)
AI vs. Human Moderators: A Comparative Evaluation of Multimodal LLMs in Content Moderation for Brand Safety
by: Levi, Adi, et al.
Published: (2025)
by: Levi, Adi, et al.
Published: (2025)
Aximorphic Perspective Projection Model for Immersive Imagery
by: Fober, Jakub Maksymilian
Published: (2021)
by: Fober, Jakub Maksymilian
Published: (2021)
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
by: Ghosh, Archishman, et al.
Published: (2026)
by: Ghosh, Archishman, et al.
Published: (2026)
Understanding LLM Agent Behaviours via Game Theory: Strategy Recognition, Biases and Multi-Agent Dynamics
by: Huynh, Trung-Kiet, et al.
Published: (2025)
by: Huynh, Trung-Kiet, et al.
Published: (2025)
More at Stake: How Payoff and Language Shape LLM Agent Strategies in Cooperation Dilemmas
by: Huynh, Trung-Kiet, et al.
Published: (2026)
by: Huynh, Trung-Kiet, et al.
Published: (2026)
A Grounded Memory System For Smart Personal Assistants
by: Ocker, Felix, et al.
Published: (2025)
by: Ocker, Felix, et al.
Published: (2025)
Sycamore: Characterizing Synthetic Personas for Evaluating Genomics Visualization Retrieval
by: Nguyen, Huyen N., et al.
Published: (2026)
by: Nguyen, Huyen N., et al.
Published: (2026)
EntGPT: Entity Linking with Generative Large Language Models
by: Ding, Yifan, et al.
Published: (2024)
by: Ding, Yifan, et al.
Published: (2024)
Similar Items
-
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
by: Quy, Nguyen Lam Phu, et al.
Published: (2025) -
Navigating Simply, Aligning Deeply: Winning Solutions for Mouse vs. AI 2025
by: Pham, Phu-Hoa, et al.
Published: (2026) -
EvoQRE: Modeling Bounded Rationality in Safety-Critical Traffic Simulation via Evolutionary Quantal Response Equilibrium
by: Pham, Phu-Hoa, et al.
Published: (2026) -
MEMRES: A Memory-Augmented Resolver with Confidence Cascade for Agentic Python Dependency Resolution
by: Minh, Dao Sy Duy, et al.
Published: (2026) -
Unlocking Compositional Generalization in Continual Few-Shot Learning
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)