Dynamic Sub-region Search in Homogeneous Collections Using CLIP
Fuente:
arXiv
Saved in:
| Main Authors: | Jäckl, Bastian, Kloda, Vojtěch, Keim, Daniel A., Lokoč, Jakub |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Experimental Evaluation of Static Image Sub-Region-Based Search Models Using CLIP
by: Jäckl, Bastian, et al.
Published: (2025)
by: Jäckl, Bastian, et al.
Published: (2025)
Evaluating Keyframe Layouts for Visual Known-Item Search in Homogeneous Collections
by: Jäckl, Bastian, et al.
Published: (2025)
by: Jäckl, Bastian, et al.
Published: (2025)
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
Visualizing the Evolution of Twitter (X.com) Conversations: A Comprehensive Methodology Applied to AI Training Discussions on ChatGPT
by: Jess, Nicole, et al.
Published: (2024)
by: Jess, Nicole, et al.
Published: (2024)
Improving Graph Embeddings in Machine Learning Using Knowledge Completion with Validation in a Case Study on COVID-19 Spread
by: Napoli, Rosario, et al.
Published: (2025)
by: Napoli, Rosario, et al.
Published: (2025)
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
by: Gondhalekar, Chinmay, et al.
Published: (2025)
by: Gondhalekar, Chinmay, et al.
Published: (2025)
Hierarchical Multi-Positive Contrastive Learning for Patent Image Retrieval
by: Kavimandan, Kshitij, et al.
Published: (2025)
by: Kavimandan, Kshitij, et al.
Published: (2025)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
A Hybrid Framework for Real-Time Data Drift and Anomaly Identification Using Hierarchical Temporal Memory and Statistical Tests
by: Bandyopadhyay, Subhadip, et al.
Published: (2025)
by: Bandyopadhyay, Subhadip, et al.
Published: (2025)
Method for Aggregating Unstructured Data Using Large Language Models
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors
by: Zhang, Junbin, et al.
Published: (2026)
by: Zhang, Junbin, et al.
Published: (2026)
MATEX: Multi-scale Attention and Text-guided Explainability of Medical Vision-Language Models
by: Imran, Muhammad, et al.
Published: (2026)
by: Imran, Muhammad, et al.
Published: (2026)
Detecting Legend Items on Historical Maps Using GPT-4o with In-Context Learning
by: Kirsanova, Sofia, et al.
Published: (2025)
by: Kirsanova, Sofia, et al.
Published: (2025)
MultiFinRAG: An Optimized Multimodal Retrieval-Augmented Generation (RAG) Framework for Financial Question Answering
by: Gondhalekar, Chinmay, et al.
Published: (2025)
by: Gondhalekar, Chinmay, et al.
Published: (2025)
Dense Video Understanding with Gated Residual Tokenization
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
Task Memory Engine (TME): Enhancing State Awareness for Multi-Step LLM Agent Tasks
by: Ye, Ye
Published: (2025)
by: Ye, Ye
Published: (2025)
CSTS: A Benchmark for the Discovery of Correlation Structures in Time Series Clustering
by: Degen, Isabella, et al.
Published: (2025)
by: Degen, Isabella, et al.
Published: (2025)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
by: Patel, Urjitkumar, et al.
Published: (2025)
by: Patel, Urjitkumar, et al.
Published: (2025)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
by: Bui, Quoc-Anh, et al.
Published: (2025)
by: Bui, Quoc-Anh, et al.
Published: (2025)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
by: Bui, Quoc-Anh, et al.
Published: (2025)
by: Bui, Quoc-Anh, et al.
Published: (2025)
Software architecture and manual for novel versatile CT image analysis toolbox -- AnatomyArchive
by: Xu, Lei, et al.
Published: (2025)
by: Xu, Lei, et al.
Published: (2025)
Deep Learning Frameworks for Cognitive Radio Networks: Review and Open Research Challenges
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
Unified Interaction Foundational Model (UIFM) for Predicting Complex User and System Behavior
by: Ethiraj, Vignesh, et al.
Published: (2025)
by: Ethiraj, Vignesh, et al.
Published: (2025)
Efficient and Privacy-Protecting Background Removal for 2D Video Streaming using iPhone 15 Pro Max LiDAR
by: Kinnevan, Jessica, et al.
Published: (2025)
by: Kinnevan, Jessica, et al.
Published: (2025)
Proximity Measure of Information Object Features for Solving the Problem of Their Identification in Information Systems
by: Yuzefovych, Volodymyr
Published: (2026)
by: Yuzefovych, Volodymyr
Published: (2026)
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
by: Yang, Zhenye, et al.
Published: (2025)
by: Yang, Zhenye, et al.
Published: (2025)
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026)
by: Du, Yuchen, et al.
Published: (2026)
Experimenting active and sequential learning in a medieval music manuscript
by: Sharma, Sachin, et al.
Published: (2025)
by: Sharma, Sachin, et al.
Published: (2025)
LUCAS-MEGA: A Large-Scale Multimodal Dataset for Representation Learning in Soil-Environment Systems
by: Leng, Kuangdai, et al.
Published: (2026)
by: Leng, Kuangdai, et al.
Published: (2026)
Quantifying and Narrowing the Unknown: Interactive Text-to-Video Retrieval via Uncertainty Minimization
by: Zhang, Bingqing, et al.
Published: (2025)
by: Zhang, Bingqing, et al.
Published: (2025)
GLEAN: Grounded Lightweight Evaluation Anchors for Contamination-Aware Tabular Reasoning
by: Wang, Qizhi
Published: (2026)
by: Wang, Qizhi
Published: (2026)
Bi-View Embedding Fusion: A Hybrid Learning Approach for Knowledge Graph's Nodes Classification Addressing Problems with Limited Data
by: Napoli, Rosario, et al.
Published: (2025)
by: Napoli, Rosario, et al.
Published: (2025)
AIDOVECL: AI-generated Dataset of Outpainted Vehicles for Eye-level Classification and Localization
by: Kazemi, Amir, et al.
Published: (2024)
by: Kazemi, Amir, et al.
Published: (2024)
Training-free Zero-shot Composed Image Retrieval via Weighted Modality Fusion and Similarity
by: Wu, Ren-Di, et al.
Published: (2024)
by: Wu, Ren-Di, et al.
Published: (2024)
Evaluating the Effect of Compression on Video Temporal Consistency Using Objective Quality Metrics
by: Zsoldos, Peter
Published: (2026)
by: Zsoldos, Peter
Published: (2026)
TexTile: A Differentiable Metric for Texture Tileability
by: Rodriguez-Pardo, Carlos, et al.
Published: (2024)
by: Rodriguez-Pardo, Carlos, et al.
Published: (2024)
GeoJEPA: Towards Eliminating Augmentation- and Sampling Bias in Multimodal Geospatial Learning
by: Lundqvist, Theodor, et al.
Published: (2025)
by: Lundqvist, Theodor, et al.
Published: (2025)
Personalised Travel Recommendation based on Location Co-occurrence
by: Clements, Maarten, et al.
Published: (2011)
by: Clements, Maarten, et al.
Published: (2011)
Similar Items
-
Experimental Evaluation of Static Image Sub-Region-Based Search Models Using CLIP
by: Jäckl, Bastian, et al.
Published: (2025) -
Evaluating Keyframe Layouts for Visual Known-Item Search in Homogeneous Collections
by: Jäckl, Bastian, et al.
Published: (2025) -
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
by: Zhang, Haichao, et al.
Published: (2025) -
Visualizing the Evolution of Twitter (X.com) Conversations: A Comprehensive Methodology Applied to AI Training Discussions on ChatGPT
by: Jess, Nicole, et al.
Published: (2024) -
Improving Graph Embeddings in Machine Learning Using Knowledge Completion with Validation in a Case Study on COVID-19 Spread
by: Napoli, Rosario, et al.
Published: (2025)