LOCORE: Image Re-ranking with Long-Context Sequence Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiao, Zilin, Suma, Pavel, Sachdeva, Ayush, Wang, Hao-Jen, Kordopatis-Zilos, Giorgos, Tolias, Giorgos, Ordonez, Vicente |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AMES: Asymmetric and Memory-Efficient Similarity Estimation for Instance-level Retrieval
von: Suma, Pavel, et al.
Veröffentlicht: (2024)
von: Suma, Pavel, et al.
Veröffentlicht: (2024)
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
von: Suma, Pavel, et al.
Veröffentlicht: (2026)
von: Suma, Pavel, et al.
Veröffentlicht: (2026)
Indexing Multimodal Language Models for Large-scale Image Retrieval
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
Instance-Level Generation for Representation Learning
von: Wu, Yankun, et al.
Veröffentlicht: (2025)
von: Wu, Yankun, et al.
Veröffentlicht: (2025)
Fusion Transformer with Object Mask Guidance for Image Forgery Analysis
von: Karageorgiou, Dimitrios, et al.
Veröffentlicht: (2024)
von: Karageorgiou, Dimitrios, et al.
Veröffentlicht: (2024)
Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
von: Ramos, Ryan, et al.
Veröffentlicht: (2025)
von: Ramos, Ryan, et al.
Veröffentlicht: (2025)
ILIAS: Instance-Level Image retrieval At Scale
von: Kordopatis-Zilos, Giorgos, et al.
Veröffentlicht: (2025)
von: Kordopatis-Zilos, Giorgos, et al.
Veröffentlicht: (2025)
InDistill: Information flow-preserving knowledge distillation for model compression
von: Sarridis, Ioannis, et al.
Veröffentlicht: (2022)
von: Sarridis, Ioannis, et al.
Veröffentlicht: (2022)
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
Label Propagation for Zero-shot Classification with Vision-Language Models
von: Stojnić, Vladan, et al.
Veröffentlicht: (2024)
von: Stojnić, Vladan, et al.
Veröffentlicht: (2024)
Three Things to Know about Deep Metric Learning
von: Patel, Yash, et al.
Veröffentlicht: (2024)
von: Patel, Yash, et al.
Veröffentlicht: (2024)
Category-level Text-to-Image Retrieval Improved: Bridging the Domain Gap with Diffusion Models and Vision Encoders
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2025)
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2025)
SPAR: Single-Pass Any-Resolution ViT for Open-vocabulary Segmentation
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
Global-Aware Edge Prioritization for Pose Graph Initialization
von: Wei, Tong, et al.
Veröffentlicht: (2026)
von: Wei, Tong, et al.
Veröffentlicht: (2026)
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
von: Stojnić, Vladan, et al.
Veröffentlicht: (2025)
von: Stojnić, Vladan, et al.
Veröffentlicht: (2025)
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2026)
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2026)
Composed Image Retrieval for Remote Sensing
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
Composed Image Retrieval for Training-Free Domain Conversion
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
Instance-Level Composed Image Retrieval
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
Prompt2Fashion: An automatically generated fashion dataset
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
Counterfactual Edits for Generative Evaluation
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
von: Papadimitriou, Christos, et al.
Veröffentlicht: (2024)
von: Papadimitriou, Christos, et al.
Veröffentlicht: (2024)
ENACT: Entropy-based Clustering of Attention Input for Reducing the Computational Needs of Object Detection Transformers
von: Savathrakis, Giorgos, et al.
Veröffentlicht: (2024)
von: Savathrakis, Giorgos, et al.
Veröffentlicht: (2024)
Benchmarking Composed Image Retrieval for Applied Earth Observation
von: Psomas, Bill, et al.
Veröffentlicht: (2026)
von: Psomas, Bill, et al.
Veröffentlicht: (2026)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
von: Petsangourakis, Giorgos, et al.
Veröffentlicht: (2025)
von: Petsangourakis, Giorgos, et al.
Veröffentlicht: (2025)
Fine-Grained ImageNet Classification in the Wild
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
Structure Your Data: Towards Semantic Graph Counterfactuals
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2024)
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2024)
U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2026)
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2026)
Through the PRISm: Importance-Aware Scene Graphs for Image Retrieval
von: Georgoulopoulos, Dimitrios, et al.
Veröffentlicht: (2025)
von: Georgoulopoulos, Dimitrios, et al.
Veröffentlicht: (2025)
Context Matters: Query-aware Dynamic Long Sequence Modeling of Gigapixel Images
von: Guo, Zhengrui, et al.
Veröffentlicht: (2025)
von: Guo, Zhengrui, et al.
Veröffentlicht: (2025)
ARPA: A Novel Hybrid Model for Advancing Visual Word Disambiguation Using Large Language Models and Transformers
von: Papastavrou, Aristi, et al.
Veröffentlicht: (2024)
von: Papastavrou, Aristi, et al.
Veröffentlicht: (2024)
Distilling Vision Transformers for Distortion-Robust Representation Learning
von: Alexis, Konstantinos, et al.
Veröffentlicht: (2026)
von: Alexis, Konstantinos, et al.
Veröffentlicht: (2026)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
The Contribution of Knowledge in Visiolinguistic Learning: A Survey on Tasks and Challenges
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
Attention, Please! Revisiting Attentive Probing Through the Lens of Efficiency
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
SeLoRA: Self-Expanding Low-Rank Adaptation of Latent Diffusion Model for Medical Image Synthesis
von: Mao, Yuchen, et al.
Veröffentlicht: (2024)
von: Mao, Yuchen, et al.
Veröffentlicht: (2024)
A Dataset for Semantic Segmentation in the Presence of Unknowns
von: Laskar, Zakaria, et al.
Veröffentlicht: (2025)
von: Laskar, Zakaria, et al.
Veröffentlicht: (2025)
Grounding Language Models for Visual Entity Recognition
von: Xiao, Zilin, et al.
Veröffentlicht: (2024)
von: Xiao, Zilin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AMES: Asymmetric and Memory-Efficient Similarity Estimation for Instance-level Retrieval
von: Suma, Pavel, et al.
Veröffentlicht: (2024) -
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
von: Suma, Pavel, et al.
Veröffentlicht: (2026) -
Indexing Multimodal Language Models for Large-scale Image Retrieval
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026) -
Instance-Level Generation for Representation Learning
von: Wu, Yankun, et al.
Veröffentlicht: (2025) -
Fusion Transformer with Object Mask Guidance for Image Forgery Analysis
von: Karageorgiou, Dimitrios, et al.
Veröffentlicht: (2024)