Instance-Level Generation for Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yankun, Laskar, Zakaria, Kordopatis-Zilos, Giorgos, Garcia, Noa, Tolias, Giorgos |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AMES: Asymmetric and Memory-Efficient Similarity Estimation for Instance-level Retrieval
by: Suma, Pavel, et al.
Published: (2024)
by: Suma, Pavel, et al.
Published: (2024)
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
by: Suma, Pavel, et al.
Published: (2026)
by: Suma, Pavel, et al.
Published: (2026)
ILIAS: Instance-Level Image retrieval At Scale
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
by: Ramos, Ryan, et al.
Published: (2025)
by: Ramos, Ryan, et al.
Published: (2025)
Indexing Multimodal Language Models for Large-scale Image Retrieval
by: Tharwat, Bahey, et al.
Published: (2026)
by: Tharwat, Bahey, et al.
Published: (2026)
Fusion Transformer with Object Mask Guidance for Image Forgery Analysis
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
LOCORE: Image Re-ranking with Long-Context Sequence Modeling
by: Xiao, Zilin, et al.
Published: (2025)
by: Xiao, Zilin, et al.
Published: (2025)
Category-level Text-to-Image Retrieval Improved: Bridging the Domain Gap with Diffusion Models and Vision Encoders
by: Khan, Faizan Farooq, et al.
Published: (2025)
by: Khan, Faizan Farooq, et al.
Published: (2025)
InDistill: Information flow-preserving knowledge distillation for model compression
by: Sarridis, Ioannis, et al.
Published: (2022)
by: Sarridis, Ioannis, et al.
Published: (2022)
Composed Image Retrieval for Training-Free Domain Conversion
by: Efthymiadis, Nikos, et al.
Published: (2024)
by: Efthymiadis, Nikos, et al.
Published: (2024)
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
by: Efthymiadis, Nikos, et al.
Published: (2024)
by: Efthymiadis, Nikos, et al.
Published: (2024)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
Three Things to Know about Deep Metric Learning
by: Patel, Yash, et al.
Published: (2024)
by: Patel, Yash, et al.
Published: (2024)
Label Propagation for Zero-shot Classification with Vision-Language Models
by: Stojnić, Vladan, et al.
Published: (2024)
by: Stojnić, Vladan, et al.
Published: (2024)
A Dataset for Semantic Segmentation in the Presence of Unknowns
by: Laskar, Zakaria, et al.
Published: (2025)
by: Laskar, Zakaria, et al.
Published: (2025)
SPAR: Single-Pass Any-Resolution ViT for Open-vocabulary Segmentation
by: Kombol, Naomi, et al.
Published: (2026)
by: Kombol, Naomi, et al.
Published: (2026)
Global-Aware Edge Prioritization for Pose Graph Initialization
by: Wei, Tong, et al.
Published: (2026)
by: Wei, Tong, et al.
Published: (2026)
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models
by: Argyrou, Georgia, et al.
Published: (2024)
by: Argyrou, Georgia, et al.
Published: (2024)
Distilling Vision Transformers for Distortion-Robust Representation Learning
by: Alexis, Konstantinos, et al.
Published: (2026)
by: Alexis, Konstantinos, et al.
Published: (2026)
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
by: Stojnić, Vladan, et al.
Published: (2025)
by: Stojnić, Vladan, et al.
Published: (2025)
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
by: Aravanis, Tilemachos, et al.
Published: (2026)
by: Aravanis, Tilemachos, et al.
Published: (2026)
Stable Diffusion Exposed: Gender Bias from Prompt to Image
by: Wu, Yankun, et al.
Published: (2023)
by: Wu, Yankun, et al.
Published: (2023)
Counterfactual Edits for Generative Evaluation
by: Lymperaiou, Maria, et al.
Published: (2023)
by: Lymperaiou, Maria, et al.
Published: (2023)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
by: Papadimitriou, Christos, et al.
Published: (2024)
by: Papadimitriou, Christos, et al.
Published: (2024)
Prompt2Fashion: An automatically generated fashion dataset
by: Argyrou, Georgia, et al.
Published: (2024)
by: Argyrou, Georgia, et al.
Published: (2024)
ENACT: Entropy-based Clustering of Attention Input for Reducing the Computational Needs of Object Detection Transformers
by: Savathrakis, Giorgos, et al.
Published: (2024)
by: Savathrakis, Giorgos, et al.
Published: (2024)
Composed Image Retrieval for Remote Sensing
by: Psomas, Bill, et al.
Published: (2024)
by: Psomas, Bill, et al.
Published: (2024)
The Contribution of Knowledge in Visiolinguistic Learning: A Survey on Tasks and Challenges
by: Lymperaiou, Maria, et al.
Published: (2023)
by: Lymperaiou, Maria, et al.
Published: (2023)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
by: Petsangourakis, Giorgos, et al.
Published: (2025)
by: Petsangourakis, Giorgos, et al.
Published: (2025)
Privacy in Image Datasets: A Case Study on Pregnancy Ultrasounds
by: Lohanimit, Rawisara, et al.
Published: (2026)
by: Lohanimit, Rawisara, et al.
Published: (2026)
HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data
by: Girtsou, Stella, et al.
Published: (2026)
by: Girtsou, Stella, et al.
Published: (2026)
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
by: Lymperaiou, Maria, et al.
Published: (2025)
by: Lymperaiou, Maria, et al.
Published: (2025)
Structure Your Data: Towards Semantic Graph Counterfactuals
by: Dimitriou, Angeliki, et al.
Published: (2024)
by: Dimitriou, Angeliki, et al.
Published: (2024)
U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations
by: Dimitriou, Angeliki, et al.
Published: (2026)
by: Dimitriou, Angeliki, et al.
Published: (2026)
Rethinking HTG Evaluation: Bridging Generation and Recognition
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
DiffusionPen: Towards Controlling the Style of Handwritten Text Generation
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
by: Spanos, Nikolaos, et al.
Published: (2025)
by: Spanos, Nikolaos, et al.
Published: (2025)
Attention, Please! Revisiting Attentive Probing Through the Lens of Efficiency
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
Benchmarking Composed Image Retrieval for Applied Earth Observation
by: Psomas, Bill, et al.
Published: (2026)
by: Psomas, Bill, et al.
Published: (2026)
ARPA: A Novel Hybrid Model for Advancing Visual Word Disambiguation Using Large Language Models and Transformers
by: Papastavrou, Aristi, et al.
Published: (2024)
by: Papastavrou, Aristi, et al.
Published: (2024)
Similar Items
-
AMES: Asymmetric and Memory-Efficient Similarity Estimation for Instance-level Retrieval
by: Suma, Pavel, et al.
Published: (2024) -
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
by: Suma, Pavel, et al.
Published: (2026) -
ILIAS: Instance-Level Image retrieval At Scale
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025) -
Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
by: Ramos, Ryan, et al.
Published: (2025) -
Indexing Multimodal Language Models for Large-scale Image Retrieval
by: Tharwat, Bahey, et al.
Published: (2026)