On Train-Test Class Overlap and Detection for Image Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Chull Hwan, Yoon, Jooyoung, Hwang, Taebaek, Choi, Shunghyun, Gu, Yeong Hyeon, Avrithis, Yannis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining
von: Song, Chull Hwan, et al.
Veröffentlicht: (2024)
von: Song, Chull Hwan, et al.
Veröffentlicht: (2024)
Composed Image Retrieval for Training-Free Domain Conversion
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024)
Empirical Analysis of Anomaly Detection on Hyperspectral Imaging Using Dimension Reduction Methods
von: Kim, Dongeon, et al.
Veröffentlicht: (2024)
von: Kim, Dongeon, et al.
Veröffentlicht: (2024)
Do You Keep an Eye on What I Ask? Mitigating Multimodal Hallucination via Attention-Guided Ensemble Decoding
von: Cho, Yeongjae, et al.
Veröffentlicht: (2025)
von: Cho, Yeongjae, et al.
Veröffentlicht: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025)
von: Park, Sangha, et al.
Veröffentlicht: (2025)
Composed Image Retrieval for Remote Sensing
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
Multi-Target Unsupervised Domain Adaptation for Semantic Segmentation without External Data
von: Xu, Yonghao, et al.
Veröffentlicht: (2024)
von: Xu, Yonghao, et al.
Veröffentlicht: (2024)
CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds
von: Kim, Keonwoo, et al.
Veröffentlicht: (2025)
von: Kim, Keonwoo, et al.
Veröffentlicht: (2025)
Instance-Level Composed Image Retrieval
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
von: Psomas, Bill, et al.
Veröffentlicht: (2025)
Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
von: Shin, Chaehun, et al.
Veröffentlicht: (2025)
von: Shin, Chaehun, et al.
Veröffentlicht: (2025)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
Benchmarking Composed Image Retrieval for Applied Earth Observation
von: Psomas, Bill, et al.
Veröffentlicht: (2026)
von: Psomas, Bill, et al.
Veröffentlicht: (2026)
Anomaly Detection by Effectively Leveraging Synthetic Images
von: Kang, Sungho, et al.
Veröffentlicht: (2025)
von: Kang, Sungho, et al.
Veröffentlicht: (2025)
Is ImageNet worth 1 video? Learning strong image encoders from 1 long unlabelled video
von: Venkataramanan, Shashanka, et al.
Veröffentlicht: (2023)
von: Venkataramanan, Shashanka, et al.
Veröffentlicht: (2023)
Bidirectional Multimodal Prompt Learning with Scale-Aware Training for Few-Shot Multi-Class Anomaly Detection
von: Lee, Yujin, et al.
Veröffentlicht: (2024)
von: Lee, Yujin, et al.
Veröffentlicht: (2024)
Opti-CAM: Optimizing saliency maps for interpretability
von: Zhang, Hanwei, et al.
Veröffentlicht: (2023)
von: Zhang, Hanwei, et al.
Veröffentlicht: (2023)
CA-Stream: Attention-based pooling for interpretable image recognition
von: Torres, Felipe, et al.
Veröffentlicht: (2024)
von: Torres, Felipe, et al.
Veröffentlicht: (2024)
Improving Diffusion-Based Generative Models via Approximated Optimal Transport
von: Kim, Daegyu, et al.
Veröffentlicht: (2024)
von: Kim, Daegyu, et al.
Veröffentlicht: (2024)
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
GradMix: Gradient-based Selective Mixup for Robust Data Augmentation in Class-Incremental Learning
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
Contrastive Language Prompting to Ease False Positives in Medical Anomaly Detection
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
Feature Attenuation of Defective Representation Can Resolve Incomplete Masking on Anomaly Detection
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
Improving Geometry in Sparse-View 3DGS via Reprojection-based DoF Separation
von: Kim, Yongsung, et al.
Veröffentlicht: (2024)
von: Kim, Yongsung, et al.
Veröffentlicht: (2024)
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation
von: Hwang, Injoon, et al.
Veröffentlicht: (2024)
von: Hwang, Injoon, et al.
Veröffentlicht: (2024)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning
von: Hwang, Chan Yeong, et al.
Veröffentlicht: (2026)
von: Hwang, Chan Yeong, et al.
Veröffentlicht: (2026)
Class-balanced Open-set Semi-supervised Object Detection for Medical Images
von: Lu, Zhanyun, et al.
Veröffentlicht: (2024)
von: Lu, Zhanyun, et al.
Veröffentlicht: (2024)
DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
Training Class-Imbalanced Diffusion Model Via Overlap Optimization
von: Yan, Divin, et al.
Veröffentlicht: (2024)
von: Yan, Divin, et al.
Veröffentlicht: (2024)
Versatile Incremental Learning: Towards Class and Domain-Agnostic Incremental Learning
von: Park, Min-Yeong, et al.
Veröffentlicht: (2024)
von: Park, Min-Yeong, et al.
Veröffentlicht: (2024)
A Learning Paradigm for Interpretable Gradients
von: Figueroa, Felipe Torres, et al.
Veröffentlicht: (2024)
von: Figueroa, Felipe Torres, et al.
Veröffentlicht: (2024)
Synchronizing Task Behavior: Aligning Multiple Tasks during Test-Time Training
von: Jeong, Wooseong, et al.
Veröffentlicht: (2025)
von: Jeong, Wooseong, et al.
Veröffentlicht: (2025)
Disentangled Motion Modeling for Video Frame Interpolation
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts
von: Hwang, Taebaek, et al.
Veröffentlicht: (2025)
von: Hwang, Taebaek, et al.
Veröffentlicht: (2025)
Correcting Class Imbalances with Self-Training for Improved Universal Lesion Detection and Tagging
von: Shieh, Alexander, et al.
Veröffentlicht: (2025)
von: Shieh, Alexander, et al.
Veröffentlicht: (2025)
DragText: Rethinking Text Embedding in Point-based Image Editing
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
DINOv3 with Test-Time Training for Medical Image Registration
von: Wang, Shansong, et al.
Veröffentlicht: (2025)
von: Wang, Shansong, et al.
Veröffentlicht: (2025)
VizECGNet: Visual ECG Image Network for Cardiovascular Diseases Classification with Multi-Modal Training and Knowledge Distillation
von: Nam, Ju-Hyeon, et al.
Veröffentlicht: (2024)
von: Nam, Ju-Hyeon, et al.
Veröffentlicht: (2024)
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
von: Lee, Jaeseong, et al.
Veröffentlicht: (2025)
von: Lee, Jaeseong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining
von: Song, Chull Hwan, et al.
Veröffentlicht: (2024) -
Composed Image Retrieval for Training-Free Domain Conversion
von: Efthymiadis, Nikos, et al.
Veröffentlicht: (2024) -
Empirical Analysis of Anomaly Detection on Hyperspectral Imaging Using Dimension Reduction Methods
von: Kim, Dongeon, et al.
Veröffentlicht: (2024) -
Do You Keep an Eye on What I Ask? Mitigating Multimodal Hallucination via Attention-Guided Ensemble Decoding
von: Cho, Yeongjae, et al.
Veröffentlicht: (2025) -
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025)