Neighbor-aware Instance Refining with Noisy Labels for Cross-Modal Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yizhi, Pu, Ruitao, Xu, Shilin, Chen, Yingke, Liu, Quan-Hui, Sun, Yuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
by: Pu, Ruitao, et al.
Published: (2025)
by: Pu, Ruitao, et al.
Published: (2025)
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
by: Liu, Hongjun, et al.
Published: (2025)
by: Liu, Hongjun, et al.
Published: (2025)
P$^2$U: Progressive Precision Update For Efficient Model Distribution
by: Afrabandpey, Homayun, et al.
Published: (2025)
by: Afrabandpey, Homayun, et al.
Published: (2025)
DeepTaxon: An Interpretable Retrieval-Augmented Multimodal Framework for Unified Species Identification and Discovery
by: Wang, Jiawei, et al.
Published: (2026)
by: Wang, Jiawei, et al.
Published: (2026)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
by: Tu, Songjun, et al.
Published: (2025)
by: Tu, Songjun, et al.
Published: (2025)
CLIP-Guided Backdoor Defense through Entropy-Based Poisoned Dataset Separation
by: Xu, Binyan, et al.
Published: (2025)
by: Xu, Binyan, et al.
Published: (2025)
Re:Draw -- Context Aware Translation as a Controllable Method for Artistic Production
by: Cardoso, Joao Liborio, et al.
Published: (2024)
by: Cardoso, Joao Liborio, et al.
Published: (2024)
PRISM: Iterative Cross-Modal Posterior Refinement for Dynamic Text-Attributed Graphs
by: Chang, Trimble, et al.
Published: (2026)
by: Chang, Trimble, et al.
Published: (2026)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
by: Niizumi, Daisuke, et al.
Published: (2026)
by: Niizumi, Daisuke, et al.
Published: (2026)
Listen to the Unexpected: Self-Supervised Surprise Detection for Efficient Viewport Prediction
by: Khah, Arman Nik, et al.
Published: (2026)
by: Khah, Arman Nik, et al.
Published: (2026)
Generative AI for Video Translation: A Scalable Architecture for Multilingual Video Conferencing
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
A Unified Optimal Transport Framework for Cross-Modal Retrieval with Noisy Labels
by: Han, Haochen, et al.
Published: (2024)
by: Han, Haochen, et al.
Published: (2024)
Hateful Meme Detection through Context-Sensitive Prompting and Fine-Grained Labeling
by: Ouyang, Rongxin, et al.
Published: (2024)
by: Ouyang, Rongxin, et al.
Published: (2024)
SRPL-SFDA: SAM-Guided Reliable Pseudo-Labels for Source-Free Domain Adaptation in Medical Image Segmentation
by: Liu, Xinya, et al.
Published: (2025)
by: Liu, Xinya, et al.
Published: (2025)
A Multimodal Symphony: Integrating Taste and Sound through Generative AI
by: Spanio, Matteo, et al.
Published: (2025)
by: Spanio, Matteo, et al.
Published: (2025)
Optimized Gradient Clipping for Noisy Label Learning
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Personalized QoE Prediction: A Demographic-Augmented Machine Learning Framework for 5G Video Streaming Networks
by: Ahmed, Syeda Zunaira, et al.
Published: (2025)
by: Ahmed, Syeda Zunaira, et al.
Published: (2025)
Start from Video-Music Retrieval: An Inter-Intra Modal Loss for Cross Modal Retrieval
by: Chen, Zeyu, et al.
Published: (2024)
by: Chen, Zeyu, et al.
Published: (2024)
Leveraging Causal Reasoning Method for Explaining Medical Image Segmentation Models
by: Jiang, Limai, et al.
Published: (2026)
by: Jiang, Limai, et al.
Published: (2026)
A Roadmap for Multilingual, Multimodal Domain Independent Deception Detection
by: Boumber, Dainis, et al.
Published: (2024)
by: Boumber, Dainis, et al.
Published: (2024)
Goal-Based Vision-Language Driving
by: Patapati, Santosh, et al.
Published: (2025)
by: Patapati, Santosh, et al.
Published: (2025)
When Labels Have Structure: Improving Image Classification with Hierarchy-Aware Cross-Entropy
by: Chan, April, et al.
Published: (2026)
by: Chan, April, et al.
Published: (2026)
Deep Reversible Consistency Learning for Cross-modal Retrieval
by: Pu, Ruitao, et al.
Published: (2025)
by: Pu, Ruitao, et al.
Published: (2025)
Annot-Mix: Learning with Noisy Class Labels from Multiple Annotators via a Mixup Extension
by: Herde, Marek, et al.
Published: (2024)
by: Herde, Marek, et al.
Published: (2024)
Evaluating Cross-Modal Reasoning Ability and Problem Characteristics with Multimodal Item Response Theory
by: Uebayashi, Shunki, et al.
Published: (2026)
by: Uebayashi, Shunki, et al.
Published: (2026)
Genflow Ad Studio: A Compound AI Architecture for Brand-Aligned, Self-Correcting Video Generation
by: Das, Debanshu, et al.
Published: (2026)
by: Das, Debanshu, et al.
Published: (2026)
Motion Attribution for Video Generation
by: Wu, Xindi, et al.
Published: (2026)
by: Wu, Xindi, et al.
Published: (2026)
Reasoning is a Modality
by: Liu, Zhiguang, et al.
Published: (2026)
by: Liu, Zhiguang, et al.
Published: (2026)
ChordSync: Conformer-Based Alignment of Chord Annotations to Music Audio
by: Poltronieri, Andrea, et al.
Published: (2024)
by: Poltronieri, Andrea, et al.
Published: (2024)
Human-Corrected Labels Learning: Enhancing Labels Quality via Human Correction of VLMs Discrepancies
by: Li, Zhongnian, et al.
Published: (2025)
by: Li, Zhongnian, et al.
Published: (2025)
CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models
by: Liu, Zhi
Published: (2026)
by: Liu, Zhi
Published: (2026)
Pinching Visuo-haptic Display: Investigating Cross-Modal Effects of Visual Textures on Electrostatic Cloth Tactile Sensations
by: Kitagishi, Takekazu, et al.
Published: (2025)
by: Kitagishi, Takekazu, et al.
Published: (2025)
Trusted Multi-view Learning under Noisy Supervision
by: Zhang, Yilin, et al.
Published: (2024)
by: Zhang, Yilin, et al.
Published: (2024)
From Articles to Canopies: Knowledge-Driven Pseudo-Labelling for Tree Species Classification using LLM Experts
by: Romaszewski, Michał, et al.
Published: (2026)
by: Romaszewski, Michał, et al.
Published: (2026)
AG-REPA: Causal Layer Selection for Representation Alignment in Audio Flow Matching
by: Zhang, Pengfei, et al.
Published: (2026)
by: Zhang, Pengfei, et al.
Published: (2026)
Step-E: A Differentiable Data Cleaning Framework for Robust Learning with Noisy Labels
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Robust Assignment of Labels for Active Learning with Sparse and Noisy Annotations
by: Kałuża, Daniel, et al.
Published: (2023)
by: Kałuża, Daniel, et al.
Published: (2023)
BadSAD: Clean-Label Backdoor Attacks against Deep Semi-Supervised Anomaly Detection
by: Cheng, He, et al.
Published: (2024)
by: Cheng, He, et al.
Published: (2024)
UPAM: Unified Prompt Attack in Text-to-Image Generation Models Against Both Textual Filters and Visual Checkers
by: Peng, Duo, et al.
Published: (2024)
by: Peng, Duo, et al.
Published: (2024)
Similar Items
-
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
by: Pu, Ruitao, et al.
Published: (2025) -
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
by: Liu, Hongjun, et al.
Published: (2025) -
P$^2$U: Progressive Precision Update For Efficient Model Distribution
by: Afrabandpey, Homayun, et al.
Published: (2025) -
DeepTaxon: An Interpretable Retrieval-Augmented Multimodal Framework for Unified Species Identification and Discovery
by: Wang, Jiawei, et al.
Published: (2026) -
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
by: Tu, Songjun, et al.
Published: (2025)