Semantic-Aware Adversarial Training for Reliable Deep Hashing Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Xu, Zhang, Zheng, Wang, Xunguang, Wu, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
von: Ouyang, Pengxiang, et al.
Veröffentlicht: (2025)
von: Ouyang, Pengxiang, et al.
Veröffentlicht: (2025)
InvZW: Invariant Feature Learning via Noise-Adversarial Training for Robust Image Zero-Watermarking
von: Tanvir, Abdullah All, et al.
Veröffentlicht: (2025)
von: Tanvir, Abdullah All, et al.
Veröffentlicht: (2025)
Adversarially Robust Deepfake Detection via Adversarial Feature Similarity Learning
von: Khan, Sarwar
Veröffentlicht: (2024)
von: Khan, Sarwar
Veröffentlicht: (2024)
SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation
von: Li, Xuewei, et al.
Veröffentlicht: (2023)
von: Li, Xuewei, et al.
Veröffentlicht: (2023)
Knowledge Bridger: Towards Training-free Missing Modality Completion
von: Ke, Guanzhou, et al.
Veröffentlicht: (2025)
von: Ke, Guanzhou, et al.
Veröffentlicht: (2025)
EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE
von: Chen, Junyi, et al.
Veröffentlicht: (2023)
von: Chen, Junyi, et al.
Veröffentlicht: (2023)
LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
Clean Image May be Dangerous: Data Poisoning Attacks Against Deep Hashing
von: Li, Shuai, et al.
Veröffentlicht: (2025)
von: Li, Shuai, et al.
Veröffentlicht: (2025)
Learning from Mistakes: Self-Regularizing Hierarchical Representations in Point Cloud Semantic Segmentation
von: Camuffo, Elena, et al.
Veröffentlicht: (2023)
von: Camuffo, Elena, et al.
Veröffentlicht: (2023)
VIVAT: Virtuous Improving VAE Training through Artifact Mitigation
von: Novitskiy, Lev, et al.
Veröffentlicht: (2025)
von: Novitskiy, Lev, et al.
Veröffentlicht: (2025)
Deep Learning-based Text-in-Image Watermarking
von: Karki, Bishwa, et al.
Veröffentlicht: (2024)
von: Karki, Bishwa, et al.
Veröffentlicht: (2024)
Storybooth: Training-free Multi-Subject Consistency for Improved Visual Storytelling
von: Singh, Jaskirat, et al.
Veröffentlicht: (2025)
von: Singh, Jaskirat, et al.
Veröffentlicht: (2025)
Omnidirectional Video Super-Resolution using Deep Learning
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
M3R: Localized Rainfall Nowcasting with Meteorology-Informed MultiModal Attention
von: Panta, Sanjeev, et al.
Veröffentlicht: (2026)
von: Panta, Sanjeev, et al.
Veröffentlicht: (2026)
Improving Generative Adversarial Network Generalization for Facial Expression Synthesis
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
Residual Prior-driven Frequency-aware Network for Image Fusion
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
BadCM: Invisible Backdoor Attack Against Cross-Modal Learning
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
Video DataFlywheel: Resolving the Impossible Data Trinity in Video-Language Understanding
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Balanced Multi-modal Federated Learning via Cross-Modal Infiltration
von: Fan, Yunfeng, et al.
Veröffentlicht: (2023)
von: Fan, Yunfeng, et al.
Veröffentlicht: (2023)
Regularized Contrastive Partial Multi-view Outlier Detection
von: Wang, Yijia, et al.
Veröffentlicht: (2024)
von: Wang, Yijia, et al.
Veröffentlicht: (2024)
Detection of Cyberbullying in GIF using AI
von: Dave, Pal, et al.
Veröffentlicht: (2025)
von: Dave, Pal, et al.
Veröffentlicht: (2025)
NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
von: Nasir, Tayyab, et al.
Veröffentlicht: (2026)
von: Nasir, Tayyab, et al.
Veröffentlicht: (2026)
Overcome Modal Bias in Multi-modal Federated Learning via Balanced Modality Selection
von: Fan, Yunfeng, et al.
Veröffentlicht: (2023)
von: Fan, Yunfeng, et al.
Veröffentlicht: (2023)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
PRINTER:Deformation-Aware Adversarial Learning for Virtual IHC Staining with In Situ Fidelity
von: Yuan, Yizhe, et al.
Veröffentlicht: (2025)
von: Yuan, Yizhe, et al.
Veröffentlicht: (2025)
4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer's Diagnosis
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
MVP: Winning Solution to SMP Challenge 2025 Video Track
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
MTCAE-DFER: Multi-Task Cascaded Autoencoder for Dynamic Facial Expression Recognition
von: Xiang, Peihao, et al.
Veröffentlicht: (2024)
von: Xiang, Peihao, et al.
Veröffentlicht: (2024)
Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval
von: Li, Jun, et al.
Veröffentlicht: (2026)
von: Li, Jun, et al.
Veröffentlicht: (2026)
LinVT: Empower Your Image-level Large Language Model to Understand Videos
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
Generalized Jersey Number Recognition Using Multi-task Learning With Orientation-guided Weight Refinement
von: Lin, Yung-Hui, et al.
Veröffentlicht: (2024)
von: Lin, Yung-Hui, et al.
Veröffentlicht: (2024)
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
von: Yin, Xiangchen, et al.
Veröffentlicht: (2025)
von: Yin, Xiangchen, et al.
Veröffentlicht: (2025)
Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
Bridging Compressed Image Latents and Multimodal Large Language Models
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
MCE: Towards a General Framework for Handling Missing Modalities under Imbalanced Missing Rates
von: Zhao, Binyu, et al.
Veröffentlicht: (2025)
von: Zhao, Binyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
von: Pu, Ruitao, et al.
Veröffentlicht: (2025) -
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
von: Ouyang, Pengxiang, et al.
Veröffentlicht: (2025) -
InvZW: Invariant Feature Learning via Noise-Adversarial Training for Robust Image Zero-Watermarking
von: Tanvir, Abdullah All, et al.
Veröffentlicht: (2025) -
Adversarially Robust Deepfake Detection via Adversarial Feature Similarity Learning
von: Khan, Sarwar
Veröffentlicht: (2024) -
SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation
von: Li, Xuewei, et al.
Veröffentlicht: (2023)