Saved in:
| Main Authors: | Liao, Chih-Ting, Chen, Zhangquan, Meng, Chunlei, Huang, Tzu-Yu, Cao, Xin, Zheng, Xu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.11895 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
UniMoCo: Unified Modality Completion for Robust Multi-Modal Embeddings
by: Qin, Jiajun, et al.
Published: (2025)
by: Qin, Jiajun, et al.
Published: (2025)
RLBind: Adversarial-Invariant Cross-Modal Alignment for Unified Robust Embeddings
by: Lu, Yuhong
Published: (2025)
by: Lu, Yuhong
Published: (2025)
Super Encoding Network: Recursive Association of Multi-Modal Encoders for Video Understanding
by: Chen, Boyu, et al.
Published: (2025)
by: Chen, Boyu, et al.
Published: (2025)
Modality Agnostic Efficient Long Range Encoder
by: Parag, Toufiq, et al.
Published: (2025)
by: Parag, Toufiq, et al.
Published: (2025)
Adversarial Robustness in RGB-Skeleton Action Recognition: Leveraging Attention Modality Reweighter
by: Liu, Chao, et al.
Published: (2024)
by: Liu, Chao, et al.
Published: (2024)
Multi-Modal Face Anti-Spoofing via Cross-Modal Feature Transitions
by: Chong, Jun-Xiong, et al.
Published: (2025)
by: Chong, Jun-Xiong, et al.
Published: (2025)
Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
by: Zhang, Dapeng, et al.
Published: (2025)
by: Zhang, Dapeng, et al.
Published: (2025)
Expanding Event Modality Applications through a Robust CLIP-Based Encoder
by: Jeong, Sungheon, et al.
Published: (2024)
by: Jeong, Sungheon, et al.
Published: (2024)
u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model
by: Xu, Jinjin, et al.
Published: (2023)
by: Xu, Jinjin, et al.
Published: (2023)
Adversarial Versus Federated: An Adversarial Learning based Multi-Modality Cross-Domain Federated Medical Segmentation
by: Zhou, You, et al.
Published: (2025)
by: Zhou, You, et al.
Published: (2025)
Hyper Adversarial Tuning for Boosting Adversarial Robustness of Pretrained Large Vision Models
by: Lv, Kangtao, et al.
Published: (2024)
by: Lv, Kangtao, et al.
Published: (2024)
EfficienT-HDR: An Efficient Transformer-Based Framework via Multi-Exposure Fusion for HDR Reconstruction
by: Huang, Yu-Shen, et al.
Published: (2025)
by: Huang, Yu-Shen, et al.
Published: (2025)
UNIT: Unifying Image and Text Recognition in One Vision Encoder
by: Zhu, Yi, et al.
Published: (2024)
by: Zhu, Yi, et al.
Published: (2024)
MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding
by: Jin, Xin, et al.
Published: (2025)
by: Jin, Xin, et al.
Published: (2025)
Physical Adversarial Camouflage through Gradient Calibration and Regularization
by: Liang, Jiawei, et al.
Published: (2025)
by: Liang, Jiawei, et al.
Published: (2025)
Robust Pedestrian Detection with Uncertain Modality
by: Bie, Qian, et al.
Published: (2026)
by: Bie, Qian, et al.
Published: (2026)
UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation
by: Huang, Jiehui, et al.
Published: (2025)
by: Huang, Jiehui, et al.
Published: (2025)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
Robust Multi-Modal Face Anti-Spoofing with Domain Adaptation: Tackling Missing Modalities, Noisy Pseudo-Labels, and Model Degradation
by: Hsu, Ming-Tsung, et al.
Published: (2025)
by: Hsu, Ming-Tsung, et al.
Published: (2025)
Perceive and Calibrate: Analyzing and Enhancing Robustness of Medical Multi-Modal Large Language Models
by: XU, Dunyuan, et al.
Published: (2025)
by: XU, Dunyuan, et al.
Published: (2025)
Robust and Efficient Adversarial Defense in SNNs via Image Purification and Joint Detection
by: Chen, Weiran, et al.
Published: (2024)
by: Chen, Weiran, et al.
Published: (2024)
Towards Imperceptible JPEG Image Hiding: Multi-range Representations-driven Adversarial Stego Generation
by: Yang, Junxue, et al.
Published: (2025)
by: Yang, Junxue, et al.
Published: (2025)
Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
Unified Map Prior Encoder for Mapping and Planning
by: Zhang, Zongzheng, et al.
Published: (2026)
by: Zhang, Zongzheng, et al.
Published: (2026)
Imperceptible Face Forgery Attack via Adversarial Semantic Mask
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Benchmarking Multi-modal Semantic Segmentation under Sensor Failures: Missing and Noisy Modality Robustness
by: Liao, Chenfei, et al.
Published: (2025)
by: Liao, Chenfei, et al.
Published: (2025)
M2IST: Multi-Modal Interactive Side-Tuning for Efficient Referring Expression Comprehension
by: Liu, Xuyang, et al.
Published: (2024)
by: Liu, Xuyang, et al.
Published: (2024)
Uni-Encoder Meets Multi-Encoders: Representation Before Fusion for Brain Tumor Segmentation with Missing Modalities
by: Song, Peibo, et al.
Published: (2026)
by: Song, Peibo, et al.
Published: (2026)
BiXFormer: A Robust Framework for Maximizing Modality Effectiveness in Multi-Modal Semantic Segmentation
by: Chen, Jialei, et al.
Published: (2025)
by: Chen, Jialei, et al.
Published: (2025)
Adaptive Calibration: A Unified Conversion Framework of Spiking Neural Networks
by: Wang, Ziqing, et al.
Published: (2023)
by: Wang, Ziqing, et al.
Published: (2023)
MRStyle: A Unified Framework for Color Style Transfer with Multi-Modality Reference
by: Huang, Jiancheng, et al.
Published: (2024)
by: Huang, Jiancheng, et al.
Published: (2024)
MemorySAM: Memorize Modalities and Semantics with Segment Anything Model 2 for Multi-modal Semantic Segmentation
by: Liao, Chenfei, et al.
Published: (2025)
by: Liao, Chenfei, et al.
Published: (2025)
Self-Calibrated Consistency can Fight Back for Adversarial Robustness in Vision-Language Models
by: Liu, Jiaxiang, et al.
Published: (2025)
by: Liu, Jiaxiang, et al.
Published: (2025)
You Only Need Two Detectors to Achieve Multi-Modal 3D Multi-Object Tracking
by: Wang, Xiyang, et al.
Published: (2023)
by: Wang, Xiyang, et al.
Published: (2023)
MAGIC++: Efficient and Resilient Modality-Agnostic Semantic Segmentation via Hierarchical Modality Selection
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation
by: Meng, Rang, et al.
Published: (2025)
by: Meng, Rang, et al.
Published: (2025)
Efficient Image-to-Image Diffusion Classifier for Adversarial Robustness
by: Mei, Hefei, et al.
Published: (2024)
by: Mei, Hefei, et al.
Published: (2024)
MM-GTUNets: Unified Multi-Modal Graph Deep Learning for Brain Disorders Prediction
by: Cai, Luhui, et al.
Published: (2024)
by: Cai, Luhui, et al.
Published: (2024)
Patch-as-Decodable-Token: Towards Unified Multi-Modal Vision Tasks in MLLMs
by: Su, Yongyi, et al.
Published: (2025)
by: Su, Yongyi, et al.
Published: (2025)
Similar Items
-
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024) -
UniMoCo: Unified Modality Completion for Robust Multi-Modal Embeddings
by: Qin, Jiajun, et al.
Published: (2025) -
RLBind: Adversarial-Invariant Cross-Modal Alignment for Unified Robust Embeddings
by: Lu, Yuhong
Published: (2025) -
Super Encoding Network: Recursive Association of Multi-Modal Encoders for Video Understanding
by: Chen, Boyu, et al.
Published: (2025) -
Modality Agnostic Efficient Long Range Encoder
by: Parag, Toufiq, et al.
Published: (2025)