CLASH: A Benchmark for Cross-Modal Contradiction Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Popordanoska, Teodora, Li, Jiameng, Blaschko, Matthew B. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DAVE: Diagnostic benchmark for Audio Visual Evaluation
by: Radevski, Gorjan, et al.
Published: (2025)
by: Radevski, Gorjan, et al.
Published: (2025)
Dice Semimetric Losses: Optimizing the Dice Score with Soft Labels
by: Wang, Zifu, et al.
Published: (2023)
by: Wang, Zifu, et al.
Published: (2023)
Jaccard Metric Losses: Optimizing the Jaccard Index with Soft Labels
by: Wang, Zifu, et al.
Published: (2023)
by: Wang, Zifu, et al.
Published: (2023)
CARE: Confidence-aware Ratio Estimation for Medical Biomarkers
by: Li, Jiameng, et al.
Published: (2025)
by: Li, Jiameng, et al.
Published: (2025)
Revisiting Reweighted Risk for Calibration: AURC, Focal, and Inverse Focal Loss
by: Zhou, Han, et al.
Published: (2025)
by: Zhou, Han, et al.
Published: (2025)
A Structured Benchmark for Text-Guided Anomaly Detection: When Language Stops Conditioning the Decision
by: Samele, Stefano, et al.
Published: (2026)
by: Samele, Stefano, et al.
Published: (2026)
Beyond Perfect Scores: Proof-by-Contradiction for Trustworthy Machine Learning
by: Wadduwage, Dushan N., et al.
Published: (2026)
by: Wadduwage, Dushan N., et al.
Published: (2026)
Cross-Modal Instructions for Robot Motion Generation
by: Barron, William, et al.
Published: (2025)
by: Barron, William, et al.
Published: (2025)
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
by: Van Landeghem, Jordy, et al.
Published: (2024)
by: Van Landeghem, Jordy, et al.
Published: (2024)
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
by: Choi, Eunjee, et al.
Published: (2025)
by: Choi, Eunjee, et al.
Published: (2025)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
by: Li, Xu, et al.
Published: (2025)
by: Li, Xu, et al.
Published: (2025)
EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data
by: Lin, Dongyan, et al.
Published: (2026)
by: Lin, Dongyan, et al.
Published: (2026)
Pre-Trained Model Recommendation for Downstream Fine-tuning
by: Bai, Jiameng, et al.
Published: (2024)
by: Bai, Jiameng, et al.
Published: (2024)
Robust Multimodal Learning via Cross-Modal Proxy Tokens
by: Reza, Md Kaykobad, et al.
Published: (2025)
by: Reza, Md Kaykobad, et al.
Published: (2025)
AmCLR: Unified Augmented Learning for Cross-Modal Representations
by: Jagannath, Ajay, et al.
Published: (2024)
by: Jagannath, Ajay, et al.
Published: (2024)
NAB: Neural Adaptive Binning for Sparse-View CT reconstruction
by: Xie, Wangduo, et al.
Published: (2026)
by: Xie, Wangduo, et al.
Published: (2026)
ProJudge: A Multi-Modal Multi-Discipline Benchmark and Instruction-Tuning Dataset for MLLM-based Process Judges
by: Ai, Jiaxin, et al.
Published: (2025)
by: Ai, Jiaxin, et al.
Published: (2025)
Multimodal Foundation Model for Cross-Modal Retrieval and Activity Recognition Tasks
by: Matsuishi, Koki, et al.
Published: (2025)
by: Matsuishi, Koki, et al.
Published: (2025)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
HeCoFuse: Cross-Modal Complementary V2X Cooperative Perception with Heterogeneous Sensors
by: Wei, Chuheng, et al.
Published: (2025)
by: Wei, Chuheng, et al.
Published: (2025)
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
by: Mistretta, Marco, et al.
Published: (2025)
by: Mistretta, Marco, et al.
Published: (2025)
Cross-Source Supervision for Bone Infection Segmentation in Dual-Modality PET-CT
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data
by: Zhang, Yuhui, et al.
Published: (2024)
by: Zhang, Yuhui, et al.
Published: (2024)
MultiOOD: Scaling Out-of-Distribution Detection for Multiple Modalities
by: Dong, Hao, et al.
Published: (2024)
by: Dong, Hao, et al.
Published: (2024)
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
by: Wang, Yanlong, et al.
Published: (2025)
by: Wang, Yanlong, et al.
Published: (2025)
Hyperdimensional Cross-Modal Alignment of Frozen Language and Image Models for Efficient Image Captioning
by: Dalvi, Abhishek, et al.
Published: (2026)
by: Dalvi, Abhishek, et al.
Published: (2026)
Detecting Dataset Bias in Medical AI: A Generalized and Modality-Agnostic Auditing Framework
by: Drenkow, Nathan, et al.
Published: (2025)
by: Drenkow, Nathan, et al.
Published: (2025)
Urban Waterlogging Detection: A Challenging Benchmark and Large-Small Model Co-Adapter
by: Song, Suqi, et al.
Published: (2024)
by: Song, Suqi, et al.
Published: (2024)
Interleaved-Modal Chain-of-Thought
by: Gao, Jun, et al.
Published: (2024)
by: Gao, Jun, et al.
Published: (2024)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
by: Ienco, Dino, et al.
Published: (2024)
by: Ienco, Dino, et al.
Published: (2024)
COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition
by: Chen, Baiyu, et al.
Published: (2025)
by: Chen, Baiyu, et al.
Published: (2025)
MI-Pruner: Crossmodal Mutual Information-guided Token Pruner for Efficient MLLMs
by: Li, Jiameng, et al.
Published: (2026)
by: Li, Jiameng, et al.
Published: (2026)
Reversed in Time: A Novel Temporal-Emphasized Benchmark for Cross-Modal Video-Text Retrieval
by: Du, Yang, et al.
Published: (2024)
by: Du, Yang, et al.
Published: (2024)
Are Anomaly Scores Telling the Whole Story? A Benchmark for Multilevel Anomaly Detection
by: Cao, Tri, et al.
Published: (2024)
by: Cao, Tri, et al.
Published: (2024)
Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription
by: Gutteridge, Benjamin, et al.
Published: (2025)
by: Gutteridge, Benjamin, et al.
Published: (2025)
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception
by: Wang, Rujia, et al.
Published: (2025)
by: Wang, Rujia, et al.
Published: (2025)
DarkVesselNet: Multi-Modal Remote Sensing and Trajectory Reasoning for Dark Vessel Detection
by: Sharma, Arun
Published: (2026)
by: Sharma, Arun
Published: (2026)
OCR is All you need: Importing Multi-Modality into Image-based Defect Detection System
by: Hsu, Chih-Chung, et al.
Published: (2024)
by: Hsu, Chih-Chung, et al.
Published: (2024)
OCT-SelfNet: A Self-Supervised Framework with Multi-Modal Datasets for Generalized and Robust Retinal Disease Detection
by: Jannat, Fatema-E, et al.
Published: (2024)
by: Jannat, Fatema-E, et al.
Published: (2024)
CrossVL: Complexity-Aware Feature Routing and Paired Curriculum for Cross-View Vision-Language Detection
by: Liu, Zhipeng, et al.
Published: (2026)
by: Liu, Zhipeng, et al.
Published: (2026)
Similar Items
-
DAVE: Diagnostic benchmark for Audio Visual Evaluation
by: Radevski, Gorjan, et al.
Published: (2025) -
Dice Semimetric Losses: Optimizing the Dice Score with Soft Labels
by: Wang, Zifu, et al.
Published: (2023) -
Jaccard Metric Losses: Optimizing the Jaccard Index with Soft Labels
by: Wang, Zifu, et al.
Published: (2023) -
CARE: Confidence-aware Ratio Estimation for Medical Biomarkers
by: Li, Jiameng, et al.
Published: (2025) -
Revisiting Reweighted Risk for Calibration: AURC, Focal, and Inverse Focal Loss
by: Zhou, Han, et al.
Published: (2025)