Saved in:
| Main Authors: | Ma, Huanhuan, Zhang, Jinghao, Liu, Qiang, Wu, Shu, Wang, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.04756 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
by: Wu, Junfei, et al.
Published: (2024)
by: Wu, Junfei, et al.
Published: (2024)
Modality-Balanced Learning for Multimedia Recommendation
by: Zhang, Jinghao, et al.
Published: (2024)
by: Zhang, Jinghao, et al.
Published: (2024)
AgriDoctor: A Multimodal Intelligent Assistant for Agriculture
by: Zhang, Mingqing, et al.
Published: (2025)
by: Zhang, Mingqing, et al.
Published: (2025)
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
by: Zhao, Chenchen, et al.
Published: (2026)
by: Zhao, Chenchen, et al.
Published: (2026)
Gradient-Regularized Out-of-Distribution Detection
by: Sharifi, Sina, et al.
Published: (2024)
by: Sharifi, Sina, et al.
Published: (2024)
ForgeryGPT: A Multimodal LLM for Interpretable Image Forgery Detection and Localization
by: Zhang, Fanrui, et al.
Published: (2024)
by: Zhang, Fanrui, et al.
Published: (2024)
Knowledge Regularized Negative Feature Tuning of Vision-Language Models for Out-of-Distribution Detection
by: Zhu, Wenjie, et al.
Published: (2025)
by: Zhu, Wenjie, et al.
Published: (2025)
MOODv2: Masked Image Modeling for Out-of-Distribution Detection
by: Li, Jingyao, et al.
Published: (2024)
by: Li, Jingyao, et al.
Published: (2024)
Vision Also You Need: Navigating Out-of-Distribution Detection with Multimodal Large Language Model
by: Xu, Haoran, et al.
Published: (2026)
by: Xu, Haoran, et al.
Published: (2026)
Interpret Your Decision: Logical Reasoning Regularization for Generalization in Visual Classification
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
Predictive Regularization Against Visual Representation Degradation in Multimodal Large Language Models
by: Wang, Enguang, et al.
Published: (2026)
by: Wang, Enguang, et al.
Published: (2026)
Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology
by: Wu, Kaiyu, et al.
Published: (2026)
by: Wu, Kaiyu, et al.
Published: (2026)
Background Prompt for Few-Shot Out-of-Distribution Detection
by: Cai, Songyue, et al.
Published: (2025)
by: Cai, Songyue, et al.
Published: (2025)
Adapting Visual-Language Models for Generalizable Anomaly Detection in Medical Images
by: Huang, Chaoqin, et al.
Published: (2024)
by: Huang, Chaoqin, et al.
Published: (2024)
Occlusion-Aware 3D Motion Interpretation for Abnormal Behavior Detection
by: Li, Su, et al.
Published: (2024)
by: Li, Su, et al.
Published: (2024)
Temporal Separation with Entropy Regularization for Knowledge Distillation in Spiking Neural Networks
by: Yu, Kairong, et al.
Published: (2025)
by: Yu, Kairong, et al.
Published: (2025)
RUNA: Object-level Out-of-Distribution Detection via Regional Uncertainty Alignment of Multimodal Representations
by: Zhang, Bin, et al.
Published: (2025)
by: Zhang, Bin, et al.
Published: (2025)
MuSLR: Multimodal Symbolic Logical Reasoning
by: Xu, Jundong, et al.
Published: (2025)
by: Xu, Jundong, et al.
Published: (2025)
TopoLogic: An Interpretable Pipeline for Lane Topology Reasoning on Driving Scenes
by: Fu, Yanping, et al.
Published: (2024)
by: Fu, Yanping, et al.
Published: (2024)
Transfer Learning of Real Image Features with Soft Contrastive Loss for Fake Image Detection
by: Liang, Ziyou, et al.
Published: (2024)
by: Liang, Ziyou, et al.
Published: (2024)
Embedding-based Retrieval in Multimodal Content Moderation
by: Liang, Hanzhong, et al.
Published: (2025)
by: Liang, Hanzhong, et al.
Published: (2025)
Interpretable Logical Anomaly Classification via Constraint Decomposition and Instruction Fine-Tuning
by: Zhang, Xufei, et al.
Published: (2026)
by: Zhang, Xufei, et al.
Published: (2026)
Patch-Discontinuity Mining for Generalized Deepfake Detection
by: Yuan, Huanhuan, et al.
Published: (2025)
by: Yuan, Huanhuan, et al.
Published: (2025)
Shedding the Facades, Connecting the Domains: Detecting Shifting Multimodal Hate Video with Test-Time Adaptation
by: Li, Jiao, et al.
Published: (2026)
by: Li, Jiao, et al.
Published: (2026)
Out-of-Distribution Detection Based on Total Variation Estimation
by: Ma, Dabiao, et al.
Published: (2026)
by: Ma, Dabiao, et al.
Published: (2026)
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
by: Cao, Yukang, et al.
Published: (2025)
by: Cao, Yukang, et al.
Published: (2025)
REAR: Rethinking Visual Autoregressive Models via Generator-Tokenizer Consistency Regularization
by: He, Qiyuan, et al.
Published: (2025)
by: He, Qiyuan, et al.
Published: (2025)
Interpretable Face Anti-Spoofing: Enhancing Generalization with Multimodal Large Language Models
by: Zhang, Guosheng, et al.
Published: (2025)
by: Zhang, Guosheng, et al.
Published: (2025)
LAD-Reasoner: Tiny Multimodal Models are Good Reasoners for Logical Anomaly Detection
by: Li, Weijia, et al.
Published: (2025)
by: Li, Weijia, et al.
Published: (2025)
Adaptive Confidence Regularization for Multimodal Failure Detection
by: Liu, Moru, et al.
Published: (2026)
by: Liu, Moru, et al.
Published: (2026)
XR-VLM: Cross-Relationship Modeling with Multi-part Prompts and Visual Features for Fine-Grained Recognition
by: Wang, Chuanming, et al.
Published: (2025)
by: Wang, Chuanming, et al.
Published: (2025)
BiCoR-Seg: Bidirectional Co-Refinement Framework for High-Resolution Remote Sensing Image Segmentation
by: Shi, Jinghao, et al.
Published: (2025)
by: Shi, Jinghao, et al.
Published: (2025)
Revisiting Logit Distributions for Reliable Out-of-Distribution Detection
by: Liang, Jiachen, et al.
Published: (2025)
by: Liang, Jiachen, et al.
Published: (2025)
ClimaOoD: Improving Anomaly Segmentation via Physically Realistic Synthetic Data
by: Liu, Yuxing, et al.
Published: (2025)
by: Liu, Yuxing, et al.
Published: (2025)
LogicOCR: Do Your Large Multimodal Models Excel at Logical Reasoning on Text-Rich Images?
by: Ye, Maoyuan, et al.
Published: (2025)
by: Ye, Maoyuan, et al.
Published: (2025)
Component-Based Out-of-Distribution Detection
by: Liu, Wenrui, et al.
Published: (2026)
by: Liu, Wenrui, et al.
Published: (2026)
FeatureLens: A Highly Generalizable and Interpretable Framework for Detecting Adversarial Examples Based on Image Features
by: Yang, Zhigang, et al.
Published: (2025)
by: Yang, Zhigang, et al.
Published: (2025)
Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation
by: Zhang, Jing, et al.
Published: (2026)
by: Zhang, Jing, et al.
Published: (2026)
WWW: Where, Which and Whatever Enhancing Interpretability in Multimodal Deepfake Detection
by: Jung, Juho, et al.
Published: (2024)
by: Jung, Juho, et al.
Published: (2024)
DPU: Dynamic Prototype Updating for Multimodal Out-of-Distribution Detection
by: Li, Shawn, et al.
Published: (2024)
by: Li, Shawn, et al.
Published: (2024)
Similar Items
-
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
by: Wu, Junfei, et al.
Published: (2024) -
Modality-Balanced Learning for Multimedia Recommendation
by: Zhang, Jinghao, et al.
Published: (2024) -
AgriDoctor: A Multimodal Intelligent Assistant for Agriculture
by: Zhang, Mingqing, et al.
Published: (2025) -
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
by: Zhao, Chenchen, et al.
Published: (2026) -
Gradient-Regularized Out-of-Distribution Detection
by: Sharifi, Sina, et al.
Published: (2024)