SIF: Semantically In-Distribution Fingerprints for Large Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yifei, Lou, Qian, Zheng, Mengxin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Curing Semantic Drift: A Dynamic Approach to Grounding Generation in Large Vision-Language Models
von: Chen, Jiahe, et al.
Veröffentlicht: (2025)
von: Chen, Jiahe, et al.
Veröffentlicht: (2025)
Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models
von: Qin, Mengxin, et al.
Veröffentlicht: (2026)
von: Qin, Mengxin, et al.
Veröffentlicht: (2026)
Temporal-Spatial Object Relations Modeling for Vision-and-Language Navigation
von: Huang, Bowen, et al.
Veröffentlicht: (2024)
von: Huang, Bowen, et al.
Veröffentlicht: (2024)
The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
DIMoE-Adapters: Dynamic Expert Evolution for Continual Learning in Vision-Language Models
von: Qin, Mengxin, et al.
Veröffentlicht: (2026)
von: Qin, Mengxin, et al.
Veröffentlicht: (2026)
VaLiD: Mitigating the Hallucination of Large Vision Language Models by Visual Layer Fusion Contrastive Decoding
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
Robust Prompt Tuning for Vision-Language Models with Mild Semantic Noise
von: Gao, Yansheng, et al.
Veröffentlicht: (2025)
von: Gao, Yansheng, et al.
Veröffentlicht: (2025)
DuSSS: Dual Semantic Similarity-Supervised Vision-Language Model for Semi-Supervised Medical Image Segmentation
von: Pan, Qingtao, et al.
Veröffentlicht: (2024)
von: Pan, Qingtao, et al.
Veröffentlicht: (2024)
Mitigating Hallucination in Large Vision-Language Models through Aligning Attention Distribution to Information Flow
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
How Does Fine-Tuning Impact Out-of-Distribution Detection for Vision-Language Models?
von: Ming, Yifei, et al.
Veröffentlicht: (2023)
von: Ming, Yifei, et al.
Veröffentlicht: (2023)
RCP: Representation Consistency Pruner for Mitigating Distribution Shift in Large Vision-Language Models
von: Zhang, Jianwei, et al.
Veröffentlicht: (2026)
von: Zhang, Jianwei, et al.
Veröffentlicht: (2026)
Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models
von: Lu, Jiaying, et al.
Veröffentlicht: (2023)
von: Lu, Jiaying, et al.
Veröffentlicht: (2023)
CosSIF: Cosine similarity-based image filtering to overcome low inter-class variation in synthetic medical image datasets
von: Islam, Mominul, et al.
Veröffentlicht: (2023)
von: Islam, Mominul, et al.
Veröffentlicht: (2023)
Harnessing Large Language and Vision-Language Models for Robust Out-of-Distribution Detection
von: Lee, Pei-Kang, et al.
Veröffentlicht: (2025)
von: Lee, Pei-Kang, et al.
Veröffentlicht: (2025)
Can Large Vision-Language Models Correct Semantic Grounding Errors By Themselves?
von: Liao, Yuan-Hong, et al.
Veröffentlicht: (2024)
von: Liao, Yuan-Hong, et al.
Veröffentlicht: (2024)
AIGCs Confuse AI Too: Investigating and Explaining Synthetic Image-induced Hallucinations in Large Vision-Language Models
von: Gao, Yifei, et al.
Veröffentlicht: (2024)
von: Gao, Yifei, et al.
Veröffentlicht: (2024)
Temporal Visual Semantics-Induced Human Motion Understanding with Large Language Models
von: Xing, Zheng, et al.
Veröffentlicht: (2025)
von: Xing, Zheng, et al.
Veröffentlicht: (2025)
Remote Sensing Large Vision-Language Model: Semantic-augmented Multi-level Alignment and Semantic-aware Expert Modeling
von: Park, Sungjune, et al.
Veröffentlicht: (2025)
von: Park, Sungjune, et al.
Veröffentlicht: (2025)
FPBench: A Comprehensive Benchmark of Multimodal Large Language Models for Fingerprint Analysis
von: Gavas, Ekta, et al.
Veröffentlicht: (2025)
von: Gavas, Ekta, et al.
Veröffentlicht: (2025)
Exploring Large Vision-Language Models for Robust and Efficient Industrial Anomaly Detection
von: Qian, Kun, et al.
Veröffentlicht: (2024)
von: Qian, Kun, et al.
Veröffentlicht: (2024)
Semantic Alignment for Multimodal Large Language Models
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization
von: Peng, Jingwei, et al.
Veröffentlicht: (2025)
von: Peng, Jingwei, et al.
Veröffentlicht: (2025)
Generalizable Prompt Tuning for Vision-Language Models
von: Zhang, Qian
Veröffentlicht: (2024)
von: Zhang, Qian
Veröffentlicht: (2024)
FOCoOp: Enhancing Out-of-Distribution Robustness in Federated Prompt Learning for Vision-Language Models
von: Liao, Xinting, et al.
Veröffentlicht: (2025)
von: Liao, Xinting, et al.
Veröffentlicht: (2025)
A Survey on Hallucination in Large Vision-Language Models
von: Liu, Hanchao, et al.
Veröffentlicht: (2024)
von: Liu, Hanchao, et al.
Veröffentlicht: (2024)
Iterated Learning Improves Compositionality in Large Vision-Language Models
von: Zheng, Chenhao, et al.
Veröffentlicht: (2024)
von: Zheng, Chenhao, et al.
Veröffentlicht: (2024)
HiMix: Reducing Computational Complexity in Large Vision-Language Models
von: Zhang, Xuange, et al.
Veröffentlicht: (2025)
von: Zhang, Xuange, et al.
Veröffentlicht: (2025)
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
von: Qian, Zefeng, et al.
Veröffentlicht: (2025)
von: Qian, Zefeng, et al.
Veröffentlicht: (2025)
Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection
von: Yang, Le, et al.
Veröffentlicht: (2024)
von: Yang, Le, et al.
Veröffentlicht: (2024)
Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2026)
Long-Tailed Distribution-Aware Router For Mixture-of-Experts in Large Vision-Language Model
von: Cai, Chaoxiang, et al.
Veröffentlicht: (2025)
von: Cai, Chaoxiang, et al.
Veröffentlicht: (2025)
Dude: Dual Distribution-Aware Context Prompt Learning For Large Vision-Language Model
von: Nguyen, Duy M. H., et al.
Veröffentlicht: (2024)
von: Nguyen, Duy M. H., et al.
Veröffentlicht: (2024)
Vision-Centric Activation and Coordination for Multimodal Large Language Models
von: Wang, Yunnan, et al.
Veröffentlicht: (2025)
von: Wang, Yunnan, et al.
Veröffentlicht: (2025)
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning
von: Park, Dongmin, et al.
Veröffentlicht: (2024)
von: Park, Dongmin, et al.
Veröffentlicht: (2024)
Can We Predict Performance of Large Models across Vision-Language Tasks?
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
Challenging Vision-Language Models with Physically Deployable Multimodal Semantic Lighting Attacks
von: Zhao, Yingying, et al.
Veröffentlicht: (2026)
von: Zhao, Yingying, et al.
Veröffentlicht: (2026)
Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models
von: Zhang, Jinrui, et al.
Veröffentlicht: (2024)
von: Zhang, Jinrui, et al.
Veröffentlicht: (2024)
FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
von: Lu, Xinhua, et al.
Veröffentlicht: (2025)
von: Lu, Xinhua, et al.
Veröffentlicht: (2025)
ReCAD: Reinforcement Learning Enhanced Parametric CAD Model Generation with Vision-Language Models
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
Beyond Perception Errors: Semantic Fixation in Large Vision-Language Models
von: Alam, Md Tanvirul
Veröffentlicht: (2026)
von: Alam, Md Tanvirul
Veröffentlicht: (2026)
Ähnliche Einträge
-
Curing Semantic Drift: A Dynamic Approach to Grounding Generation in Large Vision-Language Models
von: Chen, Jiahe, et al.
Veröffentlicht: (2025) -
Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models
von: Qin, Mengxin, et al.
Veröffentlicht: (2026) -
Temporal-Spatial Object Relations Modeling for Vision-and-Language Navigation
von: Huang, Bowen, et al.
Veröffentlicht: (2024) -
The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024) -
DIMoE-Adapters: Dynamic Expert Evolution for Continual Learning in Vision-Language Models
von: Qin, Mengxin, et al.
Veröffentlicht: (2026)