MTRE: Multi-Token Reliability Estimation for Hallucination Detection in VLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Zollicoffer, Geigh, Vu, Minh, Bhattarai, Manish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Topological Signatures of Adversaries in Multimodal Alignments
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
LoRID: Low-Rank Iterative Diffusion for Adversarial Purification
by: Zollicoffer, Geigh, et al.
Published: (2024)
by: Zollicoffer, Geigh, et al.
Published: (2024)
LaFA: Latent Feature Attacks on Non-negative Matrix Factorization
by: Vu, Minh, et al.
Published: (2024)
by: Vu, Minh, et al.
Published: (2024)
Sanity Checks for Long-Form Hallucination Detection
by: Zollicoffer, Geigh, et al.
Published: (2026)
by: Zollicoffer, Geigh, et al.
Published: (2026)
Trustworthy Agentic AI Requires Deterministic Architectural Boundaries
by: Bhattarai, Manish, et al.
Published: (2026)
by: Bhattarai, Manish, et al.
Published: (2026)
HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
by: Bhatt, Manish
Published: (2026)
by: Bhatt, Manish
Published: (2026)
Privacy Enhanced PEFT: Tensor Train Decomposition Improves Privacy Utility Tradeoffs under DP-SGD
by: Kunwar, Pradip, et al.
Published: (2026)
by: Kunwar, Pradip, et al.
Published: (2026)
Theoretically Unmasking Inference Attacks Against LDP-Protected Clients in Federated Vision Models
by: Nguyen, Quan, et al.
Published: (2025)
by: Nguyen, Quan, et al.
Published: (2025)
GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning
by: Liu, Yue, et al.
Published: (2025)
by: Liu, Yue, et al.
Published: (2025)
VisualDAN: Exposing Vulnerabilities in VLMs with Visual-Driven DAN Commands
by: Liu, Aofan, et al.
Published: (2025)
by: Liu, Aofan, et al.
Published: (2025)
How Far Are VLMs from Privacy Awareness in the Physical World? An Empirical Study
by: Wang, Junran, et al.
Published: (2026)
by: Wang, Junran, et al.
Published: (2026)
AI-Powered Anomaly Detection with Blockchain for Real-Time Security and Reliability in Autonomous Vehicles
by: Shit, Rathin Chandra, et al.
Published: (2025)
by: Shit, Rathin Chandra, et al.
Published: (2025)
Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models
by: Braun, Tobias, et al.
Published: (2026)
by: Braun, Tobias, et al.
Published: (2026)
Hallucination as Exploit: Evidence-Carrying Multimodal Agents
by: Zhang, Guijia, et al.
Published: (2026)
by: Zhang, Guijia, et al.
Published: (2026)
Contextualized AI for Cyber Defense: An Automated Survey using LLMs
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
Hallucination-Resistant Security Planning with a Large Language Model
by: Hammar, Kim, et al.
Published: (2026)
by: Hammar, Kim, et al.
Published: (2026)
Is Mamba Reliable for Medical Imaging?
by: Latibari, Banafsheh Saber, et al.
Published: (2026)
by: Latibari, Banafsheh Saber, et al.
Published: (2026)
Token-Efficient Prompt Injection Attack: Provoking Cessation in LLM Reasoning via Adaptive Token Compression
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Hallucinating AI Hijacking Attack: Large Language Models and Malicious Code Recommenders
by: Noever, David, et al.
Published: (2024)
by: Noever, David, et al.
Published: (2024)
Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Models
by: Truong, Vu Tuan, et al.
Published: (2026)
by: Truong, Vu Tuan, et al.
Published: (2026)
Tracking Software Security Topics
by: Vu, Phong Minh, et al.
Published: (2024)
by: Vu, Phong Minh, et al.
Published: (2024)
Incident Response Planning Using a Lightweight Large Language Model with Reduced Hallucination
by: Hammar, Kim, et al.
Published: (2025)
by: Hammar, Kim, et al.
Published: (2025)
Latent Adversarial Detection: Adaptive Probing of LLM Activations for Multi-Turn Attack Detection
by: Kulkarni, Prashant
Published: (2026)
by: Kulkarni, Prashant
Published: (2026)
BadSNN: Backdoor Attacks on Spiking Neural Networks via Adversarial Spiking Neuron
by: Miah, Abdullah Arafat, et al.
Published: (2026)
by: Miah, Abdullah Arafat, et al.
Published: (2026)
Membership Inference Attacks on Tokenizers of Large Language Models
by: Tong, Meng, et al.
Published: (2025)
by: Tong, Meng, et al.
Published: (2025)
CORVUS: Red-Teaming Hallucination Detectors via Internal Signal Camouflage in Large Language Models
by: Min, Nay Myat, et al.
Published: (2026)
by: Min, Nay Myat, et al.
Published: (2026)
On the Reliability and Stability of Selective Methods in Malware Classification Tasks
by: Herzog, Alexander, et al.
Published: (2025)
by: Herzog, Alexander, et al.
Published: (2025)
Domain-Adapted Granger Causality for Real-Time Cross-Slice Attack Attribution in 6G Networks
by: Quan, Minh K., et al.
Published: (2025)
by: Quan, Minh K., et al.
Published: (2025)
Certified Causal Attribution for Real-Time Attack Forensics in 6G Network Slicing
by: Quan, Minh K., et al.
Published: (2026)
by: Quan, Minh K., et al.
Published: (2026)
Unsupervised Threat Hunting using Continuous Bag-of-Terms-and-Time (CBoTT)
by: Kayhan, Varol, et al.
Published: (2024)
by: Kayhan, Varol, et al.
Published: (2024)
MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection
by: Xue, Yinuo, et al.
Published: (2025)
by: Xue, Yinuo, et al.
Published: (2025)
Use of Multi-CNNs for Section Analysis in Static Malware Detection
by: Quertier, Tony, et al.
Published: (2024)
by: Quertier, Tony, et al.
Published: (2024)
iSeal: Encrypted Fingerprinting for Reliable LLM Ownership Verification
by: Xiong, Zixun, et al.
Published: (2025)
by: Xiong, Zixun, et al.
Published: (2025)
Lost in Modality: Evaluating the Effectiveness of Text-Based Membership Inference Attacks on Large Multimodal Models
by: Tong, Ziyi, et al.
Published: (2025)
by: Tong, Ziyi, et al.
Published: (2025)
AI-Governed Agent Architecture for Web-Trustworthy Tokenization of Alternative Assets
by: Borjigin, Ailiya, et al.
Published: (2025)
by: Borjigin, Ailiya, et al.
Published: (2025)
Understanding Secret Leakage Risks in Code LLMs: A Tokenization Perspective
by: Chen, Meifang, et al.
Published: (2026)
by: Chen, Meifang, et al.
Published: (2026)
Privacy Guard & Token Parsimony by Prompt and Context Handling and LLM Routing
by: Langiu, Alessio
Published: (2026)
by: Langiu, Alessio
Published: (2026)
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
by: Cai, Jiacheng, et al.
Published: (2024)
by: Cai, Jiacheng, et al.
Published: (2024)
TokenMark: A Modality-Agnostic Watermark for Pre-trained Transformers
by: Xu, Hengyuan, et al.
Published: (2024)
by: Xu, Hengyuan, et al.
Published: (2024)
Similar Items
-
Topological Signatures of Adversaries in Multimodal Alignments
by: Vu, Minh, et al.
Published: (2025) -
LoRID: Low-Rank Iterative Diffusion for Adversarial Purification
by: Zollicoffer, Geigh, et al.
Published: (2024) -
LaFA: Latent Feature Attacks on Non-negative Matrix Factorization
by: Vu, Minh, et al.
Published: (2024) -
Sanity Checks for Long-Form Hallucination Detection
by: Zollicoffer, Geigh, et al.
Published: (2026) -
Trustworthy Agentic AI Requires Deterministic Architectural Boundaries
by: Bhattarai, Manish, et al.
Published: (2026)