Building Trustworthy Multimodal AI: A Review of Fairness, Transparency, and Ethics in Vision-Language Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Saleh, Mohammad, Tabatabaei, Azadeh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Meta-Sealing: A Revolutionizing Integrity Assurance Protocol for Transparent, Tamper-Proof, and Trustworthy AI System
by: Krishnamoorthy, Mahesh Vaijainthymala
Published: (2024)
by: Krishnamoorthy, Mahesh Vaijainthymala
Published: (2024)
Toward Trustworthy Agentic AI: A Multimodal Framework for Preventing Prompt Injection Attacks
by: Syed, Toqeer Ali, et al.
Published: (2025)
by: Syed, Toqeer Ali, et al.
Published: (2025)
Is On-Device AI Broken and Exploitable? Assessing the Trust and Ethics in Small Language Models
by: Nakka, Kalyan, et al.
Published: (2024)
by: Nakka, Kalyan, et al.
Published: (2024)
Security-First AI: Foundations for Robust and Trustworthy Systems
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
Trustworthy AI-Generative Content for Intelligent Network Service: Robustness, Security, and Fairness
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
AI-Governed Agent Architecture for Web-Trustworthy Tokenization of Alternative Assets
by: Borjigin, Ailiya, et al.
Published: (2025)
by: Borjigin, Ailiya, et al.
Published: (2025)
Trustworthiness Calibration Framework for Phishing Email Detection Using Large Language Models
by: Ganiuly, Daniyal, et al.
Published: (2025)
by: Ganiuly, Daniyal, et al.
Published: (2025)
Tackling Cyberattacks through AI-based Reactive Systems: A Holistic Review and Future Vision
by: Molina, Sergio Bernardez, et al.
Published: (2023)
by: Molina, Sergio Bernardez, et al.
Published: (2023)
Efficient and Trustworthy Block Propagation for Blockchain-enabled Mobile Embodied AI Networks: A Graph Resfusion Approach
by: Kang, Jiawen, et al.
Published: (2025)
by: Kang, Jiawen, et al.
Published: (2025)
On the Foundations of Trustworthy Artificial Intelligence
by: Dunham, TJ
Published: (2026)
by: Dunham, TJ
Published: (2026)
MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models
by: Xu, Chejian, et al.
Published: (2025)
by: Xu, Chejian, et al.
Published: (2025)
Red-teaming the Multimodal Reasoning: Jailbreaking Vision-Language Models via Cross-modal Entanglement Attacks
by: Yan, Yu, et al.
Published: (2026)
by: Yan, Yu, et al.
Published: (2026)
Building A Secure Agentic AI Application Leveraging A2A Protocol
by: Habler, Idan, et al.
Published: (2025)
by: Habler, Idan, et al.
Published: (2025)
TinyGuard:A lightweight Byzantine Defense for Resource-Constrained Federated Learning via Statistical Update Fingerprints
by: Mahdavi, Ali, et al.
Published: (2026)
by: Mahdavi, Ali, et al.
Published: (2026)
Alphabet Index Mapping: Jailbreaking LLMs through Semantic Dissimilarity
by: Husain, Bilal Saleh
Published: (2025)
by: Husain, Bilal Saleh
Published: (2025)
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models
by: Wang, Yanbo, et al.
Published: (2025)
by: Wang, Yanbo, et al.
Published: (2025)
TrajAD: Trajectory Anomaly Detection for Trustworthy LLM Agents
by: Liu, Yibing, et al.
Published: (2026)
by: Liu, Yibing, et al.
Published: (2026)
Atlas: A Framework for ML Lifecycle Provenance & Transparency
by: Spoczynski, Marcin, et al.
Published: (2025)
by: Spoczynski, Marcin, et al.
Published: (2025)
Auspex: Building Threat Modeling Tradecraft into an Artificial Intelligence-based Copilot
by: Crossman, Andrew, et al.
Published: (2025)
by: Crossman, Andrew, et al.
Published: (2025)
Hidden in the Metadata: Stealth Poisoning Attacks on Multimodal Retrieval-Augmented Generation
by: Edemacu, Kennedy, et al.
Published: (2026)
by: Edemacu, Kennedy, et al.
Published: (2026)
Identifying the Supply Chain of AI for Trustworthiness and Risk Management in Critical Applications
by: Sheh, Raymond K., et al.
Published: (2025)
by: Sheh, Raymond K., et al.
Published: (2025)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
by: Qayyum, Adnan, et al.
Published: (2022)
by: Qayyum, Adnan, et al.
Published: (2022)
Zero-Knowledge Federated Learning: A New Trustworthy and Privacy-Preserving Distributed Learning Paradigm
by: Wang, Taotao, et al.
Published: (2025)
by: Wang, Taotao, et al.
Published: (2025)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
by: Wang, Chenrui, et al.
Published: (2025)
by: Wang, Chenrui, et al.
Published: (2025)
Enabling Trustworthy Federated Learning via Remote Attestation for Mitigating Byzantine Threats
by: Zhang, Chaoyu, et al.
Published: (2025)
by: Zhang, Chaoyu, et al.
Published: (2025)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
by: Hall, Peter, et al.
Published: (2025)
by: Hall, Peter, et al.
Published: (2025)
Penetration Testing of Agentic AI: A Comparative Security Analysis Across Models and Frameworks
by: Nguyen, Viet K., et al.
Published: (2025)
by: Nguyen, Viet K., et al.
Published: (2025)
Align is not Enough: Multimodal Universal Jailbreak Attack against Multimodal Large Language Models
by: Wang, Youze, et al.
Published: (2025)
by: Wang, Youze, et al.
Published: (2025)
Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models
by: Teng, Ma, et al.
Published: (2024)
by: Teng, Ma, et al.
Published: (2024)
Adversarial Attacks on Multimodal Large Language Models: A Comprehensive Survey
by: Jain, Bhavuk, et al.
Published: (2026)
by: Jain, Bhavuk, et al.
Published: (2026)
DiffuseTrace: A Transparent and Flexible Watermarking Scheme for Latent Diffusion Model
by: Lei, Liangqi, et al.
Published: (2024)
by: Lei, Liangqi, et al.
Published: (2024)
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
Gen-AI for User Safety: A Survey
by: Desai, Akshar Prabhu, et al.
Published: (2024)
by: Desai, Akshar Prabhu, et al.
Published: (2024)
Concept-Guided Backdoor Attack on Vision Language Models
by: Shen, Haoyu, et al.
Published: (2025)
by: Shen, Haoyu, et al.
Published: (2025)
Membership Inference Attacks Against Vision-Language Models
by: Hu, Yuke, et al.
Published: (2025)
by: Hu, Yuke, et al.
Published: (2025)
Adversarial attacks against Modern Vision-Language Models
by: La Torre, Alejandro Paredes
Published: (2026)
by: La Torre, Alejandro Paredes
Published: (2026)
Towards Trustworthy AI: Secure Deepfake Detection using CNNs and Zero-Knowledge Proofs
by: Islam, H M Mohaimanul, et al.
Published: (2025)
by: Islam, H M Mohaimanul, et al.
Published: (2025)
Less Is More -- Until It Breaks: Security Pitfalls of Vision Token Compression in Large Vision-Language Models
by: Zhang, Xiaomei, et al.
Published: (2026)
by: Zhang, Xiaomei, et al.
Published: (2026)
Text Steganography with Dynamic Codebook and Multimodal Large Language Model
by: Gao, Jianxin, et al.
Published: (2026)
by: Gao, Jianxin, et al.
Published: (2026)
Multimodal Large Language Models for Phishing Webpage Detection and Identification
by: Lee, Jehyun, et al.
Published: (2024)
by: Lee, Jehyun, et al.
Published: (2024)
Similar Items
-
Meta-Sealing: A Revolutionizing Integrity Assurance Protocol for Transparent, Tamper-Proof, and Trustworthy AI System
by: Krishnamoorthy, Mahesh Vaijainthymala
Published: (2024) -
Toward Trustworthy Agentic AI: A Multimodal Framework for Preventing Prompt Injection Attacks
by: Syed, Toqeer Ali, et al.
Published: (2025) -
Is On-Device AI Broken and Exploitable? Assessing the Trust and Ethics in Small Language Models
by: Nakka, Kalyan, et al.
Published: (2024) -
Security-First AI: Foundations for Robust and Trustworthy Systems
by: Tallam, Krti
Published: (2025) -
Trustworthy AI-Generative Content for Intelligent Network Service: Robustness, Security, and Fairness
by: Li, Siyuan, et al.
Published: (2024)