VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Raza, Shaina, Vayani, Ashmal, Jain, Aditya, Narayanan, Aravind, Khazaie, Vahid Reza, Bashir, Syed Raza, Dolatabadi, Elham, Uddin, Gias, Emmanouilidis, Christos, Qureshi, Rizwan, Shah, Mubarak |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
by: Narayanan, Aravind, et al.
Published: (2025)
by: Narayanan, Aravind, et al.
Published: (2025)
HumaniBench: A Human-Centric Framework for Large Multimodal Models Evaluation
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
by: Raval, Ananya, et al.
Published: (2025)
by: Raval, Ananya, et al.
Published: (2025)
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
by: Bashir, Syed Raza, et al.
Published: (2024)
by: Bashir, Syed Raza, et al.
Published: (2024)
Just as Humans Need Vaccines, So Do Models: Model Immunization to Combat Falsehoods
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
by: Salimian, Sina, et al.
Published: (2025)
by: Salimian, Sina, et al.
Published: (2025)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Beyond Content: How Grammatical Gender Shapes Visual Representation in Text-to-Image Models
by: Saeed, Muhammed, et al.
Published: (2025)
by: Saeed, Muhammed, et al.
Published: (2025)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
by: Kohankhaki, Farnaz, et al.
Published: (2024)
by: Kohankhaki, Farnaz, et al.
Published: (2024)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
by: Radwan, Ahmed Y., et al.
Published: (2026)
by: Radwan, Ahmed Y., et al.
Published: (2026)
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
by: Raza, Shaina
Published: (2026)
by: Raza, Shaina
Published: (2026)
Position: Beyond Assistance -- Reimagining LLMs as Ethical and Adaptive Co-Creators in Mental Health Care
by: Badawi, Abeer, et al.
Published: (2025)
by: Badawi, Abeer, et al.
Published: (2025)
BBQ-V: Benchmarking Visual Stereotype Bias in Large Multimodal Models
by: Narnaware, Vishal, et al.
Published: (2025)
by: Narnaware, Vishal, et al.
Published: (2025)
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Academic case reports lack diversity: Assessing the presence and diversity of sociodemographic and behavioral factors related to Post COVID-19 Condition
by: Florez, Juan Andres Medina, et al.
Published: (2025)
by: Florez, Juan Andres Medina, et al.
Published: (2025)
MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis
by: Vayani, Ashmal, et al.
Published: (2026)
by: Vayani, Ashmal, et al.
Published: (2026)
Learning to Share: Selective Memory for Efficient Parallel Agentic Systems
by: Fioresi, Joseph, et al.
Published: (2026)
by: Fioresi, Joseph, et al.
Published: (2026)
Who is Responsible? The Data, Models, Users or Regulations? A Comprehensive Survey on Responsible Generative AI for a Sustainable Future
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning
by: Roy, Shuvendu, et al.
Published: (2024)
by: Roy, Shuvendu, et al.
Published: (2024)
Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models
by: Khan, Sumra, et al.
Published: (2026)
by: Khan, Sumra, et al.
Published: (2026)
GAEA: A Geolocation Aware Conversational Assistant
by: Campos, Ron, et al.
Published: (2025)
by: Campos, Ron, et al.
Published: (2025)
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
by: Qureshi, Rizwan, et al.
Published: (2025)
by: Qureshi, Rizwan, et al.
Published: (2025)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
by: Chatrath, Veronica, et al.
Published: (2024)
by: Chatrath, Veronica, et al.
Published: (2024)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories
by: Ho, Jessee, et al.
Published: (2026)
by: Ho, Jessee, et al.
Published: (2026)
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
by: Raza, Shaina, et al.
Published: (2023)
by: Raza, Shaina, et al.
Published: (2023)
Equity in Healthcare: Analyzing Disparities in Machine Learning Predictions of Diabetic Patient Readmissions
by: Al-Zanbouri, Zainab, et al.
Published: (2024)
by: Al-Zanbouri, Zainab, et al.
Published: (2024)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Optimized Log Parsing with Syntactic Modifications
by: Enan, Nafid, et al.
Published: (2025)
by: Enan, Nafid, et al.
Published: (2025)
Five‐In‐One Antibacterial Strategy: A Mn(I) Complex Lights Up the Fight Against Tuberculosis
by: Maryam Bashir, et al.
Published: (2025)
by: Maryam Bashir, et al.
Published: (2025)
Advancing Medical Representation Learning Through High-Quality Data
by: Baghbanzadeh, Negin, et al.
Published: (2025)
by: Baghbanzadeh, Negin, et al.
Published: (2025)
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding
by: Mahmood, Ahmad, et al.
Published: (2024)
by: Mahmood, Ahmad, et al.
Published: (2024)
Can Generative Models Improve Self-Supervised Representation Learning?
by: Ayromlou, Sana, et al.
Published: (2024)
by: Ayromlou, Sana, et al.
Published: (2024)
From Features to Actions: Explainability in Traditional and Agentic AI Systems
by: Chaduvula, Sindhuja, et al.
Published: (2026)
by: Chaduvula, Sindhuja, et al.
Published: (2026)
A Flexible Fairness Framework with Surrogate Loss Reweighting for Addressing Sociodemographic Disparities
by: Xu, Wen, et al.
Published: (2025)
by: Xu, Wen, et al.
Published: (2025)
Similar Items
-
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
by: Narayanan, Aravind, et al.
Published: (2025) -
HumaniBench: A Human-Centric Framework for Large Multimodal Models Evaluation
by: Raza, Shaina, et al.
Published: (2025) -
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
by: Raval, Ananya, et al.
Published: (2025) -
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
by: Bashir, Syed Raza, et al.
Published: (2024) -
Just as Humans Need Vaccines, So Do Models: Model Immunization to Combat Falsehoods
by: Raza, Shaina, et al.
Published: (2025)