Just as Humans Need Vaccines, So Do Models: Model Immunization to Combat Falsehoods
Fuente:
arXiv
Guardado en:
| Autores principales: | Raza, Shaina, Qureshi, Rizwan, Farooq, Azib, Lotif, Marcelo, Chadha, Aman, Pandya, Deval, Emmanouilidis, Christos |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024)
por: Chatrath, Veronica, et al.
Publicado: (2024)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
por: Radwan, Ahmed Y., et al.
Publicado: (2026)
por: Radwan, Ahmed Y., et al.
Publicado: (2026)
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
por: Chaduvula, Sindhuja, et al.
Publicado: (2026)
por: Chaduvula, Sindhuja, et al.
Publicado: (2026)
FairSense-AI: Responsible AI Meets Sustainability
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
por: Sapkota, Ranjan, et al.
Publicado: (2025)
por: Sapkota, Ranjan, et al.
Publicado: (2025)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
por: Kohankhaki, Farnaz, et al.
Publicado: (2024)
por: Kohankhaki, Farnaz, et al.
Publicado: (2024)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
por: Garg, Muskan, et al.
Publicado: (2024)
por: Garg, Muskan, et al.
Publicado: (2024)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
AI in Pakistani Schools: Adoption, Usage, and Perceived Impact among Educators
por: Raza, Syed Hassan, et al.
Publicado: (2025)
por: Raza, Syed Hassan, et al.
Publicado: (2025)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
por: Raval, Ananya, et al.
Publicado: (2025)
por: Raval, Ananya, et al.
Publicado: (2025)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
The Psychology of Falsehood: A Human-Centric Survey of Misinformation Detection
por: Nandi, Arghodeep, et al.
Publicado: (2025)
por: Nandi, Arghodeep, et al.
Publicado: (2025)
The Energy of Falsehood: Detecting Hallucinations via Diffusion Model Likelihoods
por: Gautam, Arpit Singh, et al.
Publicado: (2026)
por: Gautam, Arpit Singh, et al.
Publicado: (2026)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
por: Sinha, Neelabh, et al.
Publicado: (2024)
por: Sinha, Neelabh, et al.
Publicado: (2024)
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
por: Bashir, Syed Raza, et al.
Publicado: (2024)
por: Bashir, Syed Raza, et al.
Publicado: (2024)
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
por: Raza, Shaina, et al.
Publicado: (2023)
por: Raza, Shaina, et al.
Publicado: (2023)
How Well Do LLMs Represent Values Across Cultures? Empirical Analysis of LLM Responses Based on Hofstede Cultural Dimensions
por: Kharchenko, Julia, et al.
Publicado: (2024)
por: Kharchenko, Julia, et al.
Publicado: (2024)
DanceText: A Training-Free Layered Framework for Controllable Multilingual Text Transformation in Images
por: Yu, Zhenyu, et al.
Publicado: (2025)
por: Yu, Zhenyu, et al.
Publicado: (2025)
HumaniBench: A Human-Centric Framework for Large Multimodal Models Evaluation
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
The Rise of Small Language Models in Healthcare: A Comprehensive Survey
por: Garg, Muskan, et al.
Publicado: (2025)
por: Garg, Muskan, et al.
Publicado: (2025)
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
por: Das, Nilanjana, et al.
Publicado: (2024)
por: Das, Nilanjana, et al.
Publicado: (2024)
Density Adaptive Attention is All You Need: Robust Parameter-Efficient Fine-Tuning Across Multiple Modalities
por: Ioannides, Georgios, et al.
Publicado: (2024)
por: Ioannides, Georgios, et al.
Publicado: (2024)
Guiding Vision-Language Model Selection for Visual Question-Answering Across Tasks, Domains, and Knowledge Types
por: Sinha, Neelabh, et al.
Publicado: (2024)
por: Sinha, Neelabh, et al.
Publicado: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
por: Chowdhury, Arijit Ghosh, et al.
Publicado: (2023)
por: Chowdhury, Arijit Ghosh, et al.
Publicado: (2023)
BEADs: Bias Evaluation Across Domains
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Breaking Language Barriers: A Question Answering Dataset for Hindi and Marathi
por: Sabane, Maithili, et al.
Publicado: (2023)
por: Sabane, Maithili, et al.
Publicado: (2023)
Unboxing Occupational Bias: Grounded Debiasing of LLMs with U.S. Labor Data
por: Gorti, Atmika, et al.
Publicado: (2024)
por: Gorti, Atmika, et al.
Publicado: (2024)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
por: Singh, Smriti, et al.
Publicado: (2024)
por: Singh, Smriti, et al.
Publicado: (2024)
Can Large Language Models Infer Causal Relationships from Real-World Text?
por: Saklad, Ryan, et al.
Publicado: (2025)
por: Saklad, Ryan, et al.
Publicado: (2025)
Fool Me, Fool Me: User Attitudes Toward LLM Falsehoods
por: Nirman, Diana Bar-Or, et al.
Publicado: (2024)
por: Nirman, Diana Bar-Or, et al.
Publicado: (2024)
Beyond Content: How Grammatical Gender Shapes Visual Representation in Text-to-Image Models
por: Saeed, Muhammed, et al.
Publicado: (2025)
por: Saeed, Muhammed, et al.
Publicado: (2025)
Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
por: Budagam, Devichand, et al.
Publicado: (2024)
por: Budagam, Devichand, et al.
Publicado: (2024)
Out-of-Distribution Detection with Attention Head Masking for Multimodal Document Classification
por: Constantinou, Christos, et al.
Publicado: (2024)
por: Constantinou, Christos, et al.
Publicado: (2024)
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
por: Salimian, Sina, et al.
Publicado: (2025)
por: Salimian, Sina, et al.
Publicado: (2025)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
por: Khoshnoodi, Mahsa, et al.
Publicado: (2024)
por: Khoshnoodi, Mahsa, et al.
Publicado: (2024)
Ejemplares similares
-
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024) -
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
por: Radwan, Ahmed Y., et al.
Publicado: (2026) -
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
por: Raza, Shaina, et al.
Publicado: (2024) -
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
por: Raza, Shaina, et al.
Publicado: (2025) -
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
por: Chaduvula, Sindhuja, et al.
Publicado: (2026)