Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
Fuente:
arXiv
Salvato in:
| Autori principali: | Raza, Shaina, Bamgbose, Oluwanifemi, Ghuge, Shardul, Tavakol, Fatemeh, Reji, Deepak John, Bashir, Syed Raza |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
di: Raza, Shaina, et al.
Pubblicazione: (2023)
di: Raza, Shaina, et al.
Pubblicazione: (2023)
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
di: Raza, Shaina, et al.
Pubblicazione: (2023)
di: Raza, Shaina, et al.
Pubblicazione: (2023)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
di: Bashir, Syed Raza, et al.
Pubblicazione: (2024)
di: Bashir, Syed Raza, et al.
Pubblicazione: (2024)
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
di: Raza, Shaina
Pubblicazione: (2026)
di: Raza, Shaina
Pubblicazione: (2026)
Evaluating Robustness of Large Language Models in Enterprise Applications: Benchmarks for Perturbation Consistency Across Formats and Languages
di: Bogavelli, Tara, et al.
Pubblicazione: (2026)
di: Bogavelli, Tara, et al.
Pubblicazione: (2026)
FakeWatch: A Framework for Detecting Fake News to Ensure Credible Elections
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Optimizing Large Language Models: Metrics, Energy Efficiency, and Case Study Insights
di: Khan, Tahniat, et al.
Pubblicazione: (2025)
di: Khan, Tahniat, et al.
Pubblicazione: (2025)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
The Rise of Small Language Models in Healthcare: A Comprehensive Survey
di: Garg, Muskan, et al.
Pubblicazione: (2025)
di: Garg, Muskan, et al.
Pubblicazione: (2025)
BEADs: Bias Evaluation Across Domains
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
di: Narayanan, Aravind, et al.
Pubblicazione: (2025)
di: Narayanan, Aravind, et al.
Pubblicazione: (2025)
Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models
di: Khan, Sumra, et al.
Pubblicazione: (2026)
di: Khan, Sumra, et al.
Pubblicazione: (2026)
Vision-Language Models on the Edge for Real-Time Robotic Perception
di: Ahmad, Sarat, et al.
Pubblicazione: (2026)
di: Ahmad, Sarat, et al.
Pubblicazione: (2026)
Can We Edit Multimodal Large Language Models?
di: Cheng, Siyuan, et al.
Pubblicazione: (2023)
di: Cheng, Siyuan, et al.
Pubblicazione: (2023)
Amnesia: Adversarial Semantic Layer Specific Activation Steering in Large Language Models
di: Raza, Ali, et al.
Pubblicazione: (2026)
di: Raza, Ali, et al.
Pubblicazione: (2026)
Generalists vs. Specialists: Evaluating Large Language Models for Urdu
di: Arif, Samee, et al.
Pubblicazione: (2024)
di: Arif, Samee, et al.
Pubblicazione: (2024)
Impact of English Language in Engineering
di: Khalid Raza
Pubblicazione: (2019)
di: Khalid Raza
Pubblicazione: (2019)
How Can We Diagnose and Treat Bias in Large Language Models for Clinical Decision-Making?
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
Semantic Fusion with Fuzzy-Membership Features for Controllable Language Modelling
di: Huang, Yongchao, et al.
Pubblicazione: (2025)
di: Huang, Yongchao, et al.
Pubblicazione: (2025)
Can Large Language Models Understand Context?
di: Zhu, Yilun, et al.
Pubblicazione: (2024)
di: Zhu, Yilun, et al.
Pubblicazione: (2024)
Can Large Language Models Understand Molecules?
di: Sadeghi, Shaghayegh, et al.
Pubblicazione: (2024)
di: Sadeghi, Shaghayegh, et al.
Pubblicazione: (2024)
CosmoCLIP: Generalizing Large Vision-Language Models for Astronomical Imaging
di: Imam, Raza, et al.
Pubblicazione: (2024)
di: Imam, Raza, et al.
Pubblicazione: (2024)
Understanding Counting Mechanisms in Large Language and Vision-Language Models
di: Hasani, Hosein, et al.
Pubblicazione: (2025)
di: Hasani, Hosein, et al.
Pubblicazione: (2025)
Can Large Language Model Agents Balance Energy Systems?
di: Ren, Xinxing, et al.
Pubblicazione: (2025)
di: Ren, Xinxing, et al.
Pubblicazione: (2025)
Can Large Language Models Understand Spatial Audio?
di: Tang, Changli, et al.
Pubblicazione: (2024)
di: Tang, Changli, et al.
Pubblicazione: (2024)
On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?
di: Imam, Raza, et al.
Pubblicazione: (2025)
di: Imam, Raza, et al.
Pubblicazione: (2025)
LLM-Agent-based Social Simulation for Attitude Diffusion
di: Reji, Deepak John
Pubblicazione: (2026)
di: Reji, Deepak John
Pubblicazione: (2026)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories
di: Ho, Jessee, et al.
Pubblicazione: (2026)
di: Ho, Jessee, et al.
Pubblicazione: (2026)
Blacks is to Anger as Whites is to Joy? Understanding Latent Affective Bias in Large Pre-trained Neural Language Models
di: Kadan, Anoop, et al.
Pubblicazione: (2023)
di: Kadan, Anoop, et al.
Pubblicazione: (2023)
Can Large Language Models Develop Gambling Addiction?
di: Lee, Seungpil, et al.
Pubblicazione: (2025)
di: Lee, Seungpil, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
di: Raza, Shaina, et al.
Pubblicazione: (2023) -
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
di: Raza, Shaina, et al.
Pubblicazione: (2024) -
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
di: Raza, Shaina, et al.
Pubblicazione: (2023) -
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
di: Raza, Shaina, et al.
Pubblicazione: (2024) -
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)