FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
Fuente:
arXiv
Salvato in:
| Autori principali: | Raza, Shaina, Ghuge, Shardul, Ding, Chen, Dolatabadi, Elham, Pandya, Deval |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
di: Raza, Shaina, et al.
Pubblicazione: (2023)
di: Raza, Shaina, et al.
Pubblicazione: (2023)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
di: Raza, Shaina, et al.
Pubblicazione: (2023)
di: Raza, Shaina, et al.
Pubblicazione: (2023)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Just as Humans Need Vaccines, So Do Models: Model Immunization to Combat Falsehoods
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
Academic case reports lack diversity: Assessing the presence and diversity of sociodemographic and behavioral factors related to Post COVID-19 Condition
di: Florez, Juan Andres Medina, et al.
Pubblicazione: (2025)
di: Florez, Juan Andres Medina, et al.
Pubblicazione: (2025)
Towards Enabling FAIR Dataspaces Using Large Language Models
di: Arnold, Benedikt T., et al.
Pubblicazione: (2024)
di: Arnold, Benedikt T., et al.
Pubblicazione: (2024)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Making Metadata More FAIR Using Large Language Models
di: Sundaram, Sowmya S., et al.
Pubblicazione: (2023)
di: Sundaram, Sowmya S., et al.
Pubblicazione: (2023)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
A FAIR and Free Prompt-based Research Assistant
di: Shamsabadi, Mahsa, et al.
Pubblicazione: (2024)
di: Shamsabadi, Mahsa, et al.
Pubblicazione: (2024)
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge
di: Laskar, Md Tahmid Rahman, et al.
Pubblicazione: (2025)
di: Laskar, Md Tahmid Rahman, et al.
Pubblicazione: (2025)
How Far Can We Extract Diverse Perspectives from Large Language Models?
di: Hayati, Shirley Anugrah, et al.
Pubblicazione: (2023)
di: Hayati, Shirley Anugrah, et al.
Pubblicazione: (2023)
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
di: Bashir, Syed Raza, et al.
Pubblicazione: (2024)
di: Bashir, Syed Raza, et al.
Pubblicazione: (2024)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
AutoFAIR : Automatic Data FAIRification via Machine Reading
di: Ma, Tingyan, et al.
Pubblicazione: (2024)
di: Ma, Tingyan, et al.
Pubblicazione: (2024)
Soft-prompt Tuning for Large Language Models to Evaluate Bias
di: Tian, Jacob-Junqi, et al.
Pubblicazione: (2023)
di: Tian, Jacob-Junqi, et al.
Pubblicazione: (2023)
The Rise of Small Language Models in Healthcare: A Comprehensive Survey
di: Garg, Muskan, et al.
Pubblicazione: (2025)
di: Garg, Muskan, et al.
Pubblicazione: (2025)
Beyond Content: How Grammatical Gender Shapes Visual Representation in Text-to-Image Models
di: Saeed, Muhammed, et al.
Pubblicazione: (2025)
di: Saeed, Muhammed, et al.
Pubblicazione: (2025)
Exploring the Influence of Label Aggregation on Minority Voices: Implications for Dataset Bias and Model Training
di: Pandya, Mugdha, et al.
Pubblicazione: (2024)
di: Pandya, Mugdha, et al.
Pubblicazione: (2024)
Can We Edit Multimodal Large Language Models?
di: Cheng, Siyuan, et al.
Pubblicazione: (2023)
di: Cheng, Siyuan, et al.
Pubblicazione: (2023)
Position: Beyond Assistance -- Reimagining LLMs as Ethical and Adaptive Co-Creators in Mental Health Care
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
SciCom Wiki: Fact-Checking and FAIR Knowledge Distribution for Scientific Videos and Podcasts
di: Wittenborg, Tim, et al.
Pubblicazione: (2025)
di: Wittenborg, Tim, et al.
Pubblicazione: (2025)
BEADs: Bias Evaluation Across Domains
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Vulnerability Detection with Code Language Models: How Far Are We?
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
di: Raza, Shaina
Pubblicazione: (2026)
di: Raza, Shaina
Pubblicazione: (2026)
Round Trip Translation Defence against Large Language Model Jailbreaking Attacks
di: Yung, Canaan, et al.
Pubblicazione: (2024)
di: Yung, Canaan, et al.
Pubblicazione: (2024)
How Much Data is Enough Data? Fine-Tuning Large Language Models for In-House Translation: Performance Evaluation Across Multiple Dataset Sizes
di: Vieira, Inacio, et al.
Pubblicazione: (2024)
di: Vieira, Inacio, et al.
Pubblicazione: (2024)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
di: Raval, Ananya, et al.
Pubblicazione: (2025)
di: Raval, Ananya, et al.
Pubblicazione: (2025)
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
di: Salimian, Sina, et al.
Pubblicazione: (2025)
di: Salimian, Sina, et al.
Pubblicazione: (2025)
Can We Predict Performance of Large Models across Vision-Language Tasks?
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
How Much is Enough? The Diminishing Returns of Tokenization Training Data
di: Reddy, Varshini, et al.
Pubblicazione: (2025)
di: Reddy, Varshini, et al.
Pubblicazione: (2025)
Can We Use Large Language Models to Fill Relevance Judgment Holes?
di: Abbasiantaeb, Zahra, et al.
Pubblicazione: (2024)
di: Abbasiantaeb, Zahra, et al.
Pubblicazione: (2024)
HumaniBench: A Human-Centric Framework for Large Multimodal Models Evaluation
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
di: Raza, Shaina, et al.
Pubblicazione: (2024) -
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
di: Raza, Shaina, et al.
Pubblicazione: (2023) -
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024) -
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
di: Raza, Shaina, et al.
Pubblicazione: (2023) -
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
di: Raza, Shaina, et al.
Pubblicazione: (2024)