Understanding the Interplay of Scale, Data, and Bias in Language Models: A Case Study with BERT
Fuente:
arXiv
Guardado en:
| Autores principales: | Ali, Muhammad, Panda, Swetasudha, Shen, Qinlan, Wick, Michael, Kobren, Ari |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
por: Nghiem, Huy, et al.
Publicado: (2025)
por: Nghiem, Huy, et al.
Publicado: (2025)
Adaptive Question Answering: Enhancing Language Model Proficiency for Addressing Knowledge Conflicts with Source Citations
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models
por: Wang, Xinyi, et al.
Publicado: (2025)
por: Wang, Xinyi, et al.
Publicado: (2025)
AccessEval: Benchmarking Disability Bias in Large Language Models
por: Panda, Srikant, et al.
Publicado: (2025)
por: Panda, Srikant, et al.
Publicado: (2025)
Data Engineering for Scaling Language Models to 128K Context
por: Fu, Yao, et al.
Publicado: (2024)
por: Fu, Yao, et al.
Publicado: (2024)
EuroBERT: Scaling Multilingual Encoders for European Languages
por: Boizard, Nicolas, et al.
Publicado: (2025)
por: Boizard, Nicolas, et al.
Publicado: (2025)
Construction Identification and Disambiguation Using BERT: A Case Study of NPN
por: Scivetti, Wesley, et al.
Publicado: (2025)
por: Scivetti, Wesley, et al.
Publicado: (2025)
Understanding or Memorizing? A Case Study of German Definite Articles in Language Models
por: Drechsel, Jonathan, et al.
Publicado: (2026)
por: Drechsel, Jonathan, et al.
Publicado: (2026)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
por: Kumar, Divyanshu, et al.
Publicado: (2024)
por: Kumar, Divyanshu, et al.
Publicado: (2024)
Relational Schemata in BERT Are Inducible, Not Emergent: A Study of Performance vs. Competence in Language Models
por: Gawin, Cole
Publicado: (2025)
por: Gawin, Cole
Publicado: (2025)
Enhancing Aspect-based Sentiment Analysis with ParsBERT in Persian Language
por: Ariai, Farid, et al.
Publicado: (2025)
por: Ariai, Farid, et al.
Publicado: (2025)
RooseBERT: A New Deal For Political Language Modelling
por: Dore, Deborah, et al.
Publicado: (2025)
por: Dore, Deborah, et al.
Publicado: (2025)
Understanding and Mitigating Tokenization Bias in Language Models
por: Phan, Buu, et al.
Publicado: (2024)
por: Phan, Buu, et al.
Publicado: (2024)
No LLM is Free From Bias: A Comprehensive Study of Bias Evaluation in Large Language Models
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
Linearly Decoding Refused Knowledge in Aligned Language Models
por: Shrivastava, Aryan, et al.
Publicado: (2025)
por: Shrivastava, Aryan, et al.
Publicado: (2025)
Utilizing Large Language Models to Generate Synthetic Data to Increase the Performance of BERT-Based Neural Networks
por: Woolsey, Chancellor R., et al.
Publicado: (2024)
por: Woolsey, Chancellor R., et al.
Publicado: (2024)
NeoBERT: A Next-Generation BERT
por: Breton, Lola Le, et al.
Publicado: (2025)
por: Breton, Lola Le, et al.
Publicado: (2025)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
KliniskVestBERT: BERT Model Specialised to Norwegian Clinical Texts
por: Autenried, Christian, et al.
Publicado: (2026)
por: Autenried, Christian, et al.
Publicado: (2026)
AraPoemBERT: A Pretrained Language Model for Arabic Poetry Analysis
por: Qarah, Faisal
Publicado: (2024)
por: Qarah, Faisal
Publicado: (2024)
SaudiBERT: A Large Language Model Pretrained on Saudi Dialect Corpora
por: Qarah, Faisal
Publicado: (2024)
por: Qarah, Faisal
Publicado: (2024)
Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT
por: Banerjee, Awritrojit, et al.
Publicado: (2025)
por: Banerjee, Awritrojit, et al.
Publicado: (2025)
Scaling Stick-Breaking Attention: An Efficient Implementation and In-depth Study
por: Tan, Shawn, et al.
Publicado: (2024)
por: Tan, Shawn, et al.
Publicado: (2024)
Optimal Turkish Subword Strategies at Scale: Systematic Evaluation of Data, Vocabulary, Morphology Interplay
por: Altinok, Duygu
Publicado: (2026)
por: Altinok, Duygu
Publicado: (2026)
BPDec: Unveiling the Potential of Masked Language Modeling Decoder in BERT pretraining
por: Liang, Wen, et al.
Publicado: (2024)
por: Liang, Wen, et al.
Publicado: (2024)
KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
por: Awlla, Kozhin muhealddin, et al.
Publicado: (2025)
por: Awlla, Kozhin muhealddin, et al.
Publicado: (2025)
BI-RADS BERT & Using Section Segmentation to Understand Radiology Reports
por: Kuling, Grey, et al.
Publicado: (2021)
por: Kuling, Grey, et al.
Publicado: (2021)
Interplay of Machine Translation, Diacritics, and Diacritization
por: Chen, Wei-Rui, et al.
Publicado: (2024)
por: Chen, Wei-Rui, et al.
Publicado: (2024)
UrduLM: A Resource-Efficient Monolingual Urdu Language Model
por: Ali, Syed Muhammad, et al.
Publicado: (2026)
por: Ali, Syed Muhammad, et al.
Publicado: (2026)
Patent Language Model Pretraining with ModernBERT
por: Yousefiramandi, Amirhossein, et al.
Publicado: (2025)
por: Yousefiramandi, Amirhossein, et al.
Publicado: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
Post-hoc Reward Calibration: A Case Study on Length Bias
por: Huang, Zeyu, et al.
Publicado: (2024)
por: Huang, Zeyu, et al.
Publicado: (2024)
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
por: Saoud, Abdulkader, et al.
Publicado: (2024)
por: Saoud, Abdulkader, et al.
Publicado: (2024)
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models
por: Liang, Chaoqi, et al.
Publicado: (2023)
por: Liang, Chaoqi, et al.
Publicado: (2023)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
por: Parihar, Shweta, et al.
Publicado: (2026)
por: Parihar, Shweta, et al.
Publicado: (2026)
Scaling Laws of Synthetic Data for Language Models
por: Qin, Zeyu, et al.
Publicado: (2025)
por: Qin, Zeyu, et al.
Publicado: (2025)
Data Generation Using Large Language Models for Text Classification: An Empirical Case Study
por: Li, Yinheng, et al.
Publicado: (2024)
por: Li, Yinheng, et al.
Publicado: (2024)
PhayaThaiBERT: Enhancing a Pretrained Thai Language Model with Unassimilated Loanwords
por: Sriwirote, Panyut, et al.
Publicado: (2023)
por: Sriwirote, Panyut, et al.
Publicado: (2023)
The Reasoning-Memorization Interplay in Language Models Is Mediated by a Single Direction
por: Hong, Yihuai, et al.
Publicado: (2025)
por: Hong, Yihuai, et al.
Publicado: (2025)
Research Trends for the Interplay between Large Language Models and Knowledge Graphs
por: Khorashadizadeh, Hanieh, et al.
Publicado: (2024)
por: Khorashadizadeh, Hanieh, et al.
Publicado: (2024)
Ejemplares similares
-
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
por: Nghiem, Huy, et al.
Publicado: (2025) -
Adaptive Question Answering: Enhancing Language Model Proficiency for Addressing Knowledge Conflicts with Source Citations
por: Shaier, Sagi, et al.
Publicado: (2024) -
Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models
por: Wang, Xinyi, et al.
Publicado: (2025) -
AccessEval: Benchmarking Disability Bias in Large Language Models
por: Panda, Srikant, et al.
Publicado: (2025) -
Data Engineering for Scaling Language Models to 128K Context
por: Fu, Yao, et al.
Publicado: (2024)