EtiCor++: Towards Understanding Etiquettical Bias in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dwivedi, Ashutosh, Singh, Siddhant Shivdutt, Modi, Ashutosh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Measuring and Modeling "Culture" in LLMs: A Survey
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
LoRMA: Low-Rank Multiplicative Adaptation for LLMs
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
Towards Quantifying Commonsense Reasoning with Mechanistic Insights
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models
von: Raj, Ashutosh
Veröffentlicht: (2026)
von: Raj, Ashutosh
Veröffentlicht: (2026)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
Geometry of Decision Making in Language Models
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
IITK at SemEval-2024 Task 4: Hierarchical Embeddings for Detection of Persuasion Techniques in Memes
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
COLD: Causal reasOning in cLosed Daily activities
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
POSESTITCH-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
von: Ko, Changgeon, et al.
Veröffentlicht: (2024)
von: Ko, Changgeon, et al.
Veröffentlicht: (2024)
Perceived Political Bias in LLMs Reduces Persuasive Abilities
von: DiGiuseppe, Matthew, et al.
Veröffentlicht: (2026)
von: DiGiuseppe, Matthew, et al.
Veröffentlicht: (2026)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
Generation and De-Identification of Indian Clinical Discharge Summaries using LLMs
von: Singh, Sanjeet, et al.
Veröffentlicht: (2024)
von: Singh, Sanjeet, et al.
Veröffentlicht: (2024)
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs
von: Wachter, Jasmin, et al.
Veröffentlicht: (2025)
von: Wachter, Jasmin, et al.
Veröffentlicht: (2025)
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
IITK at SemEval-2024 Task 1: Contrastive Learning and Autoencoders for Semantic Textual Relatedness in Multilingual Texts
von: Basak, Udvas, et al.
Veröffentlicht: (2024)
von: Basak, Udvas, et al.
Veröffentlicht: (2024)
Gender Bias in LLMs: Preliminary Evidence from Shared Parenting Scenario in Czech Family Law
von: Harasta, Jakub, et al.
Veröffentlicht: (2026)
von: Harasta, Jakub, et al.
Veröffentlicht: (2026)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
von: Shin, Jisu, et al.
Veröffentlicht: (2024)
von: Shin, Jisu, et al.
Veröffentlicht: (2024)
CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models
von: Sun, Guangzhi, et al.
Veröffentlicht: (2025)
von: Sun, Guangzhi, et al.
Veröffentlicht: (2025)
The Homogenization Problem in LLMs: Towards Meaningful Diversity in AI Safety
von: Rios-Sialer, Ian
Veröffentlicht: (2026)
von: Rios-Sialer, Ian
Veröffentlicht: (2026)
The Ethics of Interaction: Mitigating Security Threats in LLMs
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
Understanding Mental Health Content on Social Media and Its Effect Towards Suicidal Ideation
von: Bhuiyan, Mohaiminul Islam, et al.
Veröffentlicht: (2025)
von: Bhuiyan, Mohaiminul Islam, et al.
Veröffentlicht: (2025)
Debating for Better Reasoning: An Unsupervised Multimodal Approach
von: Adhikari, Ashutosh, et al.
Veröffentlicht: (2025)
von: Adhikari, Ashutosh, et al.
Veröffentlicht: (2025)
No Free Lunch in Language Model Bias Mitigation? Targeted Bias Reduction Can Exacerbate Unmitigated LLM Biases
von: Chand, Shireen, et al.
Veröffentlicht: (2025)
von: Chand, Shireen, et al.
Veröffentlicht: (2025)
Social Bias in Popular Question-Answering Benchmarks
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
Cross-Language Bias Examination in Large Language Models
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
Which English Do LLMs Prefer? Triangulating Structural Bias Towards American English in Foundation Models
von: Nayeem, Mir Tafseer, et al.
Veröffentlicht: (2026)
von: Nayeem, Mir Tafseer, et al.
Veröffentlicht: (2026)
BookSQL: A Large Scale Text-to-SQL Dataset for Accounting Domain
von: Kumar, Rahul, et al.
Veröffentlicht: (2024)
von: Kumar, Rahul, et al.
Veröffentlicht: (2024)
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
The Lifecycle of "Facts": A Survey of Social Bias in Knowledge Graphs
von: Kraft, Angelie, et al.
Veröffentlicht: (2022)
von: Kraft, Angelie, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Towards Measuring and Modeling "Culture" in LLMs: A Survey
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024) -
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024) -
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025) -
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024) -
LoRMA: Low-Rank Multiplicative Adaptation for LLMs
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)