Gespeichert in:
| Hauptverfasser: | Solaiman, Irene, Talat, Zeerak, Agnew, William, Ahmad, Lama, Baker, Dylan, Blodgett, Su Lin, Chen, Canyu, Daumé III, Hal, Dodge, Jesse, Duan, Isabella, Evans, Ellie, Friedrich, Felix, Ghosh, Avijit, Gohar, Usman, Hooker, Sara, Jernite, Yacine, Kalluri, Ria, Lusoli, Alberto, Leidinger, Alina, Lin, Michelle, Lin, Xiuzhu, Luccioni, Sasha, Mickel, Jennifer, Mitchell, Margaret, Newman, Jessica, Ovalle, Anaelia, Png, Marie-Therese, Singh, Shubham, Strait, Andrew, Struppek, Lukas, Subramonian, Arjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2306.05949 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
Understanding "Democratization" in NLP and ML Research
von: Subramonian, Arjun, et al.
Veröffentlicht: (2024)
von: Subramonian, Arjun, et al.
Veröffentlicht: (2024)
CIVICS: Building a Dataset for Examining Culturally-Informed Values in Large Language Models
von: Pistilli, Giada, et al.
Veröffentlicht: (2024)
von: Pistilli, Giada, et al.
Veröffentlicht: (2024)
Power Hungry Processing: Watts Driving the Cost of AI Deployment?
von: Luccioni, Alexandra Sasha, et al.
Veröffentlicht: (2023)
von: Luccioni, Alexandra Sasha, et al.
Veröffentlicht: (2023)
A Capabilities Approach to Studying Bias and Harm in Language Technologies
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
Exploitation All the Way Down: Calling out the Root Cause of Bad Online Experiences for Users of the "Majority World"
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
Online Learning Defense against Iterative Jailbreak Attacks via Prompt Optimization
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
Big AI's Regulatory Capture: Mapping Industry Interference and Government Complicity
von: Birhane, Abeba, et al.
Veröffentlicht: (2026)
von: Birhane, Abeba, et al.
Veröffentlicht: (2026)
A Necessary Step toward Faithfulness: Measuring and Improving Consistency in Free-Text Explanations
von: Zhao, Lingjun, et al.
Veröffentlicht: (2025)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2025)
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models
von: Nghiem, Huy, et al.
Veröffentlicht: (2024)
von: Nghiem, Huy, et al.
Veröffentlicht: (2024)
When Stereotypes GTG: The Impact of Predictive Text Suggestions on Gender Bias in Human-AI Co-Writing
von: Baumler, Connor, et al.
Veröffentlicht: (2024)
von: Baumler, Connor, et al.
Veröffentlicht: (2024)
Steering Safely or Off a Cliff? Rethinking Specificity and Robustness in Inference-Time Interventions
von: Goyal, Navita, et al.
Veröffentlicht: (2026)
von: Goyal, Navita, et al.
Veröffentlicht: (2026)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
Impoverished Language Technology: The Lack of (Social) Class in NLP
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
Beyond Release: Access Considerations for Generative AI Systems
von: Solaiman, Irene, et al.
Veröffentlicht: (2025)
von: Solaiman, Irene, et al.
Veröffentlicht: (2025)
Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research
von: Wong, Taryn, et al.
Veröffentlicht: (2026)
von: Wong, Taryn, et al.
Veröffentlicht: (2026)
Classist Tools: Social Class Correlates with Performance in NLP
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
von: Curry, Amanda Cercas, et al.
Veröffentlicht: (2024)
FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data
von: Abdelkadir, Nuredin Ali, et al.
Veröffentlicht: (2026)
von: Abdelkadir, Nuredin Ali, et al.
Veröffentlicht: (2026)
Survey of Bias In Text-to-Image Generation: Definition, Evaluation, and Mitigation
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models
von: Nghiem, Huy, et al.
Veröffentlicht: (2025)
von: Nghiem, Huy, et al.
Veröffentlicht: (2025)
Successfully Guiding Humans with Imperfect Instructions by Highlighting Potential Errors and Suggesting Corrections
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
Language Models Predict Empathy Gaps Between Social In-groups and Out-groups
von: Hou, Yu, et al.
Veröffentlicht: (2025)
von: Hou, Yu, et al.
Veröffentlicht: (2025)
Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
Who Gets Heard? Rethinking Fairness in AI for Music Systems
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
Dialect prejudice predicts AI decisions about people's character, employability, and criminality
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review
von: Agro, Maha Tufail, et al.
Veröffentlicht: (2025)
von: Agro, Maha Tufail, et al.
Veröffentlicht: (2025)
Dehumanizing Machines: Mitigating Anthropomorphic Behaviors in Text Generation Systems
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
Causal Effect of Group Diversity on Redundancy and Coverage in Peer-Reviewing
von: Goyal, Navita, et al.
Veröffentlicht: (2024)
von: Goyal, Navita, et al.
Veröffentlicht: (2024)
Pragmatics Meets Culture: Culturally-adapted Artwork Description Generation and Evaluation
von: Zhao, Lingjun, et al.
Veröffentlicht: (2026)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2026)
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations
von: Nghiem, Huy, et al.
Veröffentlicht: (2024)
von: Nghiem, Huy, et al.
Veröffentlicht: (2024)
The Impact of Explanations on Fairness in Human-AI Decision-Making: Protected vs Proxy Features
von: Goyal, Navita, et al.
Veröffentlicht: (2023)
von: Goyal, Navita, et al.
Veröffentlicht: (2023)
"One-Size-Fits-All"? Examining Expectations around What Constitute "Fair" or "Good" NLG System Behaviors
von: Lucy, Li, et al.
Veröffentlicht: (2023)
von: Lucy, Li, et al.
Veröffentlicht: (2023)
AI Automatons: AI Systems Intended to Imitate Humans
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA
von: Gor, Maharshi, et al.
Veröffentlicht: (2024)
von: Gor, Maharshi, et al.
Veröffentlicht: (2024)
Building Competence into Micro Systems.
von: Cheney, Hal
Veröffentlicht: (1986)
von: Cheney, Hal
Veröffentlicht: (1986)
PRISE: LLM-Style Sequence Compression for Learning Temporal Action Abstractions in Control
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
Non-Ground Congruence Closure
von: Leidinger, Hendrik, et al.
Veröffentlicht: (2024)
von: Leidinger, Hendrik, et al.
Veröffentlicht: (2024)
How Are LLMs Mitigating Stereotyping Harms? Learning from Search Engine Studies
von: Leidinger, Alina, et al.
Veröffentlicht: (2024)
von: Leidinger, Alina, et al.
Veröffentlicht: (2024)
Gefangenschaft, Revolution, Heimkehr
von: Moritz, Verena, et al.
Veröffentlicht: (2015)
von: Moritz, Verena, et al.
Veröffentlicht: (2015)
The Cake that is Intelligence and Who Gets to Bake it: An AI Analogy and its Implications for Participation
von: Mundt, Martin, et al.
Veröffentlicht: (2025)
von: Mundt, Martin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels
von: Fleisig, Eve, et al.
Veröffentlicht: (2024) -
Understanding "Democratization" in NLP and ML Research
von: Subramonian, Arjun, et al.
Veröffentlicht: (2024) -
CIVICS: Building a Dataset for Examining Culturally-Informed Values in Large Language Models
von: Pistilli, Giada, et al.
Veröffentlicht: (2024) -
Power Hungry Processing: Watts Driving the Cost of AI Deployment?
von: Luccioni, Alexandra Sasha, et al.
Veröffentlicht: (2023) -
A Capabilities Approach to Studying Bias and Harm in Language Technologies
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)