The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels
Fuente:
arXiv
Salvato in:
| Autori principali: | Fleisig, Eve, Blodgett, Su Lin, Klein, Dan, Talat, Zeerak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Capabilities Approach to Studying Bias and Harm in Language Technologies
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
di: Fleisig, Eve, et al.
Pubblicazione: (2024)
di: Fleisig, Eve, et al.
Pubblicazione: (2024)
Understanding "Democratization" in NLP and ML Research
di: Subramonian, Arjun, et al.
Pubblicazione: (2024)
di: Subramonian, Arjun, et al.
Pubblicazione: (2024)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
Impoverished Language Technology: The Lack of (Social) Class in NLP
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data
di: Abdelkadir, Nuredin Ali, et al.
Pubblicazione: (2026)
di: Abdelkadir, Nuredin Ali, et al.
Pubblicazione: (2026)
Balancing Quality and Variation: Spam Filtering Distorts Data Label Distributions
di: Fleisig, Eve, et al.
Pubblicazione: (2025)
di: Fleisig, Eve, et al.
Pubblicazione: (2025)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
di: Fleisig, Eve, et al.
Pubblicazione: (2023)
di: Fleisig, Eve, et al.
Pubblicazione: (2023)
MLLM-as-a-Judge for Image Safety without Human Labeling
di: Wang, Zhenting, et al.
Pubblicazione: (2024)
di: Wang, Zhenting, et al.
Pubblicazione: (2024)
Exploitation All the Way Down: Calling out the Root Cause of Bad Online Experiences for Users of the "Majority World"
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
di: Li, Jing-Jing, et al.
Pubblicazione: (2026)
di: Li, Jing-Jing, et al.
Pubblicazione: (2026)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
di: Borchers, Conrad, et al.
Pubblicazione: (2025)
di: Borchers, Conrad, et al.
Pubblicazione: (2025)
Protected group bias and stereotypes in Large Language Models
di: Kotek, Hadas, et al.
Pubblicazione: (2024)
di: Kotek, Hadas, et al.
Pubblicazione: (2024)
Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communities
di: Jarvis, Devon, et al.
Pubblicazione: (2026)
di: Jarvis, Devon, et al.
Pubblicazione: (2026)
Questionnaire Responses Do not Capture the Safety of AI Agents
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2026)
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2026)
Ghostbuster: Detecting Text Ghostwritten by Large Language Models
di: Verma, Vivek, et al.
Pubblicazione: (2023)
di: Verma, Vivek, et al.
Pubblicazione: (2023)
LLM Analysis of 150+ years of German Parliamentary Debates on Migration Reveals Shift from Post-War Solidarity to Anti-Solidarity in the Last Decade
di: Kostikova, Aida, et al.
Pubblicazione: (2025)
di: Kostikova, Aida, et al.
Pubblicazione: (2025)
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination
di: Li, Zihao
Pubblicazione: (2023)
di: Li, Zihao
Pubblicazione: (2023)
Operationalizing Automated Essay Scoring: A Human-Aware Approach
di: Plasencia-Calaña, Yenisel
Pubblicazione: (2025)
di: Plasencia-Calaña, Yenisel
Pubblicazione: (2025)
Mapping Social Choice Theory to RLHF
di: Dai, Jessica, et al.
Pubblicazione: (2024)
di: Dai, Jessica, et al.
Pubblicazione: (2024)
Hypothesis Testing for Quantifying LLM-Human Misalignment in Multiple Choice Settings
di: Hong, Harbin, et al.
Pubblicazione: (2025)
di: Hong, Harbin, et al.
Pubblicazione: (2025)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
di: Yang, Xilin
Pubblicazione: (2024)
di: Yang, Xilin
Pubblicazione: (2024)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
di: Weerasooriya, Tharindu Cyril, et al.
Pubblicazione: (2023)
di: Weerasooriya, Tharindu Cyril, et al.
Pubblicazione: (2023)
Challenging Assumptions in Learning Generic Text Style Embeddings
di: Ostheimer, Phil, et al.
Pubblicazione: (2025)
di: Ostheimer, Phil, et al.
Pubblicazione: (2025)
KPoEM: A Human-Annotated Dataset for Emotion Classification and RAG-Based Poetry Generation in Korean Modern Poetry
di: Lim, Iro, et al.
Pubblicazione: (2025)
di: Lim, Iro, et al.
Pubblicazione: (2025)
Simulated Adoption: Decoupling Magnitude and Direction in LLM In-Context Conflict Resolution
di: Zhang, Long, et al.
Pubblicazione: (2026)
di: Zhang, Long, et al.
Pubblicazione: (2026)
Ethics Whitepaper: Whitepaper on Ethical Research into Large Language Models
di: Ungless, Eddie L., et al.
Pubblicazione: (2024)
di: Ungless, Eddie L., et al.
Pubblicazione: (2024)
The Geometric Price of Discrete Logic: Context-driven Manifold Dynamics of Number Representations
di: Zhang, Long, et al.
Pubblicazione: (2026)
di: Zhang, Long, et al.
Pubblicazione: (2026)
Online Learning Defense against Iterative Jailbreak Attacks via Prompt Optimization
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
Big AI's Regulatory Capture: Mapping Industry Interference and Government Complicity
di: Birhane, Abeba, et al.
Pubblicazione: (2026)
di: Birhane, Abeba, et al.
Pubblicazione: (2026)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
Foundational Challenges in Assuring Alignment and Safety of Large Language Models
di: Anwar, Usman, et al.
Pubblicazione: (2024)
di: Anwar, Usman, et al.
Pubblicazione: (2024)
Embracing Imperfection: Simulating Students with Diverse Cognitive Levels Using LLM-based Agents
di: Wu, Tao, et al.
Pubblicazione: (2025)
di: Wu, Tao, et al.
Pubblicazione: (2025)
"I Am the One and Only, Your Cyber BFF": Understanding the Impact of GenAI Requires Understanding the Impact of Anthropomorphic AI
di: Cheng, Myra, et al.
Pubblicazione: (2024)
di: Cheng, Myra, et al.
Pubblicazione: (2024)
Understanding and Meeting Practitioner Needs When Measuring Representational Harms Caused by LLM-Based Systems
di: Harvey, Emma, et al.
Pubblicazione: (2025)
di: Harvey, Emma, et al.
Pubblicazione: (2025)
ELMES: An Automated Framework for Evaluating Large Language Models in Educational Scenarios
di: Wei, Shou'ang, et al.
Pubblicazione: (2025)
di: Wei, Shou'ang, et al.
Pubblicazione: (2025)
LLM Defenses Are Not Robust to Multi-Turn Human Jailbreaks Yet
di: Li, Nathaniel, et al.
Pubblicazione: (2024)
di: Li, Nathaniel, et al.
Pubblicazione: (2024)
PsychoGAT: A Novel Psychological Measurement Paradigm through Interactive Fiction Games with LLM Agents
di: Yang, Qisen, et al.
Pubblicazione: (2024)
di: Yang, Qisen, et al.
Pubblicazione: (2024)
AI-Augmented Predictions: LLM Assistants Improve Human Forecasting Accuracy
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
di: Zhou, Han, et al.
Pubblicazione: (2024)
di: Zhou, Han, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Capabilities Approach to Studying Bias and Harm in Language Technologies
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024) -
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
di: Fleisig, Eve, et al.
Pubblicazione: (2024) -
Understanding "Democratization" in NLP and ML Research
di: Subramonian, Arjun, et al.
Pubblicazione: (2024) -
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024) -
Impoverished Language Technology: The Lack of (Social) Class in NLP
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)