Wisdom of Instruction-Tuned Language Model Crowds. Exploring Model Label Variation
Fuente:
arXiv
Salvato in:
| Autori principali: | Plaza-del-Arco, Flor Miriam, Nozza, Debora, Hovy, Dirk |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring Subjective Tasks in Farsi: A Survey Analysis and Evaluation of Language Models
di: Rooein, Donya, et al.
Pubblicazione: (2025)
di: Rooein, Donya, et al.
Pubblicazione: (2025)
Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion Attribution
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
The Pluralistic Moral Gap: Understanding Judgment and Value Differences between Humans and Large Language Models
di: Russo, Giuseppe, et al.
Pubblicazione: (2025)
di: Russo, Giuseppe, et al.
Pubblicazione: (2025)
No for Some, Yes for Others: Persona Prompts and Other Sources of False Refusal in Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2025)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2025)
Emotion Analysis in NLP: Trends, Gaps and Roadmap for Future Directions
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
Biased Tales: Cultural and Topic Bias in Generating Children's Stories
di: Rooein, Donya, et al.
Pubblicazione: (2025)
di: Rooein, Donya, et al.
Pubblicazione: (2025)
Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
di: Baumann, Joachim, et al.
Pubblicazione: (2025)
di: Baumann, Joachim, et al.
Pubblicazione: (2025)
Language Model Council: Democratically Benchmarking Foundation Models on Highly Subjective Tasks
di: Zhao, Justin, et al.
Pubblicazione: (2024)
di: Zhao, Justin, et al.
Pubblicazione: (2024)
PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
di: Rooein, Donya, et al.
Pubblicazione: (2026)
di: Rooein, Donya, et al.
Pubblicazione: (2026)
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
SINAI at eRisk@CLEF 2023: Approaching Early Detection of Gambling with Natural Language Processing
di: Marmol-Romero, Alba Maria, et al.
Pubblicazione: (2025)
di: Marmol-Romero, Alba Maria, et al.
Pubblicazione: (2025)
Do Large Language Models Adapt to Language Variation across Socioeconomic Status?
di: Bassignana, Elisa, et al.
Pubblicazione: (2026)
di: Bassignana, Elisa, et al.
Pubblicazione: (2026)
Consistency is Key: Disentangling Label Variation in Natural Language Processing with Intra-Annotator Agreement
di: Abercrombie, Gavin, et al.
Pubblicazione: (2023)
di: Abercrombie, Gavin, et al.
Pubblicazione: (2023)
FLANS at SemEval-2026 Task 7: RAG with Open-Sourced Smaller LLMs for Everyday Knowledge Across Diverse Languages and Cultures
di: Bogdanova, Liliia, et al.
Pubblicazione: (2026)
di: Bogdanova, Liliia, et al.
Pubblicazione: (2026)
FairBelief -- Assessing Harmful Beliefs in Language Models
di: Setzu, Mattia, et al.
Pubblicazione: (2024)
di: Setzu, Mattia, et al.
Pubblicazione: (2024)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
From Chatbots to Confidants: A Cross-Cultural Study of LLM Adoption for Emotional Support
di: Amat-Lefort, Natalia, et al.
Pubblicazione: (2026)
di: Amat-Lefort, Natalia, et al.
Pubblicazione: (2026)
Compromesso! Italian Many-Shot Jailbreaks Undermine the Safety of Large Language Models
di: Pernisi, Fabio, et al.
Pubblicazione: (2024)
di: Pernisi, Fabio, et al.
Pubblicazione: (2024)
Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
di: Liu, Shang, et al.
Pubblicazione: (2024)
di: Liu, Shang, et al.
Pubblicazione: (2024)
CrowdSelect: Synthetic Instruction Data Selection with Multi-LLM Wisdom
di: Li, Yisen, et al.
Pubblicazione: (2025)
di: Li, Yisen, et al.
Pubblicazione: (2025)
Measuring Pragmatic Influence in Large Language Model Instructions
di: Geng, Yilin, et al.
Pubblicazione: (2026)
di: Geng, Yilin, et al.
Pubblicazione: (2026)
WisdomBot: Tuning Large Language Models with Artificial Intelligence Knowledge
di: Chen, Jingyuan, et al.
Pubblicazione: (2025)
di: Chen, Jingyuan, et al.
Pubblicazione: (2025)
MSTS: A Multimodal Safety Test Suite for Vision-Language Models
di: Röttger, Paul, et al.
Pubblicazione: (2025)
di: Röttger, Paul, et al.
Pubblicazione: (2025)
COPOS: Corpus Of Patient Opinions in Spanish. Application of Sentiment Analysis Techniques
di: Flor Miriam Plaza-del-Arco
Pubblicazione: (2016)
di: Flor Miriam Plaza-del-Arco
Pubblicazione: (2016)
Conversations as a Source for Teaching Scientific Concepts at Different Education Levels
di: Rooein, Donya, et al.
Pubblicazione: (2024)
di: Rooein, Donya, et al.
Pubblicazione: (2024)
Narratives at Conflict: Computational Analysis of News Framing in Multilingual Disinformation Campaigns
di: Sinelnik, Antonina, et al.
Pubblicazione: (2024)
di: Sinelnik, Antonina, et al.
Pubblicazione: (2024)
SafetyPrompts: a Systematic Review of Open Datasets for Evaluating and Improving Large Language Model Safety
di: Röttger, Paul, et al.
Pubblicazione: (2024)
di: Röttger, Paul, et al.
Pubblicazione: (2024)
The AI Gap: How Socioeconomic Status Affects Language Technology Interactions
di: Bassignana, Elisa, et al.
Pubblicazione: (2025)
di: Bassignana, Elisa, et al.
Pubblicazione: (2025)
Twists, Humps, and Pebbles: Multilingual Speech Recognition Models Exhibit Gender Performance Gaps
di: Attanasio, Giuseppe, et al.
Pubblicazione: (2024)
di: Attanasio, Giuseppe, et al.
Pubblicazione: (2024)
Wisdom of the Silicon Crowd: LLM Ensemble Prediction Capabilities Rival Human Crowd Accuracy
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
The Call for Socially Aware Language Technologies
di: Yang, Diyi, et al.
Pubblicazione: (2024)
di: Yang, Diyi, et al.
Pubblicazione: (2024)
What Is The Political Content in LLMs' Pre- and Post-Training Data?
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
SINAI at eRisk@CLEF 2022: Approaching Early Detection of Gambling and Eating Disorders with Natural Language Processing
di: Marmol-Romero, Alba Maria, et al.
Pubblicazione: (2025)
di: Marmol-Romero, Alba Maria, et al.
Pubblicazione: (2025)
The Wisdom of Partisan Crowds: Comparing Collective Intelligence in Humans and LLM-based Agents
di: Chuang, Yun-Shiuan, et al.
Pubblicazione: (2023)
di: Chuang, Yun-Shiuan, et al.
Pubblicazione: (2023)
Diffusion Language Models Are Natively Length-Aware
di: Rossi, Vittorio, et al.
Pubblicazione: (2026)
di: Rossi, Vittorio, et al.
Pubblicazione: (2026)
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
di: Ngueajio, Mikel K., et al.
Pubblicazione: (2025)
di: Ngueajio, Mikel K., et al.
Pubblicazione: (2025)
Impoverished Language Technology: The Lack of (Social) Class in NLP
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
Chain-of-Instructions: Compositional Instruction Tuning on Large Language Models
di: Hayati, Shirley Anugrah, et al.
Pubblicazione: (2024)
di: Hayati, Shirley Anugrah, et al.
Pubblicazione: (2024)
Control Illusion: The Failure of Instruction Hierarchies in Large Language Models
di: Geng, Yilin, et al.
Pubblicazione: (2025)
di: Geng, Yilin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Exploring Subjective Tasks in Farsi: A Survey Analysis and Evaluation of Language Models
di: Rooein, Donya, et al.
Pubblicazione: (2025) -
Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion Attribution
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024) -
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024) -
The Pluralistic Moral Gap: Understanding Judgment and Value Differences between Humans and Large Language Models
di: Russo, Giuseppe, et al.
Pubblicazione: (2025) -
No for Some, Yes for Others: Persona Prompts and Other Sources of False Refusal in Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2025)