Sensitivity, Performance, Robustness: Deconstructing the Effect of Sociodemographic Prompting
Fuente:
arXiv
Guardado en:
| Autores principales: | Beck, Tilman, Schuff, Hendrik, Lauscher, Anne, Gurevych, Iryna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
How are Prompts Different in Terms of Sensitivity?
por: Lu, Sheng, et al.
Publicado: (2023)
por: Lu, Sheng, et al.
Publicado: (2023)
LLM Roleplay: Simulating Human-Chatbot Interaction
por: Tamoyan, Hovhannes, et al.
Publicado: (2024)
por: Tamoyan, Hovhannes, et al.
Publicado: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
por: Paul, Indraneil, et al.
Publicado: (2024)
por: Paul, Indraneil, et al.
Publicado: (2024)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
por: Baumgärtner, Tim, et al.
Publicado: (2026)
por: Baumgärtner, Tim, et al.
Publicado: (2026)
Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling
por: Tamoyan, Hovhannes, et al.
Publicado: (2025)
por: Tamoyan, Hovhannes, et al.
Publicado: (2025)
Preemptive Detection and Correction of Misaligned Actions in LLM Agents
por: Fang, Haishuo, et al.
Publicado: (2024)
por: Fang, Haishuo, et al.
Publicado: (2024)
Towards Privacy-aware Mental Health AI Models: Advances, Challenges, and Opportunities
por: Mandal, Aishik, et al.
Publicado: (2025)
por: Mandal, Aishik, et al.
Publicado: (2025)
MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions
por: Mandal, Aishik, et al.
Publicado: (2025)
por: Mandal, Aishik, et al.
Publicado: (2025)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
por: Waldis, Andreas, et al.
Publicado: (2024)
por: Waldis, Andreas, et al.
Publicado: (2024)
Decision-Making with Deliberation: Meta-reviewing as a Document-grounded Dialogue
por: Purkayastha, Sukannya, et al.
Publicado: (2025)
por: Purkayastha, Sukannya, et al.
Publicado: (2025)
SoS: Analysis of Surface over Semantics in Multilingual Text-To-Image Generation
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
Multimodal Large Language Models to Support Real-World Fact-Checking
por: Geng, Jiahui, et al.
Publicado: (2024)
por: Geng, Jiahui, et al.
Publicado: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
por: Baumgärtner, Tim, et al.
Publicado: (2025)
por: Baumgärtner, Tim, et al.
Publicado: (2025)
Thought Flow Nets: From Single Predictions to Trains of Model Thought
por: Schuff, Hendrik, et al.
Publicado: (2021)
por: Schuff, Hendrik, et al.
Publicado: (2021)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
por: Vida, Karina, et al.
Publicado: (2024)
por: Vida, Karina, et al.
Publicado: (2024)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
por: Sadallah, Abdelrahman, et al.
Publicado: (2025)
por: Sadallah, Abdelrahman, et al.
Publicado: (2025)
DOCE: Finding the Sweet Spot for Execution-Based Code Generation
por: Li, Haau-Sing, et al.
Publicado: (2024)
por: Li, Haau-Sing, et al.
Publicado: (2024)
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
por: Puerto, Haritz, et al.
Publicado: (2026)
por: Puerto, Haritz, et al.
Publicado: (2026)
Hypothesis-Driven Feature Manifold Analysis in LLMs via Supervised Multi-Dimensional Scaling
por: Tiblias, Federico, et al.
Publicado: (2025)
por: Tiblias, Federico, et al.
Publicado: (2025)
The Curious Case of Factual (Mis)Alignment between LLMs' Short- and Long-Form Answers
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition
por: Holtermann, Carolin, et al.
Publicado: (2024)
por: Holtermann, Carolin, et al.
Publicado: (2024)
Evaluating the Elementary Multilingual Capabilities of Large Language Models with MultiQ
por: Holtermann, Carolin, et al.
Publicado: (2024)
por: Holtermann, Carolin, et al.
Publicado: (2024)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
SPARE: Single-Pass Annotation with Reference-Guided Evaluation for Automatic Process Supervision and Reward Modelling
por: Rizvi, Md Imbesat Hassan, et al.
Publicado: (2025)
por: Rizvi, Md Imbesat Hassan, et al.
Publicado: (2025)
SpaRC and SpaRP: Spatial Reasoning Characterization and Path Generation for Understanding Spatial Reasoning Capability of Large Language Models
por: Rizvi, Md Imbesat Hassan, et al.
Publicado: (2024)
por: Rizvi, Md Imbesat Hassan, et al.
Publicado: (2024)
Uncertainty-Aware Decoding with Minimum Bayes Risk
por: Daheim, Nico, et al.
Publicado: (2025)
por: Daheim, Nico, et al.
Publicado: (2025)
CORE-T: COherent REtrieval of Tables for Text-to-SQL
por: Soliman, Hassan, et al.
Publicado: (2026)
por: Soliman, Hassan, et al.
Publicado: (2026)
Aligned Probing: Relating Toxic Behavior and Model Internals
por: Waldis, Andreas, et al.
Publicado: (2025)
por: Waldis, Andreas, et al.
Publicado: (2025)
LazyReview A Dataset for Uncovering Lazy Thinking in NLP Peer Reviews
por: Purkayastha, Sukannya, et al.
Publicado: (2025)
por: Purkayastha, Sukannya, et al.
Publicado: (2025)
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
por: Purkayastha, Sukannya, et al.
Publicado: (2026)
por: Purkayastha, Sukannya, et al.
Publicado: (2026)
A Comprehensive Review of Datasets for Clinical Mental Health AI Systems
por: Mandal, Aishik, et al.
Publicado: (2025)
por: Mandal, Aishik, et al.
Publicado: (2025)
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning
por: Niu, Jingcheng, et al.
Publicado: (2025)
por: Niu, Jingcheng, et al.
Publicado: (2025)
Towards Automated Error Discovery: A Study in Conversational AI
por: Petrak, Dominic, et al.
Publicado: (2025)
por: Petrak, Dominic, et al.
Publicado: (2025)
TempViz: On the Evaluation of Temporal Knowledge in Text-to-Image Models
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
A Survey of Confidence Estimation and Calibration in Large Language Models
por: Geng, Jiahui, et al.
Publicado: (2023)
por: Geng, Jiahui, et al.
Publicado: (2023)
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
por: Dinucu-Jianu, David, et al.
Publicado: (2025)
por: Dinucu-Jianu, David, et al.
Publicado: (2025)
Stepwise Verification and Remediation of Student Reasoning Errors with Large Language Model Tutors
por: Daheim, Nico, et al.
Publicado: (2024)
por: Daheim, Nico, et al.
Publicado: (2024)
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation Grounding
por: Paul, Indraneil, et al.
Publicado: (2025)
por: Paul, Indraneil, et al.
Publicado: (2025)
Large Language Models for Human-Machine Collaborative Particle Accelerator Tuning through Natural Language
por: Kaiser, Jan, et al.
Publicado: (2024)
por: Kaiser, Jan, et al.
Publicado: (2024)
Towards Ethical Multi-Agent Systems of Large Language Models: A Mechanistic Interpretability Perspective
por: Lee, Jae Hee, et al.
Publicado: (2025)
por: Lee, Jae Hee, et al.
Publicado: (2025)
Ejemplares similares
-
How are Prompts Different in Terms of Sensitivity?
por: Lu, Sheng, et al.
Publicado: (2023) -
LLM Roleplay: Simulating Human-Chatbot Interaction
por: Tamoyan, Hovhannes, et al.
Publicado: (2024) -
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
por: Paul, Indraneil, et al.
Publicado: (2024) -
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
por: Baumgärtner, Tim, et al.
Publicado: (2026) -
Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling
por: Tamoyan, Hovhannes, et al.
Publicado: (2025)