Aligning NLP Models with Target Population Perspectives using PAIR: Population-Aligned Instance Replication
Fuente:
arXiv
Saved in:
| Main Authors: | Eckman, Stephanie, Ma, Bolei, Kern, Christoph, Chew, Rob, Plank, Barbara, Kreuter, Frauke |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Annotation Sensitivity: Training Data Collection Methods Affect Model Performance
by: Kern, Christoph, et al.
Published: (2023)
by: Kern, Christoph, et al.
Published: (2023)
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
by: Chew, Robert, et al.
Published: (2026)
by: Chew, Robert, et al.
Published: (2026)
Position: Insights from Survey Methodology can Improve Training Data
by: Eckman, Stephanie, et al.
Published: (2024)
by: Eckman, Stephanie, et al.
Published: (2024)
The Missing Link: Allocation Performance in Causal Machine Learning
by: Fischer-Abaigar, Unai, et al.
Published: (2024)
by: Fischer-Abaigar, Unai, et al.
Published: (2024)
Bias in the Loop: How Humans Evaluate AI-Generated Suggestions
by: Beck, Jacob, et al.
Published: (2025)
by: Beck, Jacob, et al.
Published: (2025)
Bridging the gap: Towards an Expanded Toolkit for AI-driven Decision-Making in the Public Sector
by: Fischer-Abaigar, Unai, et al.
Published: (2023)
by: Fischer-Abaigar, Unai, et al.
Published: (2023)
Too Open for Opinion? Embracing Open-Endedness in Large Language Models for Social Simulation
by: Ma, Bolei, et al.
Published: (2025)
by: Ma, Bolei, et al.
Published: (2025)
The Potential and Challenges of Evaluating Attitudes, Opinions, and Values in Large Language Models
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Algorithmic Fidelity of Large Language Models in Generating Synthetic German Public Opinions: A Case Study
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
by: Wang, Xinpeng, et al.
Published: (2024)
by: Wang, Xinpeng, et al.
Published: (2024)
Pragmatics in the Era of Large Language Models: A Survey on Datasets, Evaluation, Opportunities and Challenges
by: Ma, Bolei, et al.
Published: (2025)
by: Ma, Bolei, et al.
Published: (2025)
To share or not to share: What risks would laypeople accept to give sensitive data to differentially-private NLP systems?
by: Weiss, Christopher, et al.
Published: (2023)
by: Weiss, Christopher, et al.
Published: (2023)
Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models
by: Nie, Ercong, et al.
Published: (2024)
by: Nie, Ercong, et al.
Published: (2024)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
by: Yuan, Chenchen, et al.
Published: (2026)
by: Yuan, Chenchen, et al.
Published: (2026)
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Refusal Direction is Universal Across Safety-Aligned Languages
by: Wang, Xinpeng, et al.
Published: (2025)
by: Wang, Xinpeng, et al.
Published: (2025)
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation
by: Sarmah, Bhaskarjit, et al.
Published: (2024)
by: Sarmah, Bhaskarjit, et al.
Published: (2024)
Quantifying and Attributing Submodel Uncertainty in Stochastic Simulation Models and Digital Twins
by: Ghasemloo, Mohammadmahdi, et al.
Published: (2026)
by: Ghasemloo, Mohammadmahdi, et al.
Published: (2026)
Variation is the Norm: Embracing Sociolinguistics in NLP
by: Lutgen, Anne-Marie, et al.
Published: (2026)
by: Lutgen, Anne-Marie, et al.
Published: (2026)
Look at the Text: Instruction-Tuned Language Models are More Robust Multiple Choice Selectors than You Think
by: Wang, Xinpeng, et al.
Published: (2024)
by: Wang, Xinpeng, et al.
Published: (2024)
A New Semisupervised Technique for Polarity Analysis using Masked Language Models
by: Watanabe, Kohei
Published: (2026)
by: Watanabe, Kohei
Published: (2026)
NLP-based detection of systematic anomalies among the narratives of consumer complaints
by: Gao, Peiheng, et al.
Published: (2023)
by: Gao, Peiheng, et al.
Published: (2023)
Understanding Jailbreak Success: A Study of Latent Space Dynamics in Large Language Models
by: Ball, Sarah, et al.
Published: (2024)
by: Ball, Sarah, et al.
Published: (2024)
Languages in Whisper-Style Speech Encoders Align Both Phonetically and Semantically
by: Shim, Ryan Soh-Eun, et al.
Published: (2025)
by: Shim, Ryan Soh-Eun, et al.
Published: (2025)
Geological Inference from Textual Data using Word Embeddings
by: Linphrachaya, Nanmanas, et al.
Published: (2025)
by: Linphrachaya, Nanmanas, et al.
Published: (2025)
Goodness of Fit for Bayesian Generative Models with Applications in Population Genetics
by: Mailloux, Guillaume Le, et al.
Published: (2025)
by: Mailloux, Guillaume Le, et al.
Published: (2025)
Automated scoring of the Ambiguous Intentions Hostility Questionnaire using fine-tuned large language models
by: Lyu, Y., et al.
Published: (2025)
by: Lyu, Y., et al.
Published: (2025)
Population-Aligned Persona Generation for LLM-based Social Simulation
by: Hu, Zhengyu, et al.
Published: (2025)
by: Hu, Zhengyu, et al.
Published: (2025)
Maximizing Signal in Human-Model Preference Alignment
by: Kraus, Kelsey, et al.
Published: (2025)
by: Kraus, Kelsey, et al.
Published: (2025)
Safety Alignment in NLP Tasks: Weakly Aligned Summarization as an In-Context Attack
by: Fu, Yu, et al.
Published: (2023)
by: Fu, Yu, et al.
Published: (2023)
Syntax-Guided Diffusion Language Models with User-Integrated Personalization
by: Zhang, Ruqian, et al.
Published: (2025)
by: Zhang, Ruqian, et al.
Published: (2025)
Add Noise, Tasks, or Layers? MaiNLP at the VarDial 2025 Shared Task on Norwegian Dialectal Slot and Intent Detection
by: Blaschke, Verena, et al.
Published: (2025)
by: Blaschke, Verena, et al.
Published: (2025)
Population Power Curves in ASCA with Permutation Testing
by: Camacho, Jose, et al.
Published: (2024)
by: Camacho, Jose, et al.
Published: (2024)
Better Aligned with Survey Respondents or Training Data? Unveiling Political Leanings of LLMs on U.S. Supreme Court Cases
by: Xu, Shanshan, et al.
Published: (2025)
by: Xu, Shanshan, et al.
Published: (2025)
Distributed Asymmetric Allocation: A Topic Model for Large Imbalanced Corpora in Social Sciences
by: Watanabe, Kohei
Published: (2025)
by: Watanabe, Kohei
Published: (2025)
An Embedded Diachronic Sense Change Model with a Case Study from Ancient Greek
by: Zafar, Schyan, et al.
Published: (2023)
by: Zafar, Schyan, et al.
Published: (2023)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
by: Xu, Shanshan, et al.
Published: (2025)
by: Xu, Shanshan, et al.
Published: (2025)
How Model Size, Temperature, and Prompt Style Affect LLM-Human Assessment Score Alignment
by: Jung, Julie, et al.
Published: (2025)
by: Jung, Julie, et al.
Published: (2025)
Leveraging text data for causal inference using electronic health records
by: Mozer, Reagan, et al.
Published: (2023)
by: Mozer, Reagan, et al.
Published: (2023)
Can Large Language Models Understand You Better? An MBTI Personality Detection Dataset Aligned with Population Traits
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Similar Items
-
Annotation Sensitivity: Training Data Collection Methods Affect Model Performance
by: Kern, Christoph, et al.
Published: (2023) -
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
by: Chew, Robert, et al.
Published: (2026) -
Position: Insights from Survey Methodology can Improve Training Data
by: Eckman, Stephanie, et al.
Published: (2024) -
The Missing Link: Allocation Performance in Causal Machine Learning
by: Fischer-Abaigar, Unai, et al.
Published: (2024) -
Bias in the Loop: How Humans Evaluate AI-Generated Suggestions
by: Beck, Jacob, et al.
Published: (2025)