Pragmatics in the Era of Large Language Models: A Survey on Datasets, Evaluation, Opportunities and Challenges
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Bolei, Li, Yuting, Zhou, Wei, Gong, Ziwei, Liu, Yang Janet, Jasinskaja, Katja, Friedrich, Annemarie, Hirschberg, Julia, Kreuter, Frauke, Plank, Barbara |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Potential and Challenges of Evaluating Attitudes, Opinions, and Values in Large Language Models
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Table Question Answering in the Era of Large Language Models: A Comprehensive Survey of Tasks, Methods, and Evaluation
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
Position: Insights from Survey Methodology can Improve Training Data
by: Eckman, Stephanie, et al.
Published: (2024)
by: Eckman, Stephanie, et al.
Published: (2024)
Too Open for Opinion? Embracing Open-Endedness in Large Language Models for Social Simulation
by: Ma, Bolei, et al.
Published: (2025)
by: Ma, Bolei, et al.
Published: (2025)
Aligning NLP Models with Target Population Perspectives using PAIR: Population-Aligned Instance Replication
by: Eckman, Stephanie, et al.
Published: (2025)
by: Eckman, Stephanie, et al.
Published: (2025)
Algorithmic Fidelity of Large Language Models in Generating Synthetic German Public Opinions: A Case Study
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Multimodal Emotion Recognition in Conversations: A Survey of Methods, Trends, Challenges and Prospects
by: Wu, Chengyan, et al.
Published: (2025)
by: Wu, Chengyan, et al.
Published: (2025)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
by: Wang, Xinpeng, et al.
Published: (2024)
by: Wang, Xinpeng, et al.
Published: (2024)
Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models
by: Nie, Ercong, et al.
Published: (2024)
by: Nie, Ercong, et al.
Published: (2024)
SURE: Synergistic Uncertainty-aware Reasoning for Multimodal Emotion Recognition in Conversations
by: Cai, Yiqiang, et al.
Published: (2026)
by: Cai, Yiqiang, et al.
Published: (2026)
Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey
by: Mondorf, Philipp, et al.
Published: (2024)
by: Mondorf, Philipp, et al.
Published: (2024)
Understanding Jailbreak Success: A Study of Latent Space Dynamics in Large Language Models
by: Ball, Sarah, et al.
Published: (2024)
by: Ball, Sarah, et al.
Published: (2024)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
by: Yuan, Chenchen, et al.
Published: (2026)
by: Yuan, Chenchen, et al.
Published: (2026)
Annotation Sensitivity: Training Data Collection Methods Affect Model Performance
by: Kern, Christoph, et al.
Published: (2023)
by: Kern, Christoph, et al.
Published: (2023)
Chapter 7 Digital Trace Data
by: Keusch, Florian, et al.
Published: (2021)
by: Keusch, Florian, et al.
Published: (2021)
Multimodal Multi-loss Fusion Network for Sentiment Analysis
by: Wu, Zehui, et al.
Published: (2023)
by: Wu, Zehui, et al.
Published: (2023)
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Look at the Text: Instruction-Tuned Language Models are More Robust Multiple Choice Selectors than You Think
by: Wang, Xinpeng, et al.
Published: (2024)
by: Wang, Xinpeng, et al.
Published: (2024)
Human Preferences in Large Language Model Latent Space: A Technical Analysis on the Reliability of Synthetic Data in Voting Outcome Prediction
by: Ball, Sarah, et al.
Published: (2025)
by: Ball, Sarah, et al.
Published: (2025)
A Survey on Open Information Extraction from Rule-based Model to Large Language Model
by: Liu, Pai, et al.
Published: (2022)
by: Liu, Pai, et al.
Published: (2022)
Huntington Disease Automatic Speech Recognition with Biomarker Supervision
by: Wang, Charles L., et al.
Published: (2026)
by: Wang, Charles L., et al.
Published: (2026)
Bias in the Loop: How Humans Evaluate AI-Generated Suggestions
by: Beck, Jacob, et al.
Published: (2025)
by: Beck, Jacob, et al.
Published: (2025)
Trustworthy Recommendation in the Era of Large Language Models: Opportunities and Challenges
by: Wang, Bohao, et al.
Published: (2026)
by: Wang, Bohao, et al.
Published: (2026)
The Imperfective Paradox in Large Language Models
by: Ma, Bolei, et al.
Published: (2026)
by: Ma, Bolei, et al.
Published: (2026)
Detecting Mental Manipulation in Speech via Synthetic Multi-Speaker Dialogue
by: Chen, Run, et al.
Published: (2026)
by: Chen, Run, et al.
Published: (2026)
The Impact of Question Framing on the Performance of Automatic Occupation Coding
by: Kononykhina, Olga, et al.
Published: (2025)
by: Kononykhina, Olga, et al.
Published: (2025)
To share or not to share: What risks would laypeople accept to give sensitive data to differentially-private NLP systems?
by: Weiss, Christopher, et al.
Published: (2023)
by: Weiss, Christopher, et al.
Published: (2023)
Sensing What Surveys Miss: Understanding and Personalizing Proactive LLM Support by User Modeling
by: Liu, Ailin, et al.
Published: (2026)
by: Liu, Ailin, et al.
Published: (2026)
NovAScore: A New Automated Metric for Evaluating Document Level Novelty
by: Ai, Lin, et al.
Published: (2024)
by: Ai, Lin, et al.
Published: (2024)
Comparative Evaluation of Expressive Japanese Character Text-to-Speech with VITS and Style-BERT-VITS2
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
Animating Language Practice: Engagement with Stylized Conversational Agents in Japanese Learning
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models
by: Ahnert, Georg, et al.
Published: (2025)
by: Ahnert, Georg, et al.
Published: (2025)
Akan Cinematic Emotions (ACE): A Multimodal Multi-party Dataset for Emotion Recognition in Movie Dialogues
by: Sasu, David, et al.
Published: (2025)
by: Sasu, David, et al.
Published: (2025)
The Missing Link: Allocation Performance in Causal Machine Learning
by: Fischer-Abaigar, Unai, et al.
Published: (2024)
by: Fischer-Abaigar, Unai, et al.
Published: (2024)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
by: Mondorf, Philipp, et al.
Published: (2024)
by: Mondorf, Philipp, et al.
Published: (2024)
Toward Understanding the Transferability of Adversarial Suffixes in Large Language Models
by: Ball, Sarah, et al.
Published: (2025)
by: Ball, Sarah, et al.
Published: (2025)
Beyond Silent Letters: Amplifying LLMs in Emotion Recognition with Vocal Nuances
by: Wu, Zehui, et al.
Published: (2024)
by: Wu, Zehui, et al.
Published: (2024)
Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language Models
by: Mondorf, Philipp, et al.
Published: (2024)
by: Mondorf, Philipp, et al.
Published: (2024)
Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry
by: Ma, Bolei, et al.
Published: (2025)
by: Ma, Bolei, et al.
Published: (2025)
CREAM: Comparison-Based Reference-Free ELO-Ranked Automatic Evaluation for Meeting Summarization
by: Gong, Ziwei, et al.
Published: (2024)
by: Gong, Ziwei, et al.
Published: (2024)
Similar Items
-
The Potential and Challenges of Evaluating Attitudes, Opinions, and Values in Large Language Models
by: Ma, Bolei, et al.
Published: (2024) -
Table Question Answering in the Era of Large Language Models: A Comprehensive Survey of Tasks, Methods, and Evaluation
by: Zhou, Wei, et al.
Published: (2025) -
Position: Insights from Survey Methodology can Improve Training Data
by: Eckman, Stephanie, et al.
Published: (2024) -
Too Open for Opinion? Embracing Open-Endedness in Large Language Models for Social Simulation
by: Ma, Bolei, et al.
Published: (2025) -
Aligning NLP Models with Target Population Perspectives using PAIR: Population-Aligned Instance Replication
by: Eckman, Stephanie, et al.
Published: (2025)