Are LLMs good pragmatic speakers?
Fuente:
arXiv
Salvato in:
| Autori principali: | Jian, Mingyue, Siddharth, N. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs
di: Aswal, Darpan, et al.
Pubblicazione: (2025)
di: Aswal, Darpan, et al.
Pubblicazione: (2025)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
di: Karia, Rushang, et al.
Pubblicazione: (2024)
di: Karia, Rushang, et al.
Pubblicazione: (2024)
Addressing speaker gender bias in large scale speech translation systems
di: Bansal, Shubham, et al.
Pubblicazione: (2025)
di: Bansal, Shubham, et al.
Pubblicazione: (2025)
Does GPT-4 surpass human performance in linguistic pragmatics?
di: Bojic, Ljubisa, et al.
Pubblicazione: (2023)
di: Bojic, Ljubisa, et al.
Pubblicazione: (2023)
$\forall$uto$\exists$val: Autonomous Assessment of LLMs in Formal Synthesis and Interpretation Tasks
di: Karia, Rushang, et al.
Pubblicazione: (2024)
di: Karia, Rushang, et al.
Pubblicazione: (2024)
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology
di: Yu, Danni, et al.
Pubblicazione: (2023)
di: Yu, Danni, et al.
Pubblicazione: (2023)
Profiling learners' affective engagement: Emotion AI, intercultural pragmatics, and language learning
di: Godwin-Jones, Robert
Pubblicazione: (2026)
di: Godwin-Jones, Robert
Pubblicazione: (2026)
Benchmarking Multimodal LLMs on Recognition and Understanding over Chemical Tables
di: Zhou, Yitong, et al.
Pubblicazione: (2025)
di: Zhou, Yitong, et al.
Pubblicazione: (2025)
Serialized EHR make for good text representations
di: Chou, Zhirong, et al.
Pubblicazione: (2025)
di: Chou, Zhirong, et al.
Pubblicazione: (2025)
Brotherhood at WMT 2024: Leveraging LLM-Generated Contextual Conversations for Cross-Lingual Image Captioning
di: Betala, Siddharth, et al.
Pubblicazione: (2024)
di: Betala, Siddharth, et al.
Pubblicazione: (2024)
HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
Evaluating LLMs for Hardware Design and Test
di: Blocklove, Jason, et al.
Pubblicazione: (2024)
di: Blocklove, Jason, et al.
Pubblicazione: (2024)
Rethinking LLM Bias Probing Using Lessons from the Social Sciences
di: Morehouse, Kirsten N., et al.
Pubblicazione: (2025)
di: Morehouse, Kirsten N., et al.
Pubblicazione: (2025)
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench
di: Narad, Reuben, et al.
Pubblicazione: (2025)
di: Narad, Reuben, et al.
Pubblicazione: (2025)
How good is GPT at writing political speeches for the White House?
di: Savoy, Jacques
Pubblicazione: (2024)
di: Savoy, Jacques
Pubblicazione: (2024)
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs
di: Luo, Hang, et al.
Pubblicazione: (2025)
di: Luo, Hang, et al.
Pubblicazione: (2025)
Can we trust AI to detect healthy multilingual English speakers among the cognitively impaired cohort in the UK? An investigation using real-world conversational speech
di: Pahar, Madhurananda, et al.
Pubblicazione: (2026)
di: Pahar, Madhurananda, et al.
Pubblicazione: (2026)
LLMs on a Budget? Say HOLA
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
REAMS: Reasoning Enhanced Algorithm for Maths Solving
di: Singh, Eishkaran, et al.
Pubblicazione: (2025)
di: Singh, Eishkaran, et al.
Pubblicazione: (2025)
The Diminishing Returns of Early-Exit Decoding in Modern LLMs
di: Wei, Rui, et al.
Pubblicazione: (2026)
di: Wei, Rui, et al.
Pubblicazione: (2026)
How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives
di: Ichmoukhamedov, Timour, et al.
Pubblicazione: (2024)
di: Ichmoukhamedov, Timour, et al.
Pubblicazione: (2024)
From Answers to Questions: EQGBench for Evaluating LLMs' Educational Question Generation
di: Zhou, Chengliang, et al.
Pubblicazione: (2025)
di: Zhou, Chengliang, et al.
Pubblicazione: (2025)
Parallelograms Strike Back: LLMs Generate Better Analogies than People
di: Liu, Qiawen Ella, et al.
Pubblicazione: (2026)
di: Liu, Qiawen Ella, et al.
Pubblicazione: (2026)
Learning Evidence Highlighting for Frozen LLMs
di: Li, Shaoang, et al.
Pubblicazione: (2026)
di: Li, Shaoang, et al.
Pubblicazione: (2026)
Hybrid-NL2SVA: Integrating RAG and Finetuning for LLM-based NL2SVA
di: Xiao, Weihua, et al.
Pubblicazione: (2025)
di: Xiao, Weihua, et al.
Pubblicazione: (2025)
Enhancing Public Speaking Skills in Engineering Students Through AI
di: Harsh, Amol, et al.
Pubblicazione: (2025)
di: Harsh, Amol, et al.
Pubblicazione: (2025)
Lived Experience Not Found: LLMs Struggle to Align with Experts on Addressing Adverse Drug Reactions from Psychiatric Medication Use
di: Chandra, Mohit, et al.
Pubblicazione: (2024)
di: Chandra, Mohit, et al.
Pubblicazione: (2024)
Discerning minds or generic tutors? Evaluating instructional guidance capabilities in Socratic LLMs
di: Liu, Ying, et al.
Pubblicazione: (2025)
di: Liu, Ying, et al.
Pubblicazione: (2025)
Training Turn-by-Turn Verifiers for Dialogue Tutoring Agents: The Curious Case of LLMs as Your Coding Tutors
di: Wang, Jian, et al.
Pubblicazione: (2025)
di: Wang, Jian, et al.
Pubblicazione: (2025)
Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness
di: Banayeeanzade, Amin, et al.
Pubblicazione: (2025)
di: Banayeeanzade, Amin, et al.
Pubblicazione: (2025)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
di: Patel, Shubham, et al.
Pubblicazione: (2024)
di: Patel, Shubham, et al.
Pubblicazione: (2024)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
di: Wei, Jianhui, et al.
Pubblicazione: (2025)
di: Wei, Jianhui, et al.
Pubblicazione: (2025)
Do LLMs Triage Like Clinicians? A Dynamic Study of Outpatient Referral
di: Liu, Xiaoxiao, et al.
Pubblicazione: (2025)
di: Liu, Xiaoxiao, et al.
Pubblicazione: (2025)
Do Compressed LLMs Forget Knowledge? An Experimental Study with Practical Implications
di: Hoang, Duc N. M, et al.
Pubblicazione: (2023)
di: Hoang, Duc N. M, et al.
Pubblicazione: (2023)
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation
di: Ouyang, Jie, et al.
Pubblicazione: (2025)
di: Ouyang, Jie, et al.
Pubblicazione: (2025)
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning
di: Zhang, Xiaotian, et al.
Pubblicazione: (2025)
di: Zhang, Xiaotian, et al.
Pubblicazione: (2025)
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
VoxHakka: A Dialectally Diverse Multi-speaker Text-to-Speech System for Taiwanese Hakka
di: Chen, Li-Wei, et al.
Pubblicazione: (2024)
di: Chen, Li-Wei, et al.
Pubblicazione: (2024)
Farther the Shift, Sparser the Representation: Analyzing OOD Mechanisms in LLMs
di: Jin, Mingyu, et al.
Pubblicazione: (2026)
di: Jin, Mingyu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs
di: Aswal, Darpan, et al.
Pubblicazione: (2025) -
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
di: Karia, Rushang, et al.
Pubblicazione: (2024) -
Addressing speaker gender bias in large scale speech translation systems
di: Bansal, Shubham, et al.
Pubblicazione: (2025) -
Does GPT-4 surpass human performance in linguistic pragmatics?
di: Bojic, Ljubisa, et al.
Pubblicazione: (2023) -
$\forall$uto$\exists$val: Autonomous Assessment of LLMs in Formal Synthesis and Interpretation Tasks
di: Karia, Rushang, et al.
Pubblicazione: (2024)