Tug-of-war between idioms' figurative and literal interpretations in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Oh, Soyoung, Huang, Xinting, Pink, Mathis, Hahn, Michael, Demberg, Vera |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
Can generative AI figure out figurative language? The influence of idioms on essay scoring by ChatGPT, Gemini, and Deepseek
di: Oğuz, Enis
Pubblicazione: (2025)
di: Oğuz, Enis
Pubblicazione: (2025)
Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning
di: Huang, Xinting, et al.
Pubblicazione: (2025)
di: Huang, Xinting, et al.
Pubblicazione: (2025)
RSA-Control: A Pragmatics-Grounded Lightweight Controllable Text Generation Framework
di: Wang, Yifan, et al.
Pubblicazione: (2024)
di: Wang, Yifan, et al.
Pubblicazione: (2024)
ChatGPT vs Human-authored Text: Insights into Controllable Text Summarization and Sentence Style Transfer
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
RST-LoRA: A Discourse-Aware Low-Rank Adaptation for Long Document Abstractive Summarization
di: Liu, Dongqi, et al.
Pubblicazione: (2024)
di: Liu, Dongqi, et al.
Pubblicazione: (2024)
Temperature-scaling surprisal estimates improve fit to human reading times -- but does it do so for the "right reasons"?
di: Liu, Tong, et al.
Pubblicazione: (2023)
di: Liu, Tong, et al.
Pubblicazione: (2023)
Can LLMs interpret figurative language as humans do?: surface-level vs representational similarity
di: Bollepally, Samhita, et al.
Pubblicazione: (2026)
di: Bollepally, Samhita, et al.
Pubblicazione: (2026)
InversionView: A General-Purpose Method for Reading Information from Neural Activations
di: Huang, Xinting, et al.
Pubblicazione: (2024)
di: Huang, Xinting, et al.
Pubblicazione: (2024)
Prompting Implicit Discourse Relation Annotation
di: Yung, Frances, et al.
Pubblicazione: (2024)
di: Yung, Frances, et al.
Pubblicazione: (2024)
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
di: Huang, Xinting, et al.
Pubblicazione: (2026)
di: Huang, Xinting, et al.
Pubblicazione: (2026)
Modeling Orthographic Variation Improves NLP Performance for Nigerian Pidgin
di: Lin, Pin-Jie, et al.
Pubblicazione: (2024)
di: Lin, Pin-Jie, et al.
Pubblicazione: (2024)
Planning Ahead with RSA: Efficient Signalling in Dynamic Environments by Projecting User Awareness across Future Timesteps
di: Das, Anwesha, et al.
Pubblicazione: (2025)
di: Das, Anwesha, et al.
Pubblicazione: (2025)
Pragmatic Reasoning improves LLM Code Generation
di: Cao, Zhuchen, et al.
Pubblicazione: (2025)
di: Cao, Zhuchen, et al.
Pubblicazione: (2025)
Explanatory Summarization with Discourse-Driven Planning
di: Liu, Dongqi, et al.
Pubblicazione: (2025)
di: Liu, Dongqi, et al.
Pubblicazione: (2025)
SciNews: From Scholarly Complexities to Public Narratives -- A Dataset for Scientific News Report Generation
di: Liu, Dongqi, et al.
Pubblicazione: (2024)
di: Liu, Dongqi, et al.
Pubblicazione: (2024)
B-cos LM: Efficiently Transforming Pre-trained Language Models for Improved Explainability
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
Atomic Calibration of LLMs in Long-Form Generations
di: Zhang, Caiqi, et al.
Pubblicazione: (2024)
di: Zhang, Caiqi, et al.
Pubblicazione: (2024)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
di: Pink, Mathis, et al.
Pubblicazione: (2024)
di: Pink, Mathis, et al.
Pubblicazione: (2024)
Tug-of-War within A Decade: Conflict Resolution in Vulnerability Analysis via Teacher-Guided Retrieval-Augmented Generations
di: Zhou, Ziyin, et al.
Pubblicazione: (2026)
di: Zhou, Ziyin, et al.
Pubblicazione: (2026)
ClashEval: Quantifying the tug-of-war between an LLM's internal prior and external evidence
di: Wu, Kevin, et al.
Pubblicazione: (2024)
di: Wu, Kevin, et al.
Pubblicazione: (2024)
Tug-of-War Between Knowledge: Exploring and Resolving Knowledge Conflicts in Retrieval-Augmented Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization
di: Saha, Anisha, et al.
Pubblicazione: (2026)
di: Saha, Anisha, et al.
Pubblicazione: (2026)
SAFE: Stepwise Atomic Feedback for Error correction in Multi-hop Reasoning
di: Kwon, Daeyong, et al.
Pubblicazione: (2026)
di: Kwon, Daeyong, et al.
Pubblicazione: (2026)
An evaluation of LLMs for political bias in Western media: Israel-Hamas and Ukraine-Russia wars
di: Chandra, Rohitash, et al.
Pubblicazione: (2026)
di: Chandra, Rohitash, et al.
Pubblicazione: (2026)
House of Cards: Massive Weights in LLMs
di: Oh, Jaehoon, et al.
Pubblicazione: (2024)
di: Oh, Jaehoon, et al.
Pubblicazione: (2024)
CBEval: A framework for evaluating and interpreting cognitive biases in LLMs
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages
di: Huang, Xuhan, et al.
Pubblicazione: (2024)
di: Huang, Xuhan, et al.
Pubblicazione: (2024)
Surrogate modeling for interpreting black-box LLMs in medical predictions
di: Han, Changho, et al.
Pubblicazione: (2026)
di: Han, Changho, et al.
Pubblicazione: (2026)
Are LLMs effective psychological assessors? Leveraging adaptive RAG for interpretable mental health screening through psychometric practice
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
di: Huang, Jianheng, et al.
Pubblicazione: (2024)
di: Huang, Jianheng, et al.
Pubblicazione: (2024)
Free(): Learning to Forget in Malloc-Only Reasoning Models
di: Zheng, Yilun, et al.
Pubblicazione: (2026)
di: Zheng, Yilun, et al.
Pubblicazione: (2026)
LoGU: Long-form Generation with Uncertainty Expressions
di: Yang, Ruihan, et al.
Pubblicazione: (2024)
di: Yang, Ruihan, et al.
Pubblicazione: (2024)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
LLMs syntactically adapt their language use to their conversational partner
di: Kandra, Florian, et al.
Pubblicazione: (2025)
di: Kandra, Florian, et al.
Pubblicazione: (2025)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
di: Cho, Hojun, et al.
Pubblicazione: (2025)
di: Cho, Hojun, et al.
Pubblicazione: (2025)
What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations
di: Liu, Dongqi, et al.
Pubblicazione: (2025)
di: Liu, Dongqi, et al.
Pubblicazione: (2025)
What Defines Good Reasoning in LLMs? Dissecting Reasoning Steps with Multi-Aspect Evaluation
di: Do, Heejin, et al.
Pubblicazione: (2025)
di: Do, Heejin, et al.
Pubblicazione: (2025)
The End of Manual Decoding: Towards Truly End-to-End Language Models
di: Wang, Zhichao, et al.
Pubblicazione: (2025)
di: Wang, Zhichao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
di: Wang, Yifan, et al.
Pubblicazione: (2025) -
Can generative AI figure out figurative language? The influence of idioms on essay scoring by ChatGPT, Gemini, and Deepseek
di: Oğuz, Enis
Pubblicazione: (2025) -
Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning
di: Huang, Xinting, et al.
Pubblicazione: (2025) -
RSA-Control: A Pragmatics-Grounded Lightweight Controllable Text Generation Framework
di: Wang, Yifan, et al.
Pubblicazione: (2024) -
ChatGPT vs Human-authored Text: Insights into Controllable Text Summarization and Sentence Style Transfer
di: Liu, Dongqi, et al.
Pubblicazione: (2023)