When Does Meaning Backfire? Investigating the Role of AMRs in NLI
Fuente:
arXiv
Saved in:
| Main Authors: | Min, Junghyun, Yang, Xiulin, Wein, Shira |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Uniform Meaning Representation Help GPT-4 Translate from Indigenous Languages?
by: Wein, Shira
Published: (2025)
by: Wein, Shira
Published: (2025)
Lost in Translationese? Reducing Translation Effect Using Abstract Meaning Representation
by: Wein, Shira, et al.
Published: (2023)
by: Wein, Shira, et al.
Published: (2023)
Generating Text from Uniform Meaning Representation
by: Markle, Emma, et al.
Published: (2025)
by: Markle, Emma, et al.
Published: (2025)
SETUP: Sentence-level English-To-Uniform Meaning Representation Parser
by: Markle, Emma, et al.
Published: (2025)
by: Markle, Emma, et al.
Published: (2025)
MASSIVE Multilingual Abstract Meaning Representation: A Dataset and Baselines for Hallucination Detection
by: Regan, Michael, et al.
Published: (2024)
by: Regan, Michael, et al.
Published: (2024)
When Thinking Backfires: Mechanistic Insights Into Reasoning-Induced Misalignment
by: Yan, Hanqi, et al.
Published: (2025)
by: Yan, Hanqi, et al.
Published: (2025)
Natural Language Processing RELIES on Linguistics
by: Opitz, Juri, et al.
Published: (2024)
by: Opitz, Juri, et al.
Published: (2024)
When Incentives Backfire, Data Stops Being Human
by: Santy, Sebastin, et al.
Published: (2025)
by: Santy, Sebastin, et al.
Published: (2025)
Language Models at the Syntax-Semantics Interface: A Case Study of the Long-Distance Binding of Chinese Reflexive ziji
by: Yang, Xiulin
Published: (2025)
by: Yang, Xiulin
Published: (2025)
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
by: Sadanandan, Binesh, et al.
Published: (2026)
by: Sadanandan, Binesh, et al.
Published: (2026)
Can LLMs Interpret and Leverage Structured Linguistic Representations? A Case Study with AMRs
by: Raut, Ankush, et al.
Published: (2025)
by: Raut, Ankush, et al.
Published: (2025)
What Does Neuro Mean to Cardio? Investigating the Role of Clinical Specialty Data in Medical LLMs
by: Yan, Xinlan, et al.
Published: (2025)
by: Yan, Xinlan, et al.
Published: (2025)
Does Less Hallucination Mean Less Creativity? An Empirical Investigation in LLMs
by: Banerjee, Mohor, et al.
Published: (2025)
by: Banerjee, Mohor, et al.
Published: (2025)
When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations
by: Aluguvelly, Avinash Goutham
Published: (2026)
by: Aluguvelly, Avinash Goutham
Published: (2026)
EconNLI: Evaluating Large Language Models on Economics Reasoning
by: Guo, Yue, et al.
Published: (2024)
by: Guo, Yue, et al.
Published: (2024)
When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling
by: Yun, Heecheol, et al.
Published: (2025)
by: Yun, Heecheol, et al.
Published: (2025)
Building UD Cairo for Old English in the Classroom
by: Levine, Lauren, et al.
Published: (2025)
by: Levine, Lauren, et al.
Published: (2025)
Entailed Between the Lines: Incorporating Implication into NLI
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Rethinking STS and NLI in Large Language Models
by: Wang, Yuxia, et al.
Published: (2023)
by: Wang, Yuxia, et al.
Published: (2023)
Exploring Continual Learning of Compositional Generalization in NLI
by: Fu, Xiyan, et al.
Published: (2024)
by: Fu, Xiyan, et al.
Published: (2024)
Semantic Gravity Wells: Why Negative Constraints Backfire
by: Rana, Shailesh
Published: (2026)
by: Rana, Shailesh
Published: (2026)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
by: Wen, Zhiyuan, et al.
Published: (2024)
by: Wen, Zhiyuan, et al.
Published: (2024)
From Disagreement to Understanding: The Case for Ambiguity Detection in NLI
by: Jayaweera, Chathuri, et al.
Published: (2025)
by: Jayaweera, Chathuri, et al.
Published: (2025)
For Generated Text, Is NLI-Neutral Text the Best Text?
by: Mersinias, Michail, et al.
Published: (2023)
by: Mersinias, Michail, et al.
Published: (2023)
Mathematical Reasoning Enhanced LLM for Formula Derivation: A Case Study on Fiber NLI Modellin
by: Zhang, Yao, et al.
Published: (2026)
by: Zhang, Yao, et al.
Published: (2026)
When Does Monolingual Data Help Multilingual Translation: The Role of Domain and Model Scale
by: Baziotis, Christos, et al.
Published: (2023)
by: Baziotis, Christos, et al.
Published: (2023)
SocialNLI: A Dialogue-Centric Social Inference Dataset
by: Deo, Akhil, et al.
Published: (2025)
by: Deo, Akhil, et al.
Published: (2025)
NLI under the Microscope: What Atomic Hypothesis Decomposition Reveals
by: Srikanth, Neha, et al.
Published: (2025)
by: Srikanth, Neha, et al.
Published: (2025)
A synthetic data approach for domain generalization of NLI models
by: Hosseini, Mohammad Javad, et al.
Published: (2024)
by: Hosseini, Mohammad Javad, et al.
Published: (2024)
Exploring Factual Entailment with NLI: A News Media Study
by: Mor-Lan, Guy, et al.
Published: (2024)
by: Mor-Lan, Guy, et al.
Published: (2024)
NL-Eye: Abductive NLI for Images
by: Ventura, Mor, et al.
Published: (2024)
by: Ventura, Mor, et al.
Published: (2024)
Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!
by: Zhou, Zhanhui, et al.
Published: (2024)
by: Zhou, Zhanhui, et al.
Published: (2024)
Self-Rationalization in the Wild: A Large Scale Out-of-Distribution Evaluation on NLI-related tasks
by: Yang, Jing, et al.
Published: (2025)
by: Yang, Jing, et al.
Published: (2025)
Punctuation Restoration Improves Structure Understanding Without Supervision
by: Min, Junghyun, et al.
Published: (2024)
by: Min, Junghyun, et al.
Published: (2024)
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
by: Bengoetxea, Jaione, et al.
Published: (2025)
by: Bengoetxea, Jaione, et al.
Published: (2025)
EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI
by: Zuo, Longfei, et al.
Published: (2025)
by: Zuo, Longfei, et al.
Published: (2025)
Medical Argument Mining: Exploitation of Scarce Data Using NLI Systems
by: Urruela, Maitane, et al.
Published: (2025)
by: Urruela, Maitane, et al.
Published: (2025)
Estimating the Causal Effects of Natural Logic Features in Neural NLI Models
by: Rozanova, Julia, et al.
Published: (2023)
by: Rozanova, Julia, et al.
Published: (2023)
How Hard is this Test Set? NLI Characterization by Exploiting Training Dynamics
by: Cosma, Adrian, et al.
Published: (2024)
by: Cosma, Adrian, et al.
Published: (2024)
MSciNLI: A Diverse Benchmark for Scientific Natural Language Inference
by: Sadat, Mobashir, et al.
Published: (2024)
by: Sadat, Mobashir, et al.
Published: (2024)
Similar Items
-
Can Uniform Meaning Representation Help GPT-4 Translate from Indigenous Languages?
by: Wein, Shira
Published: (2025) -
Lost in Translationese? Reducing Translation Effect Using Abstract Meaning Representation
by: Wein, Shira, et al.
Published: (2023) -
Generating Text from Uniform Meaning Representation
by: Markle, Emma, et al.
Published: (2025) -
SETUP: Sentence-level English-To-Uniform Meaning Representation Parser
by: Markle, Emma, et al.
Published: (2025) -
MASSIVE Multilingual Abstract Meaning Representation: A Dataset and Baselines for Hallucination Detection
by: Regan, Michael, et al.
Published: (2024)