LLMs Struggle with Abstract Meaning Comprehension More Than Expected
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alhazmi, Hamoud, Jiang, Jiachen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Algebra of Meaning: Why Machines Need Montague More Than Moore's Law
von: Jeong, Cheonkam, et al.
Veröffentlicht: (2025)
von: Jeong, Cheonkam, et al.
Veröffentlicht: (2025)
LLMs Know More Than Words: A Genre Study with Syntax, Metaphor & Phonetics
von: Shi, Weiye, et al.
Veröffentlicht: (2025)
von: Shi, Weiye, et al.
Veröffentlicht: (2025)
From Instruction to Output: The Role of Prompting in Modern NLG
von: Zaib, Munazza, et al.
Veröffentlicht: (2026)
von: Zaib, Munazza, et al.
Veröffentlicht: (2026)
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2026)
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2026)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
Long-context LLMs Struggle with Long In-context Learning
von: Li, Tianle, et al.
Veröffentlicht: (2024)
von: Li, Tianle, et al.
Veröffentlicht: (2024)
LLMs Position Themselves as More Rational Than Humans: Emergence of AI Self-Awareness Measured Through Game Theory
von: Kim, Kyung-Hoon
Veröffentlicht: (2025)
von: Kim, Kyung-Hoon
Veröffentlicht: (2025)
Many LLMs Are More Utilitarian Than One
von: Keshmirian, Anita, et al.
Veröffentlicht: (2025)
von: Keshmirian, Anita, et al.
Veröffentlicht: (2025)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
von: Jiang, Dongwei, et al.
Veröffentlicht: (2024)
von: Jiang, Dongwei, et al.
Veröffentlicht: (2024)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
von: Pal, Soumyadeep, et al.
Veröffentlicht: (2025)
von: Pal, Soumyadeep, et al.
Veröffentlicht: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
More Than Meets the Eye? Uncovering the Reasoning-Planning Disconnect in Training Vision-Language Driving Models
von: Song, Xurui, et al.
Veröffentlicht: (2025)
von: Song, Xurui, et al.
Veröffentlicht: (2025)
Trivial Vocabulary Bans Improve LLM Reasoning More Than Deep Linguistic Constraints
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
More Than a Quick Glance: Overcoming the Greedy Bias in KV-Cache Compression
von: Sood, Aryan, et al.
Veröffentlicht: (2026)
von: Sood, Aryan, et al.
Veröffentlicht: (2026)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
von: Malek, Alan, et al.
Veröffentlicht: (2025)
von: Malek, Alan, et al.
Veröffentlicht: (2025)
Debating with More Persuasive LLMs Leads to More Truthful Answers
von: Khan, Akbir, et al.
Veröffentlicht: (2024)
von: Khan, Akbir, et al.
Veröffentlicht: (2024)
More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection
von: Sun, Runze, et al.
Veröffentlicht: (2026)
von: Sun, Runze, et al.
Veröffentlicht: (2026)
Tokenization Is More Than Compression
von: Schmidt, Craig W., et al.
Veröffentlicht: (2024)
von: Schmidt, Craig W., et al.
Veröffentlicht: (2024)
Abstract Meaning Representation-Based Logic-Driven Data Augmentation for Logical Reasoning
von: Bao, Qiming, et al.
Veröffentlicht: (2023)
von: Bao, Qiming, et al.
Veröffentlicht: (2023)
Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models
von: Bakhshi, Asim D.
Veröffentlicht: (2026)
von: Bakhshi, Asim D.
Veröffentlicht: (2026)
Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversations
von: Liu, Zijie, et al.
Veröffentlicht: (2025)
von: Liu, Zijie, et al.
Veröffentlicht: (2025)
Why Low-Resource NLP Needs More Than Cross-Lingual Transfer: Lessons Learned from Luxembourgish
von: Philippy, Fred, et al.
Veröffentlicht: (2026)
von: Philippy, Fred, et al.
Veröffentlicht: (2026)
Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models
von: Dong, Dota Tianai, et al.
Veröffentlicht: (2025)
von: Dong, Dota Tianai, et al.
Veröffentlicht: (2025)
Catch Me If You Can? Not Yet: LLMs Still Struggle to Imitate the Implicit Writing Styles of Everyday Authors
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2025)
SWAN: Semantic Watermarking with Abstract Meaning Representation
von: Ye, Ziping, et al.
Veröffentlicht: (2026)
von: Ye, Ziping, et al.
Veröffentlicht: (2026)
Retrieval Augmented Generation (RAG) and Beyond: A Comprehensive Survey on How to Make your LLMs use External Data More Wisely
von: Zhao, Siyun, et al.
Veröffentlicht: (2024)
von: Zhao, Siyun, et al.
Veröffentlicht: (2024)
More Than Bits: Multi-Envelope Double Binary Factorization for Extreme Quantization
von: Ichikawa, Yuma, et al.
Veröffentlicht: (2025)
von: Ichikawa, Yuma, et al.
Veröffentlicht: (2025)
Discrete Audio Tokens: More Than a Survey!
von: Mousavi, Pooneh, et al.
Veröffentlicht: (2025)
von: Mousavi, Pooneh, et al.
Veröffentlicht: (2025)
Mobile-Bench-v2: A More Realistic and Comprehensive Benchmark for VLM-based Mobile Agents
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
From Struggle (06-2024) to Mastery (02-2025) LLMs Conquer Advanced Algorithm Exams and Pave the Way for Editorial Generation
von: Dumitran, Adrian Marius, et al.
Veröffentlicht: (2025)
von: Dumitran, Adrian Marius, et al.
Veröffentlicht: (2025)
PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2024)
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2024)
Rethinking Layer Redundancy: Calibration Matters More Than Search in LLM Depth Pruning
von: Kim, Minkyu, et al.
Veröffentlicht: (2026)
von: Kim, Minkyu, et al.
Veröffentlicht: (2026)
MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning
von: Rüdiger, Sten, et al.
Veröffentlicht: (2026)
von: Rüdiger, Sten, et al.
Veröffentlicht: (2026)
CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
von: Shen, Junzhe, et al.
Veröffentlicht: (2026)
von: Shen, Junzhe, et al.
Veröffentlicht: (2026)
Safety Alignment Should Be Made More Than Just A Few Attention Heads
von: Huang, Chao, et al.
Veröffentlicht: (2025)
von: Huang, Chao, et al.
Veröffentlicht: (2025)
More Than One Teacher: Adaptive Multi-Guidance Policy Optimization for Diverse Exploration
von: Yuan, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Yuan, Xiaoyang, et al.
Veröffentlicht: (2025)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
Multimodal LLMs Struggle with Basic Visual Network Analysis: a VNA Benchmark
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
AraSTEM: A Native Arabic Multiple Choice Question Benchmark for Evaluating LLMs Knowledge In STEM Subjects
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2024)
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2024)
Language Models Struggle to Use Representations Learned In-Context
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Algebra of Meaning: Why Machines Need Montague More Than Moore's Law
von: Jeong, Cheonkam, et al.
Veröffentlicht: (2025) -
LLMs Know More Than Words: A Genre Study with Syntax, Metaphor & Phonetics
von: Shi, Weiye, et al.
Veröffentlicht: (2025) -
From Instruction to Output: The Role of Prompting in Modern NLG
von: Zaib, Munazza, et al.
Veröffentlicht: (2026) -
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2026) -
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
von: McGovern, Hope, et al.
Veröffentlicht: (2026)