Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
Fuente:
arXiv
Saved in:
| Main Authors: | Oguz, Metehan, Bakman, Yavuz, Yaldiz, Duygu Nur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
by: Oğuz, Metehan, et al.
Published: (2024)
by: Oğuz, Metehan, et al.
Published: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
by: Wang, Nien-Shao, et al.
Published: (2025)
by: Wang, Nien-Shao, et al.
Published: (2025)
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)
by: Ziashahabi, Amir, et al.
Published: (2025)
Introducing TrGLUE and SentiTurca: A Comprehensive Benchmark for Turkish General Language Understanding and Sentiment Analysis
by: Altinok, Duygu
Published: (2025)
by: Altinok, Duygu
Published: (2025)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026)
by: Bakman, Yavuz, et al.
Published: (2026)
Does Context Matter? ContextualJudgeBench for Evaluating LLM-based Judges in Contextual Settings
by: Xu, Austin, et al.
Published: (2025)
by: Xu, Austin, et al.
Published: (2025)
Smooth Operators: LLMs Translating Imperfect Hints into Disfluency-Rich Transcripts
by: Altinok, Duygu
Published: (2025)
by: Altinok, Duygu
Published: (2025)
Whispering Context: Distilling Syntax and Semantics for Long Speech Transcripts
by: Altinok, Duygu
Published: (2025)
by: Altinok, Duygu
Published: (2025)
Optimal Turkish Subword Strategies at Scale: Systematic Evaluation of Data, Vocabulary, Morphology Interplay
by: Altinok, Duygu
Published: (2026)
by: Altinok, Duygu
Published: (2026)
SFR-RAG: Towards Contextually Faithful LLMs
by: Nguyen, Xuan-Phi, et al.
Published: (2024)
by: Nguyen, Xuan-Phi, et al.
Published: (2024)
COBIAS: Assessing the Contextual Reliability of Bias Benchmarks for Language Models
by: Govil, Priyanshul, et al.
Published: (2024)
by: Govil, Priyanshul, et al.
Published: (2024)
Can generative AI figure out figurative language? The influence of idioms on essay scoring by ChatGPT, Gemini, and Deepseek
by: Oğuz, Enis
Published: (2025)
by: Oğuz, Enis
Published: (2025)
Contextual Categorization Enhancement through LLMs Latent-Space
by: Bettouche, Zineddine, et al.
Published: (2024)
by: Bettouche, Zineddine, et al.
Published: (2024)
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
Rethinking the Understanding Ability across LLMs through Mutual Information
by: Wang, Shaojie, et al.
Published: (2025)
by: Wang, Shaojie, et al.
Published: (2025)
Understanding Mental Health Content on Social Media and Its Effect Towards Suicidal Ideation
by: Bhuiyan, Mohaiminul Islam, et al.
Published: (2025)
by: Bhuiyan, Mohaiminul Islam, et al.
Published: (2025)
Enhancing Contextual Understanding in Large Language Models through Contrastive Decoding
by: Zhao, Zheng, et al.
Published: (2024)
by: Zhao, Zheng, et al.
Published: (2024)
Bridging Context Gaps: Leveraging Coreference Resolution for Long Contextual Understanding
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
LOLgorithm: Integrating Semantic,Syntactic and Contextual Elements for Humor Classification
by: Khurana, Tanisha, et al.
Published: (2024)
by: Khurana, Tanisha, et al.
Published: (2024)
SportsMetrics: Blending Text and Numerical Data to Understand Information Fusion in LLMs
by: Hu, Yebowen, et al.
Published: (2024)
by: Hu, Yebowen, et al.
Published: (2024)
Assessing and Understanding Creativity in Large Language Models
by: Zhao, Yunpu, et al.
Published: (2024)
by: Zhao, Yunpu, et al.
Published: (2024)
Assessing the Capability of LLMs in Solving POSCOMP Questions
by: Viegas, Cayo, et al.
Published: (2025)
by: Viegas, Cayo, et al.
Published: (2025)
A Framework to Assess Multilingual Vulnerabilities of LLMs
by: Tang, Likai, et al.
Published: (2025)
by: Tang, Likai, et al.
Published: (2025)
Assessing LLMs Suitability for Knowledge Graph Completion
by: Iga, Vasile Ionut Remus, et al.
Published: (2024)
by: Iga, Vasile Ionut Remus, et al.
Published: (2024)
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
by: Matys, Piotr, et al.
Published: (2025)
by: Matys, Piotr, et al.
Published: (2025)
Contextual ASR Error Handling with LLMs Augmentation for Goal-Oriented Conversational AI
by: Asano, Yuya, et al.
Published: (2025)
by: Asano, Yuya, et al.
Published: (2025)
From Oracle to Noisy Context: Mitigating Contextual Exposure Bias in Speech-LLMs
by: Guo, Xiaoyong, et al.
Published: (2026)
by: Guo, Xiaoyong, et al.
Published: (2026)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
by: Zhou, Chulun, et al.
Published: (2025)
by: Zhou, Chulun, et al.
Published: (2025)
Fine-Tuning Medical Language Models for Enhanced Long-Contextual Understanding and Domain Expertise
by: Yang, Qimin, et al.
Published: (2024)
by: Yang, Qimin, et al.
Published: (2024)
Benchmarking Contextual and Paralinguistic Reasoning in Speech-LLMs: A Case Study with In-the-Wild Data
by: Wang, Qiongqiong, et al.
Published: (2025)
by: Wang, Qiongqiong, et al.
Published: (2025)
Causal Understanding by LLMs: The Role of Uncertainty
by: Lithgow-Serrano, Oscar, et al.
Published: (2025)
by: Lithgow-Serrano, Oscar, et al.
Published: (2025)
Understanding the Collapse of LLMs in Model Editing
by: Yang, Wanli, et al.
Published: (2024)
by: Yang, Wanli, et al.
Published: (2024)
Understanding Chain-of-Thought in LLMs through Information Theory
by: Ton, Jean-Francois, et al.
Published: (2024)
by: Ton, Jean-Francois, et al.
Published: (2024)
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?
by: Mavi, John, et al.
Published: (2024)
by: Mavi, John, et al.
Published: (2024)
MediEval: A Unified Medical Benchmark for Patient-Contextual and Knowledge-Grounded Reasoning in LLMs
by: Qu, Zhan, et al.
Published: (2025)
by: Qu, Zhan, et al.
Published: (2025)
Similar Items
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
by: Oğuz, Metehan, et al.
Published: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025) -
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024) -
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
by: Wang, Nien-Shao, et al.
Published: (2025) -
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024)