Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bajpai, Ashutosh, Chakraborty, Tanmoy |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
par: Bajpai, Ashutosh, et autres
Publié: (2024)
par: Bajpai, Ashutosh, et autres
Publié: (2024)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
par: Bajpai, Ashutosh, et autres
Publié: (2026)
par: Bajpai, Ashutosh, et autres
Publié: (2026)
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
par: Bajpai, Ashutosh, et autres
Publié: (2026)
par: Bajpai, Ashutosh, et autres
Publié: (2026)
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
par: Levine, Lauren, et autres
Publié: (2026)
par: Levine, Lauren, et autres
Publié: (2026)
Predicting Word Similarity in Context with Referential Translation Machines
par: Biçici, Ergun
Publié: (2024)
par: Biçici, Ergun
Publié: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
par: Saji, Alan, et autres
Publié: (2025)
par: Saji, Alan, et autres
Publié: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
par: Ashuach, Tomer, et autres
Publié: (2025)
par: Ashuach, Tomer, et autres
Publié: (2025)
Question Answering Over Spatio-Temporal Knowledge Graph
par: Dai, Xinbang, et autres
Publié: (2024)
par: Dai, Xinbang, et autres
Publié: (2024)
Identifying Intensity of the Structure and Content in Tweets and the Discriminative Power of Attributes in Context with Referential Translation Machines
par: Biçici, Ergun
Publié: (2024)
par: Biçici, Ergun
Publié: (2024)
How Well Do LLMs Imitate Human Writing Style?
par: Jemama, Rebira, et autres
Publié: (2025)
par: Jemama, Rebira, et autres
Publié: (2025)
MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
par: Moosa, Ibraheem Muhammad, et autres
Publié: (2024)
par: Moosa, Ibraheem Muhammad, et autres
Publié: (2024)
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models
par: Fuad, Kazi Ahmed Asif, et autres
Publié: (2024)
par: Fuad, Kazi Ahmed Asif, et autres
Publié: (2024)
When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models
par: Bae, Ji Ho
Publié: (2026)
par: Bae, Ji Ho
Publié: (2026)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
par: Evelo, Bart, et autres
Publié: (2026)
par: Evelo, Bart, et autres
Publié: (2026)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
par: Simhi, Adi, et autres
Publié: (2024)
par: Simhi, Adi, et autres
Publié: (2024)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
par: Chang, Hoyeon, et autres
Publié: (2024)
par: Chang, Hoyeon, et autres
Publié: (2024)
GanitBench: A bi-lingual benchmark for evaluating mathematical reasoning in Vision Language Models
par: Bandooni, Ashutosh, et autres
Publié: (2025)
par: Bandooni, Ashutosh, et autres
Publié: (2025)
LLMs Are Not Scorers: Rethinking MT Evaluation with Generation-Based Methods
par: Cui, Hyang
Publié: (2025)
par: Cui, Hyang
Publié: (2025)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
par: Chehbouni, Khaoula, et autres
Publié: (2025)
par: Chehbouni, Khaoula, et autres
Publié: (2025)
MIRIAD: Augmenting LLMs with millions of medical query-response pairs
par: Zheng, Qinyue, et autres
Publié: (2025)
par: Zheng, Qinyue, et autres
Publié: (2025)
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
par: Qi, Jinhu, et autres
Publié: (2024)
par: Qi, Jinhu, et autres
Publié: (2024)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
par: Bouchekif, Abdessalam, et autres
Publié: (2026)
par: Bouchekif, Abdessalam, et autres
Publié: (2026)
LLMs and the Human Condition
par: Wallis, Peter
Publié: (2024)
par: Wallis, Peter
Publié: (2024)
ManagerBench: Evaluating the Safety-Pragmatism Trade-off in Autonomous LLMs
par: Simhi, Adi, et autres
Publié: (2025)
par: Simhi, Adi, et autres
Publié: (2025)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
par: CH-Wang, Sky, et autres
Publié: (2025)
par: CH-Wang, Sky, et autres
Publié: (2025)
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly
par: Hosseini, Peyman, et autres
Publié: (2024)
par: Hosseini, Peyman, et autres
Publié: (2024)
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
par: Šindelář, Pavel, et autres
Publié: (2025)
par: Šindelář, Pavel, et autres
Publié: (2025)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
par: Ewais, Ahmed, et autres
Publié: (2026)
par: Ewais, Ahmed, et autres
Publié: (2026)
MALT: Mechanistic Ablation of Lossy Translation in LLMs for a Low-Resource Language: Urdu
par: Bajwa, Taaha Saleem
Publié: (2025)
par: Bajwa, Taaha Saleem
Publié: (2025)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
par: Simhi, Adi, et autres
Publié: (2025)
par: Simhi, Adi, et autres
Publié: (2025)
Improving the OOD Performance of Closed-Source LLMs on NLI Through Strategic Data Selection
par: Stacey, Joe, et autres
Publié: (2025)
par: Stacey, Joe, et autres
Publié: (2025)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
par: Akavarapu, V. S. D. S. Mahesh, et autres
Publié: (2025)
par: Akavarapu, V. S. D. S. Mahesh, et autres
Publié: (2025)
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
par: Saha, Soumadeep, et autres
Publié: (2025)
par: Saha, Soumadeep, et autres
Publié: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025)
par: Peters, Sydney, et autres
Publié: (2025)
Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs
par: Cui, Wanyun, et autres
Publié: (2025)
par: Cui, Wanyun, et autres
Publié: (2025)
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
par: Mamidanna, Siddarth, et autres
Publié: (2025)
par: Mamidanna, Siddarth, et autres
Publié: (2025)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
par: Wang, Yuxia, et autres
Publié: (2024)
par: Wang, Yuxia, et autres
Publié: (2024)
Improving LLMs with a knowledge from databases
par: Máša, Petr
Publié: (2025)
par: Máša, Petr
Publié: (2025)
EduGuardBench: A Holistic Benchmark for Evaluating the Pedagogical Fidelity and Adversarial Safety of LLMs as Simulated Teachers
par: Jiang, Yilin, et autres
Publié: (2025)
par: Jiang, Yilin, et autres
Publié: (2025)
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
par: Sorstkins, Andrejs
Publié: (2025)
par: Sorstkins, Andrejs
Publié: (2025)
Documents similaires
-
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
par: Bajpai, Ashutosh, et autres
Publié: (2024) -
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
par: Bajpai, Ashutosh, et autres
Publié: (2026) -
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
par: Bajpai, Ashutosh, et autres
Publié: (2026) -
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
par: Levine, Lauren, et autres
Publié: (2026) -
Predicting Word Similarity in Context with Referential Translation Machines
par: Biçici, Ergun
Publié: (2024)