Remember This Event That Year? Assessing Temporal Information and Reasoning in Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Beniwal, Himanshu, Patel, Dishant, D, Kowsik Nandagopan, Ladia, Hritik, Yadav, Ankit, Singh, Mayank |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Cross-lingual Editing in Multilingual Language Models
par: Beniwal, Himanshu, et autres
Publié: (2024)
par: Beniwal, Himanshu, et autres
Publié: (2024)
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
par: Yadav, Ankit, et autres
Publié: (2024)
par: Yadav, Ankit, et autres
Publié: (2024)
Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models
par: Beniwal, Himanshu, et autres
Publié: (2026)
par: Beniwal, Himanshu, et autres
Publié: (2026)
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
par: Sheth, Rajvee, et autres
Publié: (2025)
par: Sheth, Rajvee, et autres
Publié: (2025)
Char-mander Use mBackdoor! A Study of Cross-lingual Backdoor Attacks in Multilingual LLMs
par: Beniwal, Himanshu, et autres
Publié: (2025)
par: Beniwal, Himanshu, et autres
Publié: (2025)
COMMENTATOR: A Code-mixed Multilingual Text Annotation Framework
par: Sheth, Rajvee, et autres
Publié: (2024)
par: Sheth, Rajvee, et autres
Publié: (2024)
Beyond Monolingual Assumptions: A Survey of Code-Switched NLP in the Era of Large Language Models across Modalities
par: Sheth, Rajvee, et autres
Publié: (2025)
par: Sheth, Rajvee, et autres
Publié: (2025)
One Instruction Does Not Fit All: How Well Do Embeddings Align Personas and Instructions in Low-Resource Indian Languages?
par: Shah, Arya, et autres
Publié: (2026)
par: Shah, Arya, et autres
Publié: (2026)
Cause and Effect: Can Large Language Models Truly Understand Causality?
par: Ashwani, Swagata, et autres
Publié: (2024)
par: Ashwani, Swagata, et autres
Publié: (2024)
UNITYAI-GUARD: Pioneering Toxicity Detection Across Low-Resource Indian Languages
par: Beniwal, Himanshu, et autres
Publié: (2025)
par: Beniwal, Himanshu, et autres
Publié: (2025)
Peering Through Preferences: Unraveling Feedback Acquisition for Aligning Large Language Models
par: Bansal, Hritik, et autres
Publié: (2023)
par: Bansal, Hritik, et autres
Publié: (2023)
HALO: An Ontology for Representing and Categorizing Hallucinations in Large Language Models
par: Nananukul, Navapat, et autres
Publié: (2023)
par: Nananukul, Navapat, et autres
Publié: (2023)
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models
par: Tang, Zhisheng, et autres
Publié: (2024)
par: Tang, Zhisheng, et autres
Publié: (2024)
Error Taxonomy-Guided Prompt Optimization
par: Singh, Mayank, et autres
Publié: (2026)
par: Singh, Mayank, et autres
Publié: (2026)
Breaking mBad! Supervised Fine-tuning for Cross-Lingual Detoxification
par: Beniwal, Himanshu, et autres
Publié: (2025)
par: Beniwal, Himanshu, et autres
Publié: (2025)
Towards Optimizing and Evaluating a Retrieval Augmented QA Chatbot using LLMs with Human in the Loop
par: Afzal, Anum, et autres
Publié: (2024)
par: Afzal, Anum, et autres
Publié: (2024)
A Comprehensive Evaluation on Event Reasoning of Large Language Models
par: Tao, Zhengwei, et autres
Publié: (2024)
par: Tao, Zhengwei, et autres
Publié: (2024)
Navigating Semantic Relations: Challenges for Language Models in Abstract Common-Sense Reasoning
par: Gawin, Cole, et autres
Publié: (2025)
par: Gawin, Cole, et autres
Publié: (2025)
No Universal Prompt: Unifying Reasoning through Adaptive Prompting for Temporal Table Reasoning
par: Rajgaria, Abhishek, et autres
Publié: (2025)
par: Rajgaria, Abhishek, et autres
Publié: (2025)
LTLBench: Towards Benchmarks for Evaluating Temporal Reasoning in Large Language Models
par: Tang, Weizhi, et autres
Publié: (2024)
par: Tang, Weizhi, et autres
Publié: (2024)
AdapTime: Enabling Adaptive Temporal Reasoning in Large Language Models
par: Deng, Yimin, et autres
Publié: (2026)
par: Deng, Yimin, et autres
Publié: (2026)
DEPART: DEcomposing PARiTy across Multilingual LLMs
par: Uppadhyay, Manan, et autres
Publié: (2026)
par: Uppadhyay, Manan, et autres
Publié: (2026)
Large Language Models-guided Dynamic Adaptation for Temporal Knowledge Graph Reasoning
par: Wang, Jiapu, et autres
Publié: (2024)
par: Wang, Jiapu, et autres
Publié: (2024)
Assessing and Enhancing the Robustness of Large Language Models with Task Structure Variations for Logical Reasoning
par: Bao, Qiming, et autres
Publié: (2023)
par: Bao, Qiming, et autres
Publié: (2023)
An Evaluation of Estimative Uncertainty in Large Language Models
par: Tang, Zhisheng, et autres
Publié: (2024)
par: Tang, Zhisheng, et autres
Publié: (2024)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
par: Bansal, Hritik, et autres
Publié: (2024)
par: Bansal, Hritik, et autres
Publié: (2024)
Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations
par: Aravindan, Ashwath Vaithinathan, et autres
Publié: (2026)
par: Aravindan, Ashwath Vaithinathan, et autres
Publié: (2026)
Assessing Large Language Models on Climate Information
par: Bulian, Jannis, et autres
Publié: (2023)
par: Bulian, Jannis, et autres
Publié: (2023)
Multilingual Information Retrieval with a Monolingual Knowledge Base
par: Zhuang, Yingying, et autres
Publié: (2025)
par: Zhuang, Yingying, et autres
Publié: (2025)
Narrative-of-Thought: Improving Temporal Reasoning of Large Language Models via Recounted Narratives
par: Zhang, Xinliang Frederick, et autres
Publié: (2024)
par: Zhang, Xinliang Frederick, et autres
Publié: (2024)
TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models
par: Chu, Zheng, et autres
Publié: (2023)
par: Chu, Zheng, et autres
Publié: (2023)
What Really Controls Temporal Reasoning in Large Language Models: Tokenisation or Representation of Time?
par: Bhatia, Gagan, et autres
Publié: (2026)
par: Bhatia, Gagan, et autres
Publié: (2026)
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference
par: Shen, Ke, et autres
Publié: (2024)
par: Shen, Ke, et autres
Publié: (2024)
Smaller, Weaker, Yet Better: Training LLM Reasoners via Compute-Optimal Sampling
par: Bansal, Hritik, et autres
Publié: (2024)
par: Bansal, Hritik, et autres
Publié: (2024)
Towards Explainable Temporal Reasoning in Large Language Models: A Structure-Aware Generative Framework
par: Jiang, Zihao, et autres
Publié: (2025)
par: Jiang, Zihao, et autres
Publié: (2025)
SemEval-2026 Task 12: Abductive Event Reasoning: Towards Real-World Event Causal Inference for Large Language Models
par: Cao, Pengfei, et autres
Publié: (2026)
par: Cao, Pengfei, et autres
Publié: (2026)
Towards a More Inclusive AI: Progress and Perspectives in Large Language Model Training for the Sámi Language
par: Paul, Ronny, et autres
Publié: (2024)
par: Paul, Ronny, et autres
Publié: (2024)
SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions
par: Suvarna, Ashima, et autres
Publié: (2026)
par: Suvarna, Ashima, et autres
Publié: (2026)
Assessing and Understanding Creativity in Large Language Models
par: Zhao, Yunpu, et autres
Publié: (2024)
par: Zhao, Yunpu, et autres
Publié: (2024)
Assessing Political Bias in Large Language Models
par: Rettenberger, Luca, et autres
Publié: (2024)
par: Rettenberger, Luca, et autres
Publié: (2024)
Documents similaires
-
Cross-lingual Editing in Multilingual Language Models
par: Beniwal, Himanshu, et autres
Publié: (2024) -
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
par: Yadav, Ankit, et autres
Publié: (2024) -
Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models
par: Beniwal, Himanshu, et autres
Publié: (2026) -
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
par: Sheth, Rajvee, et autres
Publié: (2025) -
Char-mander Use mBackdoor! A Study of Cross-lingual Backdoor Attacks in Multilingual LLMs
par: Beniwal, Himanshu, et autres
Publié: (2025)