Beyond Known Facts: Generating Unseen Temporal Knowledge to Address Data Contamination in LLM Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Amalvy, Arthur, Huang, Hen-Hsen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing
by: Amalvy, Arthur, et al.
Published: (2026)
by: Amalvy, Arthur, et al.
Published: (2026)
Democratizing LLM Efficiency: From Hyperscale Optimizations to Universal Deployability
by: Huang, Hen-Hsen
Published: (2025)
by: Huang, Hen-Hsen
Published: (2025)
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks
by: Chan, Brian J, et al.
Published: (2024)
by: Chan, Brian J, et al.
Published: (2024)
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
by: Lin, Wei-Hsiang, et al.
Published: (2025)
by: Lin, Wei-Hsiang, et al.
Published: (2025)
Annotation Guidelines for Corpus Novelties: Part 2 -- Alias Resolution Version 1.0
by: Amalvy, Arthur, et al.
Published: (2024)
by: Amalvy, Arthur, et al.
Published: (2024)
Annotation Guidelines for Corpus Novelties: Part 1 -- Named Entity Recognition
by: Amalvy, Arthur, et al.
Published: (2024)
by: Amalvy, Arthur, et al.
Published: (2024)
Diagnosing Model Editing via Knowledge Spectrum
by: Pan, Tsung-Hsuan, et al.
Published: (2025)
by: Pan, Tsung-Hsuan, et al.
Published: (2025)
Strategy-Induct: Task-Level Strategy Induction for Instruction Generation
by: Chen, Po-Chun, et al.
Published: (2026)
by: Chen, Po-Chun, et al.
Published: (2026)
Co-Trained Retriever-Generator Framework for Question Generation in Earnings Calls
by: Juan, Yining, et al.
Published: (2024)
by: Juan, Yining, et al.
Published: (2024)
Diverge to Induce Prompting: Multi-Rationale Induction for Zero-Shot Reasoning
by: Chen, Po-Chun, et al.
Published: (2026)
by: Chen, Po-Chun, et al.
Published: (2026)
DINA: A Dual Defense Framework Against Internal Noise and External Attacks in Natural Language Processing
by: Chuang, Ko-Wei, et al.
Published: (2025)
by: Chuang, Ko-Wei, et al.
Published: (2025)
Learning to Rank Context for Named Entity Recognition Using a Synthetic Dataset
by: Amalvy, Arthur, et al.
Published: (2023)
by: Amalvy, Arthur, et al.
Published: (2023)
The Role of Natural Language Processing Tasks in Automatic Literary Character Network Construction
by: Amalvy, Arthur, et al.
Published: (2024)
by: Amalvy, Arthur, et al.
Published: (2024)
Renard: A Modular Pipeline for Extracting Character Networks from Narrative Texts
by: Amalvy, Arthur, et al.
Published: (2024)
by: Amalvy, Arthur, et al.
Published: (2024)
The Role of Global and Local Context in Named Entity Recognition
by: Amalvy, Arthur, et al.
Published: (2023)
by: Amalvy, Arthur, et al.
Published: (2023)
Personalized Graph-Empowered Large Language Model for Proactive Information Access
by: Chang, Chia Cheng, et al.
Published: (2026)
by: Chang, Chia Cheng, et al.
Published: (2026)
No One Fits All: From Fixed Prompting to Learned Routing in Multilingual LLMs
by: Wu, Wei-Chi, et al.
Published: (2026)
by: Wu, Wei-Chi, et al.
Published: (2026)
Pre-Finetuning with Impact Duration Awareness for Stock Movement Prediction
by: Chiu, Chr-Jr, et al.
Published: (2024)
by: Chiu, Chr-Jr, et al.
Published: (2024)
"Why" Has the Least Side Effect on Model Editing
by: Pan, Tsung-Hsuan, et al.
Published: (2024)
by: Pan, Tsung-Hsuan, et al.
Published: (2024)
Beyond Translation: LLM-Based Data Generation for Multilingual Fact-Checking
by: Chung, Yi-Ling, et al.
Published: (2025)
by: Chung, Yi-Ling, et al.
Published: (2025)
Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models
by: Wei, Sheng-Lun, et al.
Published: (2024)
by: Wei, Sheng-Lun, et al.
Published: (2024)
Visual Lifelog Retrieval through Captioning-Enhanced Interpretation
by: Shih, Yu-Fei, et al.
Published: (2025)
by: Shih, Yu-Fei, et al.
Published: (2025)
Efficient Beam Search for Large Language Models Using Trie-Based Decoding
by: Chan, Brian J, et al.
Published: (2025)
by: Chan, Brian J, et al.
Published: (2025)
AR$^2$: Adversarial Reinforcement Learning for Abstract Reasoning in Large Language Models
by: Yeh, Cheng-Kai, et al.
Published: (2025)
by: Yeh, Cheng-Kai, et al.
Published: (2025)
Enhancing Investment Opinion Ranking through Argument-Based Sentiment Analysis
by: Chen, Chung-Chi, et al.
Published: (2024)
by: Chen, Chung-Chi, et al.
Published: (2024)
On Large Language Models' Hallucination with Regard to Known Facts
by: Jiang, Che, et al.
Published: (2024)
by: Jiang, Che, et al.
Published: (2024)
Beyond Memorization: Testing LLM Reasoning on Unseen Theory of Computation Tasks
by: Shelat, Shlok, et al.
Published: (2026)
by: Shelat, Shlok, et al.
Published: (2026)
Using Contextually Aligned Online Reviews to Measure LLMs' Performance Disparities Across Language Varieties
by: Tang, Zixin, et al.
Published: (2025)
by: Tang, Zixin, et al.
Published: (2025)
ChronoFact: Timeline-based Temporal Fact Verification
by: Barik, Anab Maulana, et al.
Published: (2024)
by: Barik, Anab Maulana, et al.
Published: (2024)
Bilingual Evaluation of Language Models on General Knowledge in University Entrance Exams with Minimal Contamination
by: Salido, Eva Sánchez, et al.
Published: (2024)
by: Salido, Eva Sánchez, et al.
Published: (2024)
Bias in the Ear of the Listener: Assessing Sensitivity in Audio Language Models Across Linguistic, Demographic, and Positional Variations
by: Wei, Sheng-Lun, et al.
Published: (2026)
by: Wei, Sheng-Lun, et al.
Published: (2026)
See the Unseen: Better Context-Consistent Knowledge-Editing by Noises
by: Huang, Youcheng, et al.
Published: (2024)
by: Huang, Youcheng, et al.
Published: (2024)
LatestEval: Addressing Data Contamination in Language Model Evaluation through Dynamic and Time-Sensitive Test Construction
by: Li, Yucheng, et al.
Published: (2023)
by: Li, Yucheng, et al.
Published: (2023)
Knowledge-to-SQL: Enhancing SQL Generation with Data Expert LLM
by: Hong, Zijin, et al.
Published: (2024)
by: Hong, Zijin, et al.
Published: (2024)
From Data to Knowledge: Evaluating How Efficiently Language Models Learn Facts
by: Christoph, Daniel, et al.
Published: (2025)
by: Christoph, Daniel, et al.
Published: (2025)
DCR: Quantifying Data Contamination in LLMs Evaluation
by: Xu, Cheng, et al.
Published: (2025)
by: Xu, Cheng, et al.
Published: (2025)
Detecting LLM Fact-conflicting Hallucinations Enhanced by Temporal-logic-based Reasoning
by: Li, Ningke, et al.
Published: (2025)
by: Li, Ningke, et al.
Published: (2025)
Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory
by: Sun, Jingwei, et al.
Published: (2026)
by: Sun, Jingwei, et al.
Published: (2026)
Interconnected Kingdoms: Comparing 'A Song of Ice and Fire' Adaptations Across Media Using Complex Networks
by: Amalvy, Arthur, et al.
Published: (2024)
by: Amalvy, Arthur, et al.
Published: (2024)
Evaluating Large Language Model Capability in Vietnamese Fact-Checking Data Generation
by: To, Long Truong, et al.
Published: (2024)
by: To, Long Truong, et al.
Published: (2024)
Similar Items
-
Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing
by: Amalvy, Arthur, et al.
Published: (2026) -
Democratizing LLM Efficiency: From Hyperscale Optimizations to Universal Deployability
by: Huang, Hen-Hsen
Published: (2025) -
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks
by: Chan, Brian J, et al.
Published: (2024) -
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
by: Lin, Wei-Hsiang, et al.
Published: (2025) -
Annotation Guidelines for Corpus Novelties: Part 2 -- Alias Resolution Version 1.0
by: Amalvy, Arthur, et al.
Published: (2024)