The Impact of Post-training on Data Contamination
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kocyigit, Muhammed Yusuf, Yildirim, Caglar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Overestimation in LLM Evaluation: A Controlled Large-Scale Study on Data Contamination's Impact on Machine Translation
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2025)
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2025)
Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models
von: Tao, Yongding, et al.
Veröffentlicht: (2025)
von: Tao, Yongding, et al.
Veröffentlicht: (2025)
Investigating Data Contamination for Pre-training Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training
von: Gwak, Minju, et al.
Veröffentlicht: (2026)
von: Gwak, Minju, et al.
Veröffentlicht: (2026)
Search-Time Data Contamination
von: Han, Ziwen, et al.
Veröffentlicht: (2025)
von: Han, Ziwen, et al.
Veröffentlicht: (2025)
Impact of Inaccurate Contamination Ratio on Robust Unsupervised Anomaly Detection
von: Masakuna, Jordan F., et al.
Veröffentlicht: (2024)
von: Masakuna, Jordan F., et al.
Veröffentlicht: (2024)
In Search for Architectures and Loss Functions in Multi-Objective Reinforcement Learning
von: Terekhov, Mikhail, et al.
Veröffentlicht: (2024)
von: Terekhov, Mikhail, et al.
Veröffentlicht: (2024)
Deep Positive-Unlabeled Anomaly Detection for Contaminated Unlabeled Data
von: Takahashi, Hiroshi, et al.
Veröffentlicht: (2024)
von: Takahashi, Hiroshi, et al.
Veröffentlicht: (2024)
RADAR: Mechanistic Pathways for Detecting Data Contamination in LLM Evaluation
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
Anomaly Detection with Adaptive and Aggressive Rejection for Contaminated Training Data
von: Lee, Jungi, et al.
Veröffentlicht: (2025)
von: Lee, Jungi, et al.
Veröffentlicht: (2025)
TSFMAudit: Data Contamination Auditing in Forecasting Time Series Foundation Models
von: Li, Hongkai, et al.
Veröffentlicht: (2026)
von: Li, Hongkai, et al.
Veröffentlicht: (2026)
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models
von: Golchin, Shahriar, et al.
Veröffentlicht: (2023)
von: Golchin, Shahriar, et al.
Veröffentlicht: (2023)
Differential Harm Propensity in Personalized LLM Agents: The Curious Case of Mental Health Disclosure
von: Yildirim, Caglar
Veröffentlicht: (2026)
von: Yildirim, Caglar
Veröffentlicht: (2026)
AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2026)
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2026)
BoA: Attention-aware Post-training Quantization without Backpropagation
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
Efficient Learning of Fuzzy Logic Systems for Large-Scale Data Using Deep Learning
von: Koklu, Ata, et al.
Veröffentlicht: (2024)
von: Koklu, Ata, et al.
Veröffentlicht: (2024)
The Role of Deep Learning Regularizations on Actors in Offline RL
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
A Generic Machine Learning Framework for Fully-Unsupervised Anomaly Detection with Contaminated Data
von: Ulmer, Markus, et al.
Veröffentlicht: (2023)
von: Ulmer, Markus, et al.
Veröffentlicht: (2023)
Generative Modeling of Networked Time-Series via Transformer Architectures
von: Elnady, Yusuf
Veröffentlicht: (2025)
von: Elnady, Yusuf
Veröffentlicht: (2025)
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training
von: Wang, Kevin, et al.
Veröffentlicht: (2026)
von: Wang, Kevin, et al.
Veröffentlicht: (2026)
Prioritized Replay for RL Post-training
von: Fatemi, Mehdi
Veröffentlicht: (2026)
von: Fatemi, Mehdi
Veröffentlicht: (2026)
Bayesian Kolmogorov Arnold Networks (Bayesian_KANs): A Probabilistic Approach to Enhance Accuracy and Interpretability
von: Hassan, Masoud Muhammed
Veröffentlicht: (2024)
von: Hassan, Masoud Muhammed
Veröffentlicht: (2024)
Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders
von: Jing, Yi, et al.
Veröffentlicht: (2026)
von: Jing, Yi, et al.
Veröffentlicht: (2026)
Beneficial Reasoning Behaviors in Agentic Search and Effective Post-training to Obtain Them
von: Jin, Jiahe, et al.
Veröffentlicht: (2025)
von: Jin, Jiahe, et al.
Veröffentlicht: (2025)
AdamS: Momentum Itself Can Be A Normalizer for LLM Pretraining and Post-training
von: Zhang, Huishuai, et al.
Veröffentlicht: (2025)
von: Zhang, Huishuai, et al.
Veröffentlicht: (2025)
RoCA: Robust Contrastive One-class Time Series Anomaly Detection with Contaminated Data
von: Mou, Xudong, et al.
Veröffentlicht: (2025)
von: Mou, Xudong, et al.
Veröffentlicht: (2025)
Beyond Surface-Level Similarity: Hierarchical Contamination Detection for Synthetic Training Data in Foundation Models
von: Mehta, Sushant
Veröffentlicht: (2025)
von: Mehta, Sushant
Veröffentlicht: (2025)
From Simulation to Enaction: Post-trained language models recognize and react to their own generations
von: G., Asvin, et al.
Veröffentlicht: (2026)
von: G., Asvin, et al.
Veröffentlicht: (2026)
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena
von: Luo, Haipeng, et al.
Veröffentlicht: (2024)
von: Luo, Haipeng, et al.
Veröffentlicht: (2024)
Automatic Pair Construction for Contrastive Post-training
von: Xu, Canwen, et al.
Veröffentlicht: (2023)
von: Xu, Canwen, et al.
Veröffentlicht: (2023)
Towards Effective Theory of LLMs: A Representation Learning Approach
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
State Contamination in Memory-Augmented LLM Agents
von: Wang, Yian, et al.
Veröffentlicht: (2026)
von: Wang, Yian, et al.
Veröffentlicht: (2026)
Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training
von: Hu, Pingbang, et al.
Veröffentlicht: (2026)
von: Hu, Pingbang, et al.
Veröffentlicht: (2026)
How Much Can We Forget about Data Contamination?
von: Bordt, Sebastian, et al.
Veröffentlicht: (2024)
von: Bordt, Sebastian, et al.
Veröffentlicht: (2024)
PBP: Post-training Backdoor Purification for Malware Classifiers
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2024)
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2024)
Post-training for Efficient Communication via Convention Formation
von: Hua, Yilun, et al.
Veröffentlicht: (2025)
von: Hua, Yilun, et al.
Veröffentlicht: (2025)
Efficient Post-training Quantization with FP8 Formats
von: Shen, Haihao, et al.
Veröffentlicht: (2023)
von: Shen, Haihao, et al.
Veröffentlicht: (2023)
Soft Contamination Means Benchmarks Test Shallow Generalization
von: Spiesberger, Ari, et al.
Veröffentlicht: (2026)
von: Spiesberger, Ari, et al.
Veröffentlicht: (2026)
Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Overestimation in LLM Evaluation: A Controlled Large-Scale Study on Data Contamination's Impact on Machine Translation
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2025) -
Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models
von: Tao, Yongding, et al.
Veröffentlicht: (2025) -
Investigating Data Contamination for Pre-training Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2024) -
LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training
von: Gwak, Minju, et al.
Veröffentlicht: (2026) -
Search-Time Data Contamination
von: Han, Ziwen, et al.
Veröffentlicht: (2025)