Adaptive Distraction: Probing LLM Contextual Robustness with Automated Tree Search

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Yanbo, Xu, Zixiang, Huang, Yue, Gao, Chujie, Wu, Siyuan, Ye, Jiayi, Chen, Pin-Yu, Chen, Xiuying, Zhang, Xiangliang
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912594976047104
author Wang, Yanbo
Xu, Zixiang
Huang, Yue
Gao, Chujie
Wu, Siyuan
Ye, Jiayi
Chen, Pin-Yu
Chen, Xiuying
Zhang, Xiangliang
author_facet Wang, Yanbo
Xu, Zixiang
Huang, Yue
Gao, Chujie
Wu, Siyuan
Ye, Jiayi
Chen, Pin-Yu
Chen, Xiuying
Zhang, Xiangliang
contents Large Language Models (LLMs) often struggle to maintain their original performance when faced with semantically coherent but task-irrelevant contextual information. Although prior studies have explored this issue using fixed-template or retrieval-based distractions, such static methods show limited effectiveness against contemporary models. To address this problem, we propose a dynamic distraction generation framework based on tree search, where the generation process is guided by model behavior. Without modifying the original question or answer, the method efficiently produces challenging adaptive distractions across multiple datasets, enabling systematic stress testing of LLMs' contextual robustness. Experiments on four benchmarks demonstrate that the generated distractions lead to an average performance drop of over 45\% for mainstream models. Further comparisons of mitigation strategies show that prompt-based optimization methods yield limited gains, whereas post-training approaches (e.g., DPO) significantly enhance the model's contextual robustness. The results indicate that these issues do not stem from knowledge deficits in LLMs, but from a fundamental inability to maintain consistent reasoning under contextual distraction, posing a major challenge to the reliability of LLMs in real-world applications. The code is publicly available at https://github.com/wyf23187/Adaptive_Distractions.
format Preprint
id arxiv_https___arxiv_org_abs_2502_01609
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Adaptive Distraction: Probing LLM Contextual Robustness with Automated Tree Search
Wang, Yanbo
Xu, Zixiang
Huang, Yue
Gao, Chujie
Wu, Siyuan
Ye, Jiayi
Chen, Pin-Yu
Chen, Xiuying
Zhang, Xiangliang
Computation and Language
Large Language Models (LLMs) often struggle to maintain their original performance when faced with semantically coherent but task-irrelevant contextual information. Although prior studies have explored this issue using fixed-template or retrieval-based distractions, such static methods show limited effectiveness against contemporary models. To address this problem, we propose a dynamic distraction generation framework based on tree search, where the generation process is guided by model behavior. Without modifying the original question or answer, the method efficiently produces challenging adaptive distractions across multiple datasets, enabling systematic stress testing of LLMs' contextual robustness. Experiments on four benchmarks demonstrate that the generated distractions lead to an average performance drop of over 45\% for mainstream models. Further comparisons of mitigation strategies show that prompt-based optimization methods yield limited gains, whereas post-training approaches (e.g., DPO) significantly enhance the model's contextual robustness. The results indicate that these issues do not stem from knowledge deficits in LLMs, but from a fundamental inability to maintain consistent reasoning under contextual distraction, posing a major challenge to the reliability of LLMs in real-world applications. The code is publicly available at https://github.com/wyf23187/Adaptive_Distractions.
title Adaptive Distraction: Probing LLM Contextual Robustness with Automated Tree Search
topic Computation and Language
url https://arxiv.org/abs/2502.01609