Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis
Fuente:
arXiv
Saved in:
| Main Authors: | Zou, Rui, Wei, Mengqi, Zhu, Yutao, Wen, Jirong, Zhao, Xin, Chen, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hide and Seek in Noise Labels: Noise-Robust Collaborative Active Learning with LLM-Powered Assistance
by: Yuan, Bo, et al.
Published: (2025)
by: Yuan, Bo, et al.
Published: (2025)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
by: Yu, Erxin, et al.
Published: (2025)
by: Yu, Erxin, et al.
Published: (2025)
Hide&Seek: Remove Image Watermarks with Negligible Cost via Pixel-wise Reconstruction
by: Chen, Huajie, et al.
Published: (2026)
by: Chen, Huajie, et al.
Published: (2026)
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
by: Iourovitski, Dmitri, et al.
Published: (2024)
by: Iourovitski, Dmitri, et al.
Published: (2024)
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
by: Park, Seongheon, et al.
Published: (2026)
by: Park, Seongheon, et al.
Published: (2026)
MACD: Multi-Agent Clinical Diagnosis with Self-Learned Knowledge for LLM
by: Li, Wenliang, et al.
Published: (2025)
by: Li, Wenliang, et al.
Published: (2025)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
by: Gan, Esther, et al.
Published: (2024)
by: Gan, Esther, et al.
Published: (2024)
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
by: He, Yidong, et al.
Published: (2026)
by: He, Yidong, et al.
Published: (2026)
Hide and Find: A Distributed Adversarial Attack on Federated Graph Learning
by: Liu, Jinshan, et al.
Published: (2026)
by: Liu, Jinshan, et al.
Published: (2026)
Hide and Seek: Investigating Redundancy in Earth Observation Imagery
by: Papazafeiropoulos, Tasos, et al.
Published: (2026)
by: Papazafeiropoulos, Tasos, et al.
Published: (2026)
GISA: A Benchmark for General Information-Seeking Assistant
by: Zhu, Yutao, et al.
Published: (2026)
by: Zhu, Yutao, et al.
Published: (2026)
Hide and Seek in Embedding Space: Geometry-based Steganography and Detection in Large Language Models
by: Westphal, Charles, et al.
Published: (2026)
by: Westphal, Charles, et al.
Published: (2026)
MentalSeek-Dx: Towards Progressive Hypothetico-Deductive Reasoning for Real-world Psychiatric Diagnosis
by: Sun, Xiao, et al.
Published: (2026)
by: Sun, Xiao, et al.
Published: (2026)
HOLMES: to Detect Adversarial Examples with Multiple Detectors
by: Wen, Jing
Published: (2024)
by: Wen, Jing
Published: (2024)
Improving LLMs' Generalized Reasoning Abilities by Graph Problems
by: Zhang, Qifan, et al.
Published: (2025)
by: Zhang, Qifan, et al.
Published: (2025)
How Do Diffusion Models Improve Adversarial Robustness?
by: Yuezhang, Liu, et al.
Published: (2025)
by: Yuezhang, Liu, et al.
Published: (2025)
Structure Enables Effective Self-Localization of Errors in LLMs
by: Samanta, Ankur, et al.
Published: (2026)
by: Samanta, Ankur, et al.
Published: (2026)
CreativeGame:Toward Mechanic-Aware Creative Game Generation
by: Ma, Hongnan, et al.
Published: (2026)
by: Ma, Hongnan, et al.
Published: (2026)
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Annealing Self-Distillation Rectification Improves Adversarial Training
by: Wu, Yu-Yu, et al.
Published: (2023)
by: Wu, Yu-Yu, et al.
Published: (2023)
Direct-Inverse Prompting: Analyzing LLMs' Discriminative Capacity in Self-Improving Generation
by: Ahn, Jihyun Janice, et al.
Published: (2024)
by: Ahn, Jihyun Janice, et al.
Published: (2024)
Zero-Shot Machine Unlearning with Proxy Adversarial Data Generation
by: Chen, Huiqiang, et al.
Published: (2025)
by: Chen, Huiqiang, et al.
Published: (2025)
Self-Improving Customer Review Response Generation Based on LLMs
by: Azov, Guy, et al.
Published: (2024)
by: Azov, Guy, et al.
Published: (2024)
Improving Fairness in LLMs Through Testing-Time Adversaries
by: Gregio, Isabela Pereira, et al.
Published: (2025)
by: Gregio, Isabela Pereira, et al.
Published: (2025)
Understanding and Improving Adversarial Robustness of Neural Probabilistic Circuits
by: Chen, Weixin, et al.
Published: (2025)
by: Chen, Weixin, et al.
Published: (2025)
Are the Hidden States Hiding Something? Testing the Limits of Factuality-Encoding Capabilities in LLMs
by: Servedio, Giovanni, et al.
Published: (2025)
by: Servedio, Giovanni, et al.
Published: (2025)
A Federated Learning Framework for Handling Subtype Confounding and Heterogeneity in Large-Scale Neuroimaging Diagnosis
by: Zhao, Xinglin, et al.
Published: (2025)
by: Zhao, Xinglin, et al.
Published: (2025)
Lost in Stories: Consistency Bugs in Long Story Generation by LLMs
by: Li, Junjie, et al.
Published: (2026)
by: Li, Junjie, et al.
Published: (2026)
Computer Environments Elicit General Agentic Intelligence in LLMs
by: Cheng, Daixuan, et al.
Published: (2026)
by: Cheng, Daixuan, et al.
Published: (2026)
Gradual Vigilance and Interval Communication: Enhancing Value Alignment in Multi-Agent Debates
by: Zou, Rui, et al.
Published: (2024)
by: Zou, Rui, et al.
Published: (2024)
ExpSeek: Self-Triggered Experience Seeking for Web Agents
by: Zhang, Wenyuan, et al.
Published: (2026)
by: Zhang, Wenyuan, et al.
Published: (2026)
StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs
by: Luo, Qijun, et al.
Published: (2025)
by: Luo, Qijun, et al.
Published: (2025)
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
by: Huang, Yue, et al.
Published: (2025)
by: Huang, Yue, et al.
Published: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
by: Belaire, Roman, et al.
Published: (2024)
by: Belaire, Roman, et al.
Published: (2024)
Adversarial Error Correction for Visual Autoregressive Generation
by: Bi, Ligong, et al.
Published: (2026)
by: Bi, Ligong, et al.
Published: (2026)
ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection
by: Li, Jiaqi, et al.
Published: (2025)
by: Li, Jiaqi, et al.
Published: (2025)
CorBenchX: Large-Scale Chest X-Ray Error Dataset and Vision-Language Model Benchmark for Report Error Correction
by: Zou, Jing, et al.
Published: (2025)
by: Zou, Jing, et al.
Published: (2025)
Comorbidity-Informed Transfer Learning for Neuro-developmental Disorder Diagnosis
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
QuantAgent: Seeking Holy Grail in Trading by Self-Improving Large Language Model
by: Wang, Saizhuo, et al.
Published: (2024)
by: Wang, Saizhuo, et al.
Published: (2024)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
by: Sun, Zhensu, et al.
Published: (2026)
by: Sun, Zhensu, et al.
Published: (2026)
Similar Items
-
Hide and Seek in Noise Labels: Noise-Robust Collaborative Active Learning with LLM-Powered Assistance
by: Yuan, Bo, et al.
Published: (2025) -
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
by: Yu, Erxin, et al.
Published: (2025) -
Hide&Seek: Remove Image Watermarks with Negligible Cost via Pixel-wise Reconstruction
by: Chen, Huajie, et al.
Published: (2026) -
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
by: Iourovitski, Dmitri, et al.
Published: (2024) -
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
by: Park, Seongheon, et al.
Published: (2026)