Don't Start What You Can't Finish: A Counterfactual Audit of Support-State Triage in LLM Agents
Fuente:
arXiv
Guardado en:
| Autor principal: | Unlu, Eren |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Preservation Is Not Enough for Width Growth: Regime-Sensitive Selection of Dense LM Warm Starts
por: Unlu, Eren
Publicado: (2026)
por: Unlu, Eren
Publicado: (2026)
Know When to Trust the Skill: Delayed Appraisal and Epistemic Vigilance for Single-Agent LLMs
por: Unlu, Eren
Publicado: (2026)
por: Unlu, Eren
Publicado: (2026)
Geotokens and Geotransformers
por: Unlu, Eren
Publicado: (2024)
por: Unlu, Eren
Publicado: (2024)
Can AI Assistants Know What They Don't Know?
por: Cheng, Qinyuan, et al.
Publicado: (2024)
por: Cheng, Qinyuan, et al.
Publicado: (2024)
Counterfactual Trace Auditing of LLM Agent Skills
por: Zhou, Xiaolin, et al.
Publicado: (2026)
por: Zhou, Xiaolin, et al.
Publicado: (2026)
xAI-Drop: Don't Use What You Cannot Explain
por: De Luca, Vincenzo Marco, et al.
Publicado: (2024)
por: De Luca, Vincenzo Marco, et al.
Publicado: (2024)
Implicit Intelligence -- Evaluating Agents on What Users Don't Say
por: Sirdeshmukh, Ved, et al.
Publicado: (2026)
por: Sirdeshmukh, Ved, et al.
Publicado: (2026)
What LLMs Think When You Don't Tell Them What to Think About?
por: Kwon, Yongchan, et al.
Publicado: (2026)
por: Kwon, Yongchan, et al.
Publicado: (2026)
What Papers Don't Tell You: Recovering Tacit Knowledge for Automated Paper Reproduction
por: Li, Lehui, et al.
Publicado: (2026)
por: Li, Lehui, et al.
Publicado: (2026)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
por: Park, Young-Jin, et al.
Publicado: (2025)
por: Park, Young-Jin, et al.
Publicado: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
por: Bajpai, Divya Jyoti, et al.
Publicado: (2025)
por: Bajpai, Divya Jyoti, et al.
Publicado: (2025)
Shadows Don't Lie and Lines Can't Bend! Generative Models don't know Projective Geometry...for now
por: Sarkar, Ayush, et al.
Publicado: (2023)
por: Sarkar, Ayush, et al.
Publicado: (2023)
Tell Me What You Don't Know: Enhancing Refusal Capabilities of Role-Playing Agents via Representation Space Analysis and Editing
por: Liu, Wenhao, et al.
Publicado: (2024)
por: Liu, Wenhao, et al.
Publicado: (2024)
Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents
por: Cheng, Liang, et al.
Publicado: (2026)
por: Cheng, Liang, et al.
Publicado: (2026)
Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems
por: Priyanshu, Aman, et al.
Publicado: (2026)
por: Priyanshu, Aman, et al.
Publicado: (2026)
s3: You Don't Need That Much Data to Train a Search Agent via RL
por: Jiang, Pengcheng, et al.
Publicado: (2025)
por: Jiang, Pengcheng, et al.
Publicado: (2025)
What We Don't C: Manifold Disentanglement for Structured Discovery
por: Rogers, Brian, et al.
Publicado: (2025)
por: Rogers, Brian, et al.
Publicado: (2025)
Don't Just Fine-tune the Agent, Tune the Environment
por: Lu, Siyuan, et al.
Publicado: (2025)
por: Lu, Siyuan, et al.
Publicado: (2025)
Don't Change My View: Ideological Bias Auditing in Large Language Models
por: Kröger, Paul, et al.
Publicado: (2025)
por: Kröger, Paul, et al.
Publicado: (2025)
I Don't Know You, But I Can Catch You: Real-Time Defense against Diverse Adversarial Patches for Object Detectors
por: Lin, Zijin, et al.
Publicado: (2024)
por: Lin, Zijin, et al.
Publicado: (2024)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
por: Parmar, Mihir, et al.
Publicado: (2022)
por: Parmar, Mihir, et al.
Publicado: (2022)
Don't Start from Scratch: Behavioral Refinement via Interpolant-based Policy Diffusion
por: Chen, Kaiqi, et al.
Publicado: (2024)
por: Chen, Kaiqi, et al.
Publicado: (2024)
Don't Make the LLM Read the Graph: Make the Graph Think
por: Sun, Yuqi, et al.
Publicado: (2026)
por: Sun, Yuqi, et al.
Publicado: (2026)
You Don't Need Prompt Engineering Anymore: The Prompting Inversion
por: Khan, Imran
Publicado: (2025)
por: Khan, Imran
Publicado: (2025)
Human Resilience in the AI Era -- What Machines Can't Replace
por: Liu, Shaoshan, et al.
Publicado: (2025)
por: Liu, Shaoshan, et al.
Publicado: (2025)
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
por: Kambhampati, Subbarao, et al.
Publicado: (2024)
por: Kambhampati, Subbarao, et al.
Publicado: (2024)
Reasoning Models Don't Always Say What They Think
por: Chen, Yanda, et al.
Publicado: (2025)
por: Chen, Yanda, et al.
Publicado: (2025)
mHC-lite: You Don't Need 20 Sinkhorn-Knopp Iterations
por: Yang, Yongyi, et al.
Publicado: (2026)
por: Yang, Yongyi, et al.
Publicado: (2026)
You Don't Know Until You Click:Automated GUI Testing for Production-Ready Software Evaluation
por: Bian, Yutong, et al.
Publicado: (2025)
por: Bian, Yutong, et al.
Publicado: (2025)
Formalize, Don't Optimize: The Heuristic Trap in LLM-Generated Combinatorial Solvers
por: Wang, Haoyu, et al.
Publicado: (2026)
por: Wang, Haoyu, et al.
Publicado: (2026)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
por: Mayne, Harry, et al.
Publicado: (2025)
por: Mayne, Harry, et al.
Publicado: (2025)
Don't Pay Attention
por: Hammoud, Mohammad, et al.
Publicado: (2025)
por: Hammoud, Mohammad, et al.
Publicado: (2025)
Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces
por: Zhang, Yilin, et al.
Publicado: (2026)
por: Zhang, Yilin, et al.
Publicado: (2026)
Equal Access, Unequal Interaction: A Counterfactual Audit of LLM Fairness
por: Amiri-Margavi, Alireza, et al.
Publicado: (2026)
por: Amiri-Margavi, Alireza, et al.
Publicado: (2026)
PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say
por: Zhang, Mingxuan, et al.
Publicado: (2026)
por: Zhang, Mingxuan, et al.
Publicado: (2026)
Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)
por: Lade, Ankit Hemant, et al.
Publicado: (2026)
por: Lade, Ankit Hemant, et al.
Publicado: (2026)
Time Blindness: Why Video-Language Models Can't See What Humans Can?
por: Upadhyay, Ujjwal, et al.
Publicado: (2025)
por: Upadhyay, Ujjwal, et al.
Publicado: (2025)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
por: Sun, Yiyou, et al.
Publicado: (2025)
por: Sun, Yiyou, et al.
Publicado: (2025)
Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation
por: Liu, Yanjiang, et al.
Publicado: (2026)
por: Liu, Yanjiang, et al.
Publicado: (2026)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
por: Khalifa, Muhammad, et al.
Publicado: (2024)
por: Khalifa, Muhammad, et al.
Publicado: (2024)
Ejemplares similares
-
Preservation Is Not Enough for Width Growth: Regime-Sensitive Selection of Dense LM Warm Starts
por: Unlu, Eren
Publicado: (2026) -
Know When to Trust the Skill: Delayed Appraisal and Epistemic Vigilance for Single-Agent LLMs
por: Unlu, Eren
Publicado: (2026) -
Geotokens and Geotransformers
por: Unlu, Eren
Publicado: (2024) -
Can AI Assistants Know What They Don't Know?
por: Cheng, Qinyuan, et al.
Publicado: (2024) -
Counterfactual Trace Auditing of LLM Agent Skills
por: Zhou, Xiaolin, et al.
Publicado: (2026)