LHAW: Controllable Underspecification for Long-Horizon Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Pu, George, Lee, Michael S., Sehwag, Udari Madhushani, Lee, David J., Zhu, Bryan, Maurya, Yash, Raghavendra, Mohit, Xue, Yuan, Denton, Samuel Marc |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
In-Context Learning with Topological Information for Knowledge Graph Completion
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
by: Pathmanathan, Pankayaraj, et al.
Published: (2024)
by: Pathmanathan, Pankayaraj, et al.
Published: (2024)
Can LLMs be Scammed? A Baseline Measurement Study
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
by: Xu, Yuancheng, et al.
Published: (2024)
by: Xu, Yuancheng, et al.
Published: (2024)
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
by: Sehwag, Udari Madhushani, et al.
Published: (2026)
by: Sehwag, Udari Madhushani, et al.
Published: (2026)
Continual Learning of Domain Knowledge from Human Feedback in Text-to-SQL
by: Cook, Thomas, et al.
Published: (2025)
by: Cook, Thomas, et al.
Published: (2025)
Defensive Refusal Bias: How Safety Alignment Fails Cyber Defenders
by: Campbell, David, et al.
Published: (2026)
by: Campbell, David, et al.
Published: (2026)
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration
by: Karthikeyan, Harish, et al.
Published: (2025)
by: Karthikeyan, Harish, et al.
Published: (2025)
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
by: Chakraborty, Souradip, et al.
Published: (2025)
by: Chakraborty, Souradip, et al.
Published: (2025)
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
by: Lee, Michael S., et al.
Published: (2026)
by: Lee, Michael S., et al.
Published: (2026)
Program Machine Policy: Addressing Long-Horizon Tasks by Integrating Program Synthesis and State Machines
by: Lin, Yu-An, et al.
Published: (2023)
by: Lin, Yu-An, et al.
Published: (2023)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
by: Chiu, Yu Ying, et al.
Published: (2025)
by: Chiu, Yu Ying, et al.
Published: (2025)
Loop Polarity Analysis to Avoid Underspecification in Deep Learning
by: Martin, Jr., Donald, et al.
Published: (2023)
by: Martin, Jr., Donald, et al.
Published: (2023)
Agentic Aggregation for Parallel Scaling of Long-Horizon Agentic Tasks
by: Lee, Yoonsang, et al.
Published: (2026)
by: Lee, Yoonsang, et al.
Published: (2026)
A Note on Harmonic Underspecification in Log-Normal Trigonometric Regression
by: Gorczyca, Michael T.
Published: (2026)
by: Gorczyca, Michael T.
Published: (2026)
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
by: Wei, Boyi, et al.
Published: (2025)
by: Wei, Boyi, et al.
Published: (2025)
Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty
by: Singh, Joykirat, et al.
Published: (2026)
by: Singh, Joykirat, et al.
Published: (2026)
A Backbone for Long-Horizon Robot Task Understanding
by: Chen, Xiaoshuai, et al.
Published: (2024)
by: Chen, Xiaoshuai, et al.
Published: (2024)
O3D: Offline Data-driven Discovery and Distillation for Sequential Decision-Making with Large Language Models
by: Xiao, Yuchen, et al.
Published: (2023)
by: Xiao, Yuchen, et al.
Published: (2023)
Underspecification in Language Modeling Tasks: A Causality-Informed Study of Gendered Pronoun Resolution
by: McMilin, Emily
Published: (2022)
by: McMilin, Emily
Published: (2022)
CHD: Coupled Hierarchical Diffusion for Long-Horizon Tasks
by: Hao, Ce, et al.
Published: (2025)
by: Hao, Ce, et al.
Published: (2025)
Capable but Unreliable: Canonical Path Deviation as a Causal Mechanism of Agent Failure in Long-Horizon Tasks
by: Lee, Wilson Y.
Published: (2026)
by: Lee, Wilson Y.
Published: (2026)
Spatially Grounded Long-Horizon Task Planning in the Wild
by: Jung, Sehun, et al.
Published: (2026)
by: Jung, Sehun, et al.
Published: (2026)
PRInTS: Reward Modeling for Long-Horizon Information Seeking
by: Lee, Jaewoo, et al.
Published: (2025)
by: Lee, Jaewoo, et al.
Published: (2025)
Accounting for Underspecification in Statistical Claims of Model Superiority
by: Sanchez, Thomas, et al.
Published: (2025)
by: Sanchez, Thomas, et al.
Published: (2025)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
by: Erdogan, Lutfi Eren, et al.
Published: (2025)
by: Erdogan, Lutfi Eren, et al.
Published: (2025)
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
by: Yang, Chenyang, et al.
Published: (2025)
by: Yang, Chenyang, et al.
Published: (2025)
Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task
by: Jung, JungHo, et al.
Published: (2025)
by: Jung, JungHo, et al.
Published: (2025)
SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks
by: Wen, Yongyan, et al.
Published: (2024)
by: Wen, Yongyan, et al.
Published: (2024)
Assessing Robustness to Spurious Correlations in Post-Training Language Models
by: Shuieh, Julia, et al.
Published: (2025)
by: Shuieh, Julia, et al.
Published: (2025)
Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation
by: Klearman, Andrew, et al.
Published: (2026)
by: Klearman, Andrew, et al.
Published: (2026)
Do Pre-Trained Language Models Detect and Understand Semantic Underspecification? Ask the DUST!
by: Wildenburg, Frank, et al.
Published: (2024)
by: Wildenburg, Frank, et al.
Published: (2024)
Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment
by: D'Oosterlinck, Karel, et al.
Published: (2024)
by: D'Oosterlinck, Karel, et al.
Published: (2024)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
by: Singh, Raghavendra
Published: (2024)
by: Singh, Raghavendra
Published: (2024)
Revisiting the Superficial Alignment Hypothesis
by: Raghavendra, Mohit, et al.
Published: (2024)
by: Raghavendra, Mohit, et al.
Published: (2024)
Balancing the Budget: Understanding Trade-offs Between Supervised and Preference-Based Finetuning
by: Raghavendra, Mohit, et al.
Published: (2025)
by: Raghavendra, Mohit, et al.
Published: (2025)
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?
by: Lee, Hwiwon, et al.
Published: (2026)
by: Lee, Hwiwon, et al.
Published: (2026)
Plan-and-Write: Structure-Guided Length Control for LLMs without Model Retraining
by: Akinfaderin, Adewale, et al.
Published: (2025)
by: Akinfaderin, Adewale, et al.
Published: (2025)
Similar Items
-
In-Context Learning with Topological Information for Knowledge Graph Completion
by: Sehwag, Udari Madhushani, et al.
Published: (2024) -
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
by: Pathmanathan, Pankayaraj, et al.
Published: (2024) -
Can LLMs be Scammed? A Baseline Measurement Study
by: Sehwag, Udari Madhushani, et al.
Published: (2024) -
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
by: Sehwag, Udari Madhushani, et al.
Published: (2025) -
GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
by: Xu, Yuancheng, et al.
Published: (2024)