PALADIN: Self-Correcting Language Model Agents to Cure Tool-Failure Cases
Fuente:
arXiv
Saved in:
| Main Authors: | Vuddanti, Sri Vatsa, Shah, Aarav, Chittiprolu, Satwik Kumar, Song, Tony, Dev, Sunishchal, Zhu, Kevin, Chaudhary, Maheep |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Recoverability Has a Law: The ERR Measure for Tool-Augmented Agents
by: Vuddanti, Sri Vatsa, et al.
Published: (2026)
by: Vuddanti, Sri Vatsa, et al.
Published: (2026)
Amortized Latent Steering: Low-Cost Alternative to Test-Time Optimization
by: Egbuna, Nathan, et al.
Published: (2025)
by: Egbuna, Nathan, et al.
Published: (2025)
Broken Chains: The Cost of Incomplete Reasoning in LLMs
by: Su, Ian, et al.
Published: (2026)
by: Su, Ian, et al.
Published: (2026)
In-Context Environments Induce Evaluation-Awareness in Language Models
by: Chaudhary, Maheep
Published: (2026)
by: Chaudhary, Maheep
Published: (2026)
SALT: Steering Activations towards Leakage-free Thinking in Chain of Thought
by: Batra, Shourya, et al.
Published: (2025)
by: Batra, Shourya, et al.
Published: (2025)
Hydra: A Modular Architecture for Efficient Long-Context Reasoning
by: Chaudhary, Siddharth, et al.
Published: (2025)
by: Chaudhary, Siddharth, et al.
Published: (2025)
FRIT: Using Causal Importance to Improve Chain-of-Thought Faithfulness
by: Swaroop, Anand, et al.
Published: (2025)
by: Swaroop, Anand, et al.
Published: (2025)
SafetyNet: Detecting Harmful Outputs in LLMs by Modeling and Monitoring Deceptive Behaviors
by: Chaudhary, Maheep, et al.
Published: (2025)
by: Chaudhary, Maheep, et al.
Published: (2025)
Evaluating Open-Source Sparse Autoencoders on Disentangling Factual Knowledge in GPT-2 Small
by: Chaudhary, Maheep, et al.
Published: (2024)
by: Chaudhary, Maheep, et al.
Published: (2024)
The Ultra Slow-Roll Phase Of Warm Inflation In Braneworld Cosmology
by: Shah, Aarav
Published: (2025)
by: Shah, Aarav
Published: (2025)
Limits of Emergent Reasoning of Large Language Models in Agentic Frameworks for Deterministic Games
by: Su, Chris, et al.
Published: (2025)
by: Su, Chris, et al.
Published: (2025)
Sumudu Neural Operator for ODEs and PDEs
by: Zelenskiy, Ben, et al.
Published: (2025)
by: Zelenskiy, Ben, et al.
Published: (2025)
MemeCLIP: Leveraging CLIP Representations for Multimodal Meme Classification
by: Shah, Siddhant Bikram, et al.
Published: (2024)
by: Shah, Siddhant Bikram, et al.
Published: (2024)
Alignment-Constrained Dynamic Pruning for LLMs: Identifying and Preserving Alignment-Critical Circuits
by: Patel, Dev, et al.
Published: (2025)
by: Patel, Dev, et al.
Published: (2025)
DuoLens: A Framework for Robust Detection of Machine-Generated Multilingual Text and Code
by: Agrawal, Shriyansh, et al.
Published: (2025)
by: Agrawal, Shriyansh, et al.
Published: (2025)
Weight space Detection of Backdoors in LoRA Adapters
by: Merenciano, David Puertolas, et al.
Published: (2026)
by: Merenciano, David Puertolas, et al.
Published: (2026)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
by: Mao, Nathan, et al.
Published: (2026)
by: Mao, Nathan, et al.
Published: (2026)
Punctuation and Predicates in Language Models
by: Chauhan, Sonakshi, et al.
Published: (2025)
by: Chauhan, Sonakshi, et al.
Published: (2025)
Modeling and Predicting Multi-Turn Answer Instability in Large Language Models
by: He, Jiahang, et al.
Published: (2025)
by: He, Jiahang, et al.
Published: (2025)
Enhancing Quantum Diffusion Models with Pairwise Bell State Entanglement
by: Shah, Shivalee, et al.
Published: (2024)
by: Shah, Shivalee, et al.
Published: (2024)
A Few Bad Neurons: Isolating and Surgically Correcting Sycophancy
by: O'Brien, Claire, et al.
Published: (2026)
by: O'Brien, Claire, et al.
Published: (2026)
AgentChangeBench: A Multi-Dimensional Evaluation Framework for Goal-Shift Robustness in Conversational AI
by: Rana, Manik, et al.
Published: (2025)
by: Rana, Manik, et al.
Published: (2025)
PALADIN : Robust Neural Fingerprinting for Text-to-Image Diffusion Models
by: L, Murthy, et al.
Published: (2025)
by: L, Murthy, et al.
Published: (2025)
Emergent Persuasion: Will LLMs Persuade Without Being Prompted?
by: Chang, Vincent, et al.
Published: (2025)
by: Chang, Vincent, et al.
Published: (2025)
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
by: Sahay, Kenji, et al.
Published: (2025)
by: Sahay, Kenji, et al.
Published: (2025)
Inference-Time Chain-of-Thought Pruning with Latent Informativeness Signals
by: Li, Sophie, et al.
Published: (2025)
by: Li, Sophie, et al.
Published: (2025)
Optimizing Chain-of-Thought Confidence via Topological and Dirichlet Risk Analysis
by: More, Abhishek, et al.
Published: (2025)
by: More, Abhishek, et al.
Published: (2025)
Peek-a-Boo Reasoning: Contrastive Region Masking in MLLMs
by: Chaturvedi, Isha, et al.
Published: (2025)
by: Chaturvedi, Isha, et al.
Published: (2025)
Evaluation Awareness Scales Predictably in Open-Weights Large Language Models
by: Chaudhary, Maheep, et al.
Published: (2025)
by: Chaudhary, Maheep, et al.
Published: (2025)
ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs
by: Thomas, Rohan Subramanian, et al.
Published: (2026)
by: Thomas, Rohan Subramanian, et al.
Published: (2026)
Quantum-Evolutionary Neural Networks for Multi-Agent Federated Learning
by: Lala, Aarav, et al.
Published: (2025)
by: Lala, Aarav, et al.
Published: (2025)
XAgen: An Explainability Tool for Identifying and Correcting Failures in Multi-Agent Workflows
by: Wang, Xinru, et al.
Published: (2025)
by: Wang, Xinru, et al.
Published: (2025)
Judge Reliability Harness: Stress Testing the Reliability of LLM Judges
by: Dev, Sunishchal, et al.
Published: (2026)
by: Dev, Sunishchal, et al.
Published: (2026)
Practical Feasibility of Sustainable Software Engineering Tools and Techniques
by: Ghanta, Satwik, et al.
Published: (2026)
by: Ghanta, Satwik, et al.
Published: (2026)
Studying Cross-cluster Modularity in Neural Networks
by: Golechha, Satvik, et al.
Published: (2025)
by: Golechha, Satvik, et al.
Published: (2025)
MANATEE: Inference-Time Lightweight Diffusion Based Safety Defense for LLMs
by: Kan, Chun Yan Ryan, et al.
Published: (2026)
by: Kan, Chun Yan Ryan, et al.
Published: (2026)
A minireview on vermicompost and vermiwash as green pesticide for sustainable crop production: Approaches, applications, and advancements
by: Tanvi Singh, et al.
Published: (2024)
by: Tanvi Singh, et al.
Published: (2024)
JOSÉ CASIMIRO ULLOA BUCELO (1829-1891), EL PALADÍN DEL GREMIO MÉDICO
by: Oswaldo Salaverry
Published: (2010)
by: Oswaldo Salaverry
Published: (2010)
Full-Field Quantitative Visualization of Shock-Driven Pore Collapse and Failure Modes in PMMA
by: Lawlor, Barry P, et al.
Published: (2024)
by: Lawlor, Barry P, et al.
Published: (2024)
Self-Supervised Learning of Synapse Types from EM Images
by: Shetty, Aarav, et al.
Published: (2025)
by: Shetty, Aarav, et al.
Published: (2025)
Similar Items
-
Recoverability Has a Law: The ERR Measure for Tool-Augmented Agents
by: Vuddanti, Sri Vatsa, et al.
Published: (2026) -
Amortized Latent Steering: Low-Cost Alternative to Test-Time Optimization
by: Egbuna, Nathan, et al.
Published: (2025) -
Broken Chains: The Cost of Incomplete Reasoning in LLMs
by: Su, Ian, et al.
Published: (2026) -
In-Context Environments Induce Evaluation-Awareness in Language Models
by: Chaudhary, Maheep
Published: (2026) -
SALT: Steering Activations towards Leakage-free Thinking in Chain of Thought
by: Batra, Shourya, et al.
Published: (2025)