RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Xiang, Yao, Yuguang, Zhang, Qi, Dong, Kaiwen, Baidya, Avinash, Guo, Ruocheng, Hasson, Hilaf, Das, Kamalika |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
by: Guo, Ruocheng, et al.
Published: (2026)
by: Guo, Ruocheng, et al.
Published: (2026)
Node-Level Uncertainty Estimation in LLM-Generated SQL
by: Hasson, Hilaf, et al.
Published: (2025)
by: Hasson, Hilaf, et al.
Published: (2025)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
by: Baidya, Avinash, et al.
Published: (2025)
by: Baidya, Avinash, et al.
Published: (2025)
OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration
by: Li, Shijun, et al.
Published: (2025)
by: Li, Shijun, et al.
Published: (2025)
ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
by: Li, Dawei, et al.
Published: (2026)
by: Li, Dawei, et al.
Published: (2026)
Customizing Language Model Responses with Contrastive In-Context Learning
by: Gao, Xiang, et al.
Published: (2024)
by: Gao, Xiang, et al.
Published: (2024)
Goal-Conditioned Supervised Learning for Multi-Objective Recommendation
by: Li, Shijun, et al.
Published: (2024)
by: Li, Shijun, et al.
Published: (2024)
Decomposing Epistemic Uncertainty for Causal Decision Making
by: Rahman, Md Musfiqur, et al.
Published: (2026)
by: Rahman, Md Musfiqur, et al.
Published: (2026)
HyQE: Ranking Contexts with Hypothetical Query Embeddings
by: Zhou, Weichao, et al.
Published: (2024)
by: Zhou, Weichao, et al.
Published: (2024)
SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models
by: Gao, Xiang, et al.
Published: (2024)
by: Gao, Xiang, et al.
Published: (2024)
Theoretical Guarantees of Learning Ensembling Strategies with Applications to Time Series Forecasting
by: Hasson, Hilaf, et al.
Published: (2023)
by: Hasson, Hilaf, et al.
Published: (2023)
Learning to Search Effective Example Sequences for In-Context Learning
by: Gao, Xiang, et al.
Published: (2025)
by: Gao, Xiang, et al.
Published: (2025)
REMem: Reasoning with Episodic Memory in Language Agent
by: Shu, Yiheng, et al.
Published: (2026)
by: Shu, Yiheng, et al.
Published: (2026)
Transaction Categorization with Relational Deep Learning in QuickBooks
by: Dong, Kaiwen, et al.
Published: (2025)
by: Dong, Kaiwen, et al.
Published: (2025)
Jailbreaks on Vision Language Model via Multimodal Reasoning
by: Noheria, Aarush, et al.
Published: (2026)
by: Noheria, Aarush, et al.
Published: (2026)
ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding
by: Liu, Xiao, et al.
Published: (2026)
by: Liu, Xiao, et al.
Published: (2026)
A Search for Good Pseudo-random Number Generators : Survey and Empirical Studies
by: Bhattacharjee, Kamalika, et al.
Published: (2018)
by: Bhattacharjee, Kamalika, et al.
Published: (2018)
PassiveQA: A Three-Action Framework for Epistemically Calibrated Question Answering via Supervised Finetuning
by: Baidya, Madhav S
Published: (2026)
by: Baidya, Madhav S
Published: (2026)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
by: Yakolli, Nivedan, et al.
Published: (2025)
by: Yakolli, Nivedan, et al.
Published: (2025)
Projectional Coderivatives and Calculus Rules
by: Yao, Wenfang, et al.
Published: (2022)
by: Yao, Wenfang, et al.
Published: (2022)
RuleAgent: Discovering Rules for Recommendation Denoising with Autonomous Language Agents
by: Wang, Zongwei, et al.
Published: (2025)
by: Wang, Zongwei, et al.
Published: (2025)
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning
by: Liu, Jiaqi, et al.
Published: (2025)
by: Liu, Jiaqi, et al.
Published: (2025)
Déjà Vu Memorization in Vision-Language Models
by: Jayaraman, Bargav, et al.
Published: (2024)
by: Jayaraman, Bargav, et al.
Published: (2024)
Detecting the Machine: A Comprehensive Benchmark of AI-Generated Text Detectors Across Architectures, Domains, and Adversarial Conditions
by: Baidya, Madhav S., et al.
Published: (2026)
by: Baidya, Madhav S., et al.
Published: (2026)
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
Learning When to Sample: Confidence-Aware Self-Consistency for Efficient LLM Chain-of-Thought Reasoning
by: Xiong, Juming, et al.
Published: (2026)
by: Xiong, Juming, et al.
Published: (2026)
Improving Deep Regression with Tightness
by: Zhang, Shihao, et al.
Published: (2025)
by: Zhang, Shihao, et al.
Published: (2025)
Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents
by: Jaiswal, Raj, et al.
Published: (2024)
by: Jaiswal, Raj, et al.
Published: (2024)
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
by: Zhang, Jiaxin, et al.
Published: (2023)
by: Zhang, Jiaxin, et al.
Published: (2023)
Discriminant Distance-Aware Representation on Deterministic Uncertainty Quantification Methods
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
Results for response and reliability-based optimization
by: Baidya, Sanjay
Published: (2026)
by: Baidya, Sanjay
Published: (2026)
Optimal parameters and structural responses for response and reliability based optimization
by: Baidya, Sanjay
Published: (2025)
by: Baidya, Sanjay
Published: (2025)
JExplore: Design Space Exploration Tool for Nvidia Jetson Boards
by: Kutukcu, Basar, et al.
Published: (2025)
by: Kutukcu, Basar, et al.
Published: (2025)
Reconstructing Abelian Varieties via Model Theory
by: Castle, Benjamin, et al.
Published: (2025)
by: Castle, Benjamin, et al.
Published: (2025)
MLLM-Tool: A Multimodal Large Language Model For Tool Agent Learning
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
RulePilot: An LLM-Powered Agent for Security Rule Generation
by: Wang, Hongtai, et al.
Published: (2025)
by: Wang, Hongtai, et al.
Published: (2025)
Revisando un órgano olvidado: Evaluación del timo en PET-CT
by: Daniel Hasson
Published: (2020)
by: Daniel Hasson
Published: (2020)
Similar Items
-
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
by: Guo, Ruocheng, et al.
Published: (2026) -
Node-Level Uncertainty Estimation in LLM-Generated SQL
by: Hasson, Hilaf, et al.
Published: (2025) -
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
by: Baidya, Avinash, et al.
Published: (2025) -
OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration
by: Li, Shijun, et al.
Published: (2025) -
ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
by: Li, Dawei, et al.
Published: (2026)