AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Shirley, Zhao, Shiyu, Huang, Qian, Huang, Kexin, Yasunaga, Michihiro, Cao, Kaidi, Ioannidis, Vassilis N., Subbian, Karthik, Leskovec, Jure, Zou, James |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
by: Wu, Shirley, et al.
Published: (2024)
by: Wu, Shirley, et al.
Published: (2024)
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
by: Wu, Shirley, et al.
Published: (2023)
by: Wu, Shirley, et al.
Published: (2023)
Large Language Models as Analogical Reasoners
by: Yasunaga, Michihiro, et al.
Published: (2023)
by: Yasunaga, Michihiro, et al.
Published: (2023)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
by: Huang, Qian, et al.
Published: (2023)
by: Huang, Qian, et al.
Published: (2023)
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
by: Chen, Tianlang, et al.
Published: (2026)
by: Chen, Tianlang, et al.
Published: (2026)
Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes
by: Bereket, Michael, et al.
Published: (2025)
by: Bereket, Michael, et al.
Published: (2025)
Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
by: Wu, Shirley, et al.
Published: (2025)
by: Wu, Shirley, et al.
Published: (2025)
Zero-shot causal learning
by: Nilforoshan, Hamed, et al.
Published: (2023)
by: Nilforoshan, Hamed, et al.
Published: (2023)
BioDiscoveryAgent: An AI Agent for Designing Genetic Perturbation Experiments
by: Roohani, Yusuf, et al.
Published: (2024)
by: Roohani, Yusuf, et al.
Published: (2024)
Learning over Positive and Negative Edges with Contrastive Message Passing
by: Pao-Huang, Peter, et al.
Published: (2026)
by: Pao-Huang, Peter, et al.
Published: (2026)
Automated Hypothesis Validation with Agentic Sequential Falsifications
by: Huang, Kexin, et al.
Published: (2025)
by: Huang, Kexin, et al.
Published: (2025)
AgentDR: Dynamic Recommendation with Implicit Item-Item Relations via LLM-based Agents
by: Yang, Mingdai, et al.
Published: (2025)
by: Yang, Mingdai, et al.
Published: (2025)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
by: Chen, Tianlang, et al.
Published: (2025)
by: Chen, Tianlang, et al.
Published: (2025)
Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures
by: Dwivedi, Vijay Prakash, et al.
Published: (2025)
by: Dwivedi, Vijay Prakash, et al.
Published: (2025)
Multimodal RewardBench: Holistic Evaluation of Reward Models for Vision Language Models
by: Yasunaga, Michihiro, et al.
Published: (2025)
by: Yasunaga, Michihiro, et al.
Published: (2025)
RelGNN: Composite Message Passing for Relational Deep Learning
by: Chen, Tianlang, et al.
Published: (2025)
by: Chen, Tianlang, et al.
Published: (2025)
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs
by: Xu, Jialiang, et al.
Published: (2024)
by: Xu, Jialiang, et al.
Published: (2024)
TimeGraphs: Graph-based Temporal Reasoning
by: Maheshwari, Paridhi, et al.
Published: (2024)
by: Maheshwari, Paridhi, et al.
Published: (2024)
CollabLLM: From Passive Responders to Active Collaborators
by: Wu, Shirley, et al.
Published: (2025)
by: Wu, Shirley, et al.
Published: (2025)
Avá
Published: (2012)
Published: (2012)
Large Language Models are Good Relational Learners
by: Wu, Fang, et al.
Published: (2025)
by: Wu, Fang, et al.
Published: (2025)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
by: Choudhary, Nurendra, et al.
Published: (2024)
by: Choudhary, Nurendra, et al.
Published: (2024)
Covering a Graph with Dense Subgraph Families, via Triangle-Rich Sets
by: Basu, Sabyasachi, et al.
Published: (2024)
by: Basu, Sabyasachi, et al.
Published: (2024)
Latency-Quality Routing for Functionally Equivalent Tools in LLM Agents
by: Chu, Kexin, et al.
Published: (2026)
by: Chu, Kexin, et al.
Published: (2026)
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning
by: Hu, Weihua, et al.
Published: (2024)
by: Hu, Weihua, et al.
Published: (2024)
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
by: Zou, Jiaru, et al.
Published: (2025)
by: Zou, Jiaru, et al.
Published: (2025)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
by: Wu, Tailin, et al.
Published: (2024)
by: Wu, Tailin, et al.
Published: (2024)
Inferring Dynamic Networks from Marginals with Iterative Proportional Fitting
by: Chang, Serina, et al.
Published: (2024)
by: Chang, Serina, et al.
Published: (2024)
Asynchronous Tool Usage for Real-Time Agents
by: Ginart, Antonio A., et al.
Published: (2024)
by: Ginart, Antonio A., et al.
Published: (2024)
R2-Router: A New Paradigm for LLM Routing with Reasoning
by: Xue, Jiaqi, et al.
Published: (2026)
by: Xue, Jiaqi, et al.
Published: (2026)
Towards Practical Tool Usage for Continually Learning LLMs
by: Huang, Jerry, et al.
Published: (2024)
by: Huang, Jerry, et al.
Published: (2024)
Tool-R0: Self-Evolving LLM Agents for Tool-Learning from Zero Data
by: Acikgoz, Emre Can, et al.
Published: (2026)
by: Acikgoz, Emre Can, et al.
Published: (2026)
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2024)
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2024)
Learning Efficient Positional Encodings with Graph Neural Networks
by: Kanatsoulis, Charilaos I., et al.
Published: (2025)
by: Kanatsoulis, Charilaos I., et al.
Published: (2025)
HumanLM: Simulating Users with State Alignment Beats Response Imitation
by: Wu, Shirley, et al.
Published: (2026)
by: Wu, Shirley, et al.
Published: (2026)
Evaluating Privilege Usage of Agents with Real-World Tools
by: Zhang, Quan, et al.
Published: (2026)
by: Zhang, Quan, et al.
Published: (2026)
CACTUS: Chemistry Agent Connecting Tool-Usage to Science
by: McNaughton, Andrew D., et al.
Published: (2024)
by: McNaughton, Andrew D., et al.
Published: (2024)
The Kinetics of Reasoning: How Chain-of-Thought Shapes Learning in Transformers?
by: Pengmei, Zihan, et al.
Published: (2025)
by: Pengmei, Zihan, et al.
Published: (2025)
Quantile Advantage Estimation: Stabilizing RLVR for LLM Reasoning
by: Wu, Junkang, et al.
Published: (2025)
by: Wu, Junkang, et al.
Published: (2025)
reWordBench: Benchmarking and Improving the Robustness of Reward Models with Transformed Inputs
by: Wu, Zhaofeng, et al.
Published: (2025)
by: Wu, Zhaofeng, et al.
Published: (2025)
Similar Items
-
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
by: Wu, Shirley, et al.
Published: (2024) -
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
by: Wu, Shirley, et al.
Published: (2023) -
Large Language Models as Analogical Reasoners
by: Yasunaga, Michihiro, et al.
Published: (2023) -
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
by: Huang, Qian, et al.
Published: (2023) -
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
by: Chen, Tianlang, et al.
Published: (2026)