AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Shirley, Zhao, Shiyu, Huang, Qian, Huang, Kexin, Yasunaga, Michihiro, Cao, Kaidi, Ioannidis, Vassilis N., Subbian, Karthik, Leskovec, Jure, Zou, James |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
di: Wu, Shirley, et al.
Pubblicazione: (2024)
di: Wu, Shirley, et al.
Pubblicazione: (2024)
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
di: Wu, Shirley, et al.
Pubblicazione: (2023)
di: Wu, Shirley, et al.
Pubblicazione: (2023)
Large Language Models as Analogical Reasoners
di: Yasunaga, Michihiro, et al.
Pubblicazione: (2023)
di: Yasunaga, Michihiro, et al.
Pubblicazione: (2023)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
di: Huang, Qian, et al.
Pubblicazione: (2023)
di: Huang, Qian, et al.
Pubblicazione: (2023)
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
di: Chen, Tianlang, et al.
Pubblicazione: (2026)
di: Chen, Tianlang, et al.
Pubblicazione: (2026)
Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes
di: Bereket, Michael, et al.
Pubblicazione: (2025)
di: Bereket, Michael, et al.
Pubblicazione: (2025)
Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
di: Wu, Shirley, et al.
Pubblicazione: (2025)
di: Wu, Shirley, et al.
Pubblicazione: (2025)
Zero-shot causal learning
di: Nilforoshan, Hamed, et al.
Pubblicazione: (2023)
di: Nilforoshan, Hamed, et al.
Pubblicazione: (2023)
BioDiscoveryAgent: An AI Agent for Designing Genetic Perturbation Experiments
di: Roohani, Yusuf, et al.
Pubblicazione: (2024)
di: Roohani, Yusuf, et al.
Pubblicazione: (2024)
Learning over Positive and Negative Edges with Contrastive Message Passing
di: Pao-Huang, Peter, et al.
Pubblicazione: (2026)
di: Pao-Huang, Peter, et al.
Pubblicazione: (2026)
Automated Hypothesis Validation with Agentic Sequential Falsifications
di: Huang, Kexin, et al.
Pubblicazione: (2025)
di: Huang, Kexin, et al.
Pubblicazione: (2025)
AgentDR: Dynamic Recommendation with Implicit Item-Item Relations via LLM-based Agents
di: Yang, Mingdai, et al.
Pubblicazione: (2025)
di: Yang, Mingdai, et al.
Pubblicazione: (2025)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures
di: Dwivedi, Vijay Prakash, et al.
Pubblicazione: (2025)
di: Dwivedi, Vijay Prakash, et al.
Pubblicazione: (2025)
Multimodal RewardBench: Holistic Evaluation of Reward Models for Vision Language Models
di: Yasunaga, Michihiro, et al.
Pubblicazione: (2025)
di: Yasunaga, Michihiro, et al.
Pubblicazione: (2025)
RelGNN: Composite Message Passing for Relational Deep Learning
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs
di: Xu, Jialiang, et al.
Pubblicazione: (2024)
di: Xu, Jialiang, et al.
Pubblicazione: (2024)
TimeGraphs: Graph-based Temporal Reasoning
di: Maheshwari, Paridhi, et al.
Pubblicazione: (2024)
di: Maheshwari, Paridhi, et al.
Pubblicazione: (2024)
CollabLLM: From Passive Responders to Active Collaborators
di: Wu, Shirley, et al.
Pubblicazione: (2025)
di: Wu, Shirley, et al.
Pubblicazione: (2025)
Avá
Pubblicazione: (2012)
Pubblicazione: (2012)
Large Language Models are Good Relational Learners
di: Wu, Fang, et al.
Pubblicazione: (2025)
di: Wu, Fang, et al.
Pubblicazione: (2025)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
Covering a Graph with Dense Subgraph Families, via Triangle-Rich Sets
di: Basu, Sabyasachi, et al.
Pubblicazione: (2024)
di: Basu, Sabyasachi, et al.
Pubblicazione: (2024)
Latency-Quality Routing for Functionally Equivalent Tools in LLM Agents
di: Chu, Kexin, et al.
Pubblicazione: (2026)
di: Chu, Kexin, et al.
Pubblicazione: (2026)
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning
di: Hu, Weihua, et al.
Pubblicazione: (2024)
di: Hu, Weihua, et al.
Pubblicazione: (2024)
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
di: Wu, Tailin, et al.
Pubblicazione: (2024)
di: Wu, Tailin, et al.
Pubblicazione: (2024)
Inferring Dynamic Networks from Marginals with Iterative Proportional Fitting
di: Chang, Serina, et al.
Pubblicazione: (2024)
di: Chang, Serina, et al.
Pubblicazione: (2024)
Asynchronous Tool Usage for Real-Time Agents
di: Ginart, Antonio A., et al.
Pubblicazione: (2024)
di: Ginart, Antonio A., et al.
Pubblicazione: (2024)
R2-Router: A New Paradigm for LLM Routing with Reasoning
di: Xue, Jiaqi, et al.
Pubblicazione: (2026)
di: Xue, Jiaqi, et al.
Pubblicazione: (2026)
Towards Practical Tool Usage for Continually Learning LLMs
di: Huang, Jerry, et al.
Pubblicazione: (2024)
di: Huang, Jerry, et al.
Pubblicazione: (2024)
Tool-R0: Self-Evolving LLM Agents for Tool-Learning from Zero Data
di: Acikgoz, Emre Can, et al.
Pubblicazione: (2026)
di: Acikgoz, Emre Can, et al.
Pubblicazione: (2026)
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models
di: Gutiérrez, Bernal Jiménez, et al.
Pubblicazione: (2024)
di: Gutiérrez, Bernal Jiménez, et al.
Pubblicazione: (2024)
Learning Efficient Positional Encodings with Graph Neural Networks
di: Kanatsoulis, Charilaos I., et al.
Pubblicazione: (2025)
di: Kanatsoulis, Charilaos I., et al.
Pubblicazione: (2025)
HumanLM: Simulating Users with State Alignment Beats Response Imitation
di: Wu, Shirley, et al.
Pubblicazione: (2026)
di: Wu, Shirley, et al.
Pubblicazione: (2026)
Evaluating Privilege Usage of Agents with Real-World Tools
di: Zhang, Quan, et al.
Pubblicazione: (2026)
di: Zhang, Quan, et al.
Pubblicazione: (2026)
CACTUS: Chemistry Agent Connecting Tool-Usage to Science
di: McNaughton, Andrew D., et al.
Pubblicazione: (2024)
di: McNaughton, Andrew D., et al.
Pubblicazione: (2024)
The Kinetics of Reasoning: How Chain-of-Thought Shapes Learning in Transformers?
di: Pengmei, Zihan, et al.
Pubblicazione: (2025)
di: Pengmei, Zihan, et al.
Pubblicazione: (2025)
Quantile Advantage Estimation: Stabilizing RLVR for LLM Reasoning
di: Wu, Junkang, et al.
Pubblicazione: (2025)
di: Wu, Junkang, et al.
Pubblicazione: (2025)
reWordBench: Benchmarking and Improving the Robustness of Reward Models with Transformed Inputs
di: Wu, Zhaofeng, et al.
Pubblicazione: (2025)
di: Wu, Zhaofeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
di: Wu, Shirley, et al.
Pubblicazione: (2024) -
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
di: Wu, Shirley, et al.
Pubblicazione: (2023) -
Large Language Models as Analogical Reasoners
di: Yasunaga, Michihiro, et al.
Pubblicazione: (2023) -
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
di: Huang, Qian, et al.
Pubblicazione: (2023) -
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
di: Chen, Tianlang, et al.
Pubblicazione: (2026)