Saved in:
| Main Authors: | Wang, Zilong, Yang, Jingfeng, Nag, Sreyashi, Varshney, Samarth, Tang, Xianfeng, Jiang, Haoming, Shang, Jingbo, Sarwar, Sheikh Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.20737 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REAPER: Reasoning based Retrieval Planning for Complex RAG Systems
by: Joshi, Ashutosh, et al.
Published: (2024)
by: Joshi, Ashutosh, et al.
Published: (2024)
Learning with Less: Knowledge Distillation from Large Language Models via Unlabeled Data
by: Li, Juanhui, et al.
Published: (2024)
by: Li, Juanhui, et al.
Published: (2024)
Query Brand Entity Linking in E-Commerce Search
by: Liu, Dong, et al.
Published: (2025)
by: Liu, Dong, et al.
Published: (2025)
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
by: Zeng, Jingying, et al.
Published: (2025)
by: Zeng, Jingying, et al.
Published: (2025)
Examples as the Prompt: A Scalable Approach for Efficient LLM Adaptation in E-Commerce
by: Zeng, Jingying, et al.
Published: (2025)
by: Zeng, Jingying, et al.
Published: (2025)
Situated Natural Language Explanations
by: Zhu, Zining, et al.
Published: (2023)
by: Zhu, Zining, et al.
Published: (2023)
Large Language Models in the Clinic: A Comprehensive Benchmark
by: Liu, Fenglin, et al.
Published: (2024)
by: Liu, Fenglin, et al.
Published: (2024)
A Tale of LLMs and Induced Small Proxies: Scalable Agents for Knowledge Mining
by: Zhang, Sipeng, et al.
Published: (2025)
by: Zhang, Sipeng, et al.
Published: (2025)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Memorize or Generalize? Evaluating LLM Code Generation with Code Rewriting
by: Zhang, Lizhe, et al.
Published: (2025)
by: Zhang, Lizhe, et al.
Published: (2025)
TIER: Trajectory-Invariant Execution Rewards for Multi-Step Tool Composition
by: Kulkarni, Anay, et al.
Published: (2026)
by: Kulkarni, Anay, et al.
Published: (2026)
Concept-Based Interpretability for Toxicity Detection
by: Garg, Samarth, et al.
Published: (2025)
by: Garg, Samarth, et al.
Published: (2025)
Protein Secondary Structure Prediction Using 3D Graphs and Relation-Aware Message Passing Transformers
by: Varshney, Disha, et al.
Published: (2025)
by: Varshney, Disha, et al.
Published: (2025)
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Characterized Diffusion Networks for Enhanced Autonomous Driving Trajectory Prediction
by: Li, Haoming
Published: (2024)
by: Li, Haoming
Published: (2024)
To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems
by: He, Pengfei, et al.
Published: (2025)
by: He, Pengfei, et al.
Published: (2025)
SE-Agent: Self-Evolution Trajectory Optimization in Multi-Step Reasoning with LLM-Based Agents
by: Lin, Jiaye, et al.
Published: (2025)
by: Lin, Jiaye, et al.
Published: (2025)
Can Language Models Follow Multiple Turns of Entangled Instructions?
by: Han, Chi, et al.
Published: (2025)
by: Han, Chi, et al.
Published: (2025)
LLM Based Bayesian Optimization for Prompt Search
by: Ballew, Adam, et al.
Published: (2025)
by: Ballew, Adam, et al.
Published: (2025)
Ares: Adaptive Reasoning Effort Selection for Efficient LLM Agents
by: Yang, Jingbo, et al.
Published: (2026)
by: Yang, Jingbo, et al.
Published: (2026)
Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling
by: Wang, Jiaxuan, et al.
Published: (2026)
by: Wang, Jiaxuan, et al.
Published: (2026)
Cuckoo: An IE Free Rider Hatched by Massive Nutrition in LLM's Nest
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Learning-based GNSS Uncertainty Quantification using Continuous-Time Factor Graph Optimization
by: Zhang, Haoming
Published: (2025)
by: Zhang, Haoming
Published: (2025)
Persona Inconstancy in Multi-Agent LLM Collaboration: Conformity, Confabulation, and Impersonation
by: Baltaji, Razan, et al.
Published: (2024)
by: Baltaji, Razan, et al.
Published: (2024)
To Call or Not to Call: A Framework to Assess and Optimize LLM Tool Calling
by: Wu, Qinyuan, et al.
Published: (2026)
by: Wu, Qinyuan, et al.
Published: (2026)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
by: Song, Yifan, et al.
Published: (2024)
by: Song, Yifan, et al.
Published: (2024)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL
by: Wang, Youting, et al.
Published: (2026)
by: Wang, Youting, et al.
Published: (2026)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
by: Sinha, Aarush, et al.
Published: (2026)
by: Sinha, Aarush, et al.
Published: (2026)
Trajectory Conditioned Cross-embodiment Skill Transfer
by: Tang, YuHang, et al.
Published: (2025)
by: Tang, YuHang, et al.
Published: (2025)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
by: Guo, Jian-Ting, et al.
Published: (2025)
by: Guo, Jian-Ting, et al.
Published: (2025)
PRO-CUA: Process-Reward Optimization for Computer Use Agents
by: He, Yifei, et al.
Published: (2026)
by: He, Yifei, et al.
Published: (2026)
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
by: Zhang, Zhiwei, et al.
Published: (2025)
by: Zhang, Zhiwei, et al.
Published: (2025)
Target Span Detection for Implicit Harmful Content
by: Jafari, Nazanin, et al.
Published: (2024)
by: Jafari, Nazanin, et al.
Published: (2024)
R^3: Replay, Reflection, and Ranking Rewards for LLM Reinforcement Learning
by: Jiang, Zhizheng, et al.
Published: (2026)
by: Jiang, Zhizheng, et al.
Published: (2026)
Reasoning with Graphs: Structuring Implicit Knowledge to Enhance LLMs Reasoning
by: Han, Haoyu, et al.
Published: (2025)
by: Han, Haoyu, et al.
Published: (2025)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
by: Mishra, Samarth, et al.
Published: (2025)
by: Mishra, Samarth, et al.
Published: (2025)
AutoRefine: From Trajectories to Reusable Expertise for Continual LLM Agent Refinement
by: Qiu, Libin, et al.
Published: (2026)
by: Qiu, Libin, et al.
Published: (2026)
Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Similar Items
-
REAPER: Reasoning based Retrieval Planning for Complex RAG Systems
by: Joshi, Ashutosh, et al.
Published: (2024) -
Learning with Less: Knowledge Distillation from Large Language Models via Unlabeled Data
by: Li, Juanhui, et al.
Published: (2024) -
Query Brand Entity Linking in E-Commerce Search
by: Liu, Dong, et al.
Published: (2025) -
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
by: Zeng, Jingying, et al.
Published: (2025) -
Examples as the Prompt: A Scalable Approach for Efficient LLM Adaptation in E-Commerce
by: Zeng, Jingying, et al.
Published: (2025)