Guardado en:
| Autores principales: | He, Jie, Neville, Jennifer, Wan, Mengting, Yang, Longqi, Liu, Hui, Xu, Xiaofeng, Song, Xia, Pan, Jeff Z., Zhou, Pei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.18990 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Group Preference Alignment: Customized LLM Response Generation from In-Situ Conversations
por: Mondal, Ishani, et al.
Publicado: (2025)
por: Mondal, Ishani, et al.
Publicado: (2025)
Corporate Communication Companion (CCC): An LLM-empowered Writing Assistant for Workplace Social Media
por: Lu, Zhuoran, et al.
Publicado: (2024)
por: Lu, Zhuoran, et al.
Publicado: (2024)
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
por: Kamoi, Ryo, et al.
Publicado: (2026)
por: Kamoi, Ryo, et al.
Publicado: (2026)
WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback
por: Shi, Taiwei, et al.
Publicado: (2024)
por: Shi, Taiwei, et al.
Publicado: (2024)
Teaching Language Models To Gather Information Proactively
por: Huang, Tenghao, et al.
Publicado: (2025)
por: Huang, Tenghao, et al.
Publicado: (2025)
Pearl: Personalizing Large Language Model Writing Assistants with Generation-Calibrated Retrievers
por: Mysore, Sheshera, et al.
Publicado: (2023)
por: Mysore, Sheshera, et al.
Publicado: (2023)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
por: Kim, Zae Myung, et al.
Publicado: (2026)
por: Kim, Zae Myung, et al.
Publicado: (2026)
Beyond Output Critique: Self-Correction via Task Distillation
por: Rahmani, Hossein A., et al.
Publicado: (2026)
por: Rahmani, Hossein A., et al.
Publicado: (2026)
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning
por: Zou, Jiaru, et al.
Publicado: (2025)
por: Zou, Jiaru, et al.
Publicado: (2025)
ToolGen: Unified Tool Retrieval and Calling via Generation
por: Wang, Renxi, et al.
Publicado: (2024)
por: Wang, Renxi, et al.
Publicado: (2024)
Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
por: Zhu, Wenhong, et al.
Publicado: (2024)
por: Zhu, Wenhong, et al.
Publicado: (2024)
Interpretable User Satisfaction Estimation for Conversational Systems with Large Language Models
por: Lin, Ying-Chun, et al.
Publicado: (2024)
por: Lin, Ying-Chun, et al.
Publicado: (2024)
ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
por: Yue, Murong, et al.
Publicado: (2025)
por: Yue, Murong, et al.
Publicado: (2025)
DP-RFT: Learning to Generate Synthetic Text via Differentially Private Reinforcement Fine-Tuning
por: Xu, Fangyuan, et al.
Publicado: (2026)
por: Xu, Fangyuan, et al.
Publicado: (2026)
ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering
por: Liu, Marianne Menglin, et al.
Publicado: (2025)
por: Liu, Marianne Menglin, et al.
Publicado: (2025)
From Sufficiency to Reflection: Reinforcement-Guided Thinking Quality in Retrieval-Augmented Reasoning for LLMs
por: He, Jie, et al.
Publicado: (2025)
por: He, Jie, et al.
Publicado: (2025)
UniArk: Improving Generalisation and Consistency for Factual Knowledge Extraction through Debiasing
por: Yang, Yijun, et al.
Publicado: (2024)
por: Yang, Yijun, et al.
Publicado: (2024)
DiLA: Enhancing LLM Tool Learning with Differential Logic Layer
por: Zhang, Yu, et al.
Publicado: (2024)
por: Zhang, Yu, et al.
Publicado: (2024)
AWPO: Enhancing Tool-Use of Large Language Models through Adaptive Integration of Reasoning Rewards
por: Lin, Zihan, et al.
Publicado: (2025)
por: Lin, Zihan, et al.
Publicado: (2025)
TAPS: Tool-Augmented Personalisation via Structured Tagging
por: Taktasheva, Ekaterina, et al.
Publicado: (2025)
por: Taktasheva, Ekaterina, et al.
Publicado: (2025)
ToolExpander: Extending the Frontiers of Tool-Using Reinforcement Learning to Weak LLMs
por: Chen, Fu, et al.
Publicado: (2025)
por: Chen, Fu, et al.
Publicado: (2025)
The Use of Generative Search Engines for Knowledge Work and Complex Tasks
por: Suri, Siddharth, et al.
Publicado: (2024)
por: Suri, Siddharth, et al.
Publicado: (2024)
N2C2: Nearest Neighbor Enhanced Confidence Calibration for Cross-Lingual In-Context Learning
por: He, Jie, et al.
Publicado: (2025)
por: He, Jie, et al.
Publicado: (2025)
Re-Invoke: Tool Invocation Rewriting for Zero-Shot Tool Retrieval
por: Chen, Yanfei, et al.
Publicado: (2024)
por: Chen, Yanfei, et al.
Publicado: (2024)
Enhancing Network-on-Chip Design with Modern Simulation Tools
por: Dr. Ammar Khalid Al-Turk, et al.
Publicado: (2020)
por: Dr. Ammar Khalid Al-Turk, et al.
Publicado: (2020)
Reinforcing Human Behavior Simulation via Verbal Feedback
por: Sun, Weiwei, et al.
Publicado: (2026)
por: Sun, Weiwei, et al.
Publicado: (2026)
ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding
por: Xia, Heming, et al.
Publicado: (2026)
por: Xia, Heming, et al.
Publicado: (2026)
Non-Collaborative User Simulators for Tool Agents
por: Shim, Jeonghoon, et al.
Publicado: (2025)
por: Shim, Jeonghoon, et al.
Publicado: (2025)
Evaluating and Safeguarding the Adversarial Robustness of Retrieval-Based In-Context Learning
por: Yu, Simon, et al.
Publicado: (2024)
por: Yu, Simon, et al.
Publicado: (2024)
TnT-LLM: Text Mining at Scale with Large Language Models
por: Wan, Mengting, et al.
Publicado: (2024)
por: Wan, Mengting, et al.
Publicado: (2024)
Visual Reasoning through Tool-supervised Reinforcement Learning
por: Dong, Qihua, et al.
Publicado: (2026)
por: Dong, Qihua, et al.
Publicado: (2026)
Investigating Tool-Memory Conflicts in Tool-Augmented LLMs
por: Cheng, Jiali, et al.
Publicado: (2026)
por: Cheng, Jiali, et al.
Publicado: (2026)
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
por: Wang, Hongru, et al.
Publicado: (2025)
por: Wang, Hongru, et al.
Publicado: (2025)
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error
por: Wang, Boshi, et al.
Publicado: (2024)
por: Wang, Boshi, et al.
Publicado: (2024)
GTM: Simulating the World of Tools for AI Agents
por: Ren, Zhenzhen, et al.
Publicado: (2025)
por: Ren, Zhenzhen, et al.
Publicado: (2025)
Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions
por: Su, Junhao, et al.
Publicado: (2025)
por: Su, Junhao, et al.
Publicado: (2025)
PlanGenLLMs: A Modern Survey of LLM Planning Capabilities
por: Wei, Hui, et al.
Publicado: (2025)
por: Wei, Hui, et al.
Publicado: (2025)
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
por: Narajala, Vineeth Sai, et al.
Publicado: (2025)
por: Narajala, Vineeth Sai, et al.
Publicado: (2025)
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
por: Xia, Shihao, et al.
Publicado: (2026)
por: Xia, Shihao, et al.
Publicado: (2026)
SC-Bench: A Large-Scale Dataset for Smart Contract Auditing
por: Xia, Shihao, et al.
Publicado: (2024)
por: Xia, Shihao, et al.
Publicado: (2024)
Ejemplares similares
-
Group Preference Alignment: Customized LLM Response Generation from In-Situ Conversations
por: Mondal, Ishani, et al.
Publicado: (2025) -
Corporate Communication Companion (CCC): An LLM-empowered Writing Assistant for Workplace Social Media
por: Lu, Zhuoran, et al.
Publicado: (2024) -
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
por: Kamoi, Ryo, et al.
Publicado: (2026) -
WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback
por: Shi, Taiwei, et al.
Publicado: (2024) -
Teaching Language Models To Gather Information Proactively
por: Huang, Tenghao, et al.
Publicado: (2025)