Fortify the Shortest Stave in Attention: Enhancing Context Awareness of Large Language Models for Effective Tool Use
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuhan, Lv, Ang, Lin, Ting-En, Chen, Changyu, Wu, Yuchuan, Huang, Fei, Li, Yongbin, Yan, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Models
von: Chen, Changyu, et al.
Veröffentlicht: (2024)
von: Chen, Changyu, et al.
Veröffentlicht: (2024)
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
von: Lin, Hongzhan, et al.
Veröffentlicht: (2024)
von: Lin, Hongzhan, et al.
Veröffentlicht: (2024)
Batch-ICL: Effective, Efficient, and Order-Agnostic In-Context Learning
von: Zhang, Kaiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Kaiyi, et al.
Veröffentlicht: (2024)
A Survey on Self-Evolution of Large Language Models
von: Tao, Zhengwei, et al.
Veröffentlicht: (2024)
von: Tao, Zhengwei, et al.
Veröffentlicht: (2024)
HoPE: A Novel Positional Encoding Without Long-Term Decay for Enhanced Context Awareness and Extrapolation
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
von: Si, Shuzheng, et al.
Veröffentlicht: (2023)
von: Si, Shuzheng, et al.
Veröffentlicht: (2023)
OpenOmni: Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignment and Real-Time Self-Aware Emotional Speech Synthesis
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
Study on the Wear and Failure Mechanism of Copper Cooling Staves in Large Blast Furnace
von: Songjian Shan, et al.
Veröffentlicht: (2026)
von: Songjian Shan, et al.
Veröffentlicht: (2026)
More is not always better? Enhancing Many-Shot In-Context Learning with Differentiated and Reweighting Objectives
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2025)
Reverse Preference Optimization for Complex Instruction Following
von: Huang, Xiang, et al.
Veröffentlicht: (2025)
von: Huang, Xiang, et al.
Veröffentlicht: (2025)
Improving Factual Consistency of News Summarization by Contrastive Preference Optimization
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
Messing with the Danzón / Zaidee Rose Stavely
von: Stavely, Zaidee Rose
von: Stavely, Zaidee Rose
Budget-Aware Tool-Use Enables Effective Agent Scaling
von: Liu, Tengxiao, et al.
Veröffentlicht: (2025)
von: Liu, Tengxiao, et al.
Veröffentlicht: (2025)
P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling
von: Zhang, Pinyi, et al.
Veröffentlicht: (2026)
von: Zhang, Pinyi, et al.
Veröffentlicht: (2026)
MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct
von: Luo, Run, et al.
Veröffentlicht: (2024)
von: Luo, Run, et al.
Veröffentlicht: (2024)
A Simple "Motivation" Can Enhance Reinforcement Finetuning of Large Reasoning Models
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
Audio-Maestro: Enhancing Large Audio-Language Models with Tool-Augmented Reasoning
von: Lee, Kuan-Yi, et al.
Veröffentlicht: (2025)
von: Lee, Kuan-Yi, et al.
Veröffentlicht: (2025)
ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
In-Context Reinforcement Learning for Tool Use in Large Language Models
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
Extreme Region Policy Distillation
von: Chen, Changyu, et al.
Veröffentlicht: (2026)
von: Chen, Changyu, et al.
Veröffentlicht: (2026)
Heat Transfer Analysis of Copper Stave in Blast Furnace Based on Slag Crust Characteristics
von: Hengbao Ma, et al.
Veröffentlicht: (2024)
von: Hengbao Ma, et al.
Veröffentlicht: (2024)
Fortified Sera and Their Use in Environmental Virology
von: Hoyt Jonathan L. and Margolin Aaron B
Veröffentlicht: (2000)
von: Hoyt Jonathan L. and Margolin Aaron B
Veröffentlicht: (2000)
ContextCache: Context-Aware Semantic Cache for Multi-Turn Queries in Large Language Models
von: Yan, Jianxin, et al.
Veröffentlicht: (2025)
von: Yan, Jianxin, et al.
Veröffentlicht: (2025)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
von: Lv, Ang, et al.
Veröffentlicht: (2024)
von: Lv, Ang, et al.
Veröffentlicht: (2024)
OmniCharacter: Towards Immersive Role-Playing Agents with Seamless Speech-Language Personality Interaction
von: Zhang, Haonan, et al.
Veröffentlicht: (2025)
von: Zhang, Haonan, et al.
Veröffentlicht: (2025)
Agentic Tool Use in Large Language Models
von: Hu, Jinchao, et al.
Veröffentlicht: (2026)
von: Hu, Jinchao, et al.
Veröffentlicht: (2026)
CPO: Addressing Reward Ambiguity in Role-playing Dialogue via Comparative Policy Optimization
von: Ye, Xinge, et al.
Veröffentlicht: (2025)
von: Ye, Xinge, et al.
Veröffentlicht: (2025)
Fortifying Ethical Boundaries in AI: Advanced Strategies for Enhancing Security in Large Language Models
von: He, Yunhong, et al.
Veröffentlicht: (2024)
von: He, Yunhong, et al.
Veröffentlicht: (2024)
On the Role of Attention Heads in Large Language Model Safety
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
von: Yan, Shaotian, et al.
Veröffentlicht: (2025)
von: Yan, Shaotian, et al.
Veröffentlicht: (2025)
Unlocking Large Language Model's Planning Capabilities with Maximum Diversity Fine-tuning
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents
von: Liang, Yijuan, et al.
Veröffentlicht: (2026)
von: Liang, Yijuan, et al.
Veröffentlicht: (2026)
Over the radio and into the woods a mexican community radio / Zaidee Stavely
von: Stavely, Zaidee
von: Stavely, Zaidee
Tag-Enriched Multi-Attention with Large Language Models for Cross-Domain Sequential Recommendation
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
Supervised Optimism Correction: Be Confident When LLMs Are Sure
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
SAGraph: A Large-Scale Social Graph Dataset with Comprehensive Context for Influencer Selection in Marketing
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2024)
ToolMind Technical Report: A Large-Scale, Reasoning-Enhanced Tool-Use Dataset
von: Yang, Chen, et al.
Veröffentlicht: (2025)
von: Yang, Chen, et al.
Veröffentlicht: (2025)
AWPO: Enhancing Tool-Use of Large Language Models through Adaptive Integration of Reasoning Rewards
von: Lin, Zihan, et al.
Veröffentlicht: (2025)
von: Lin, Zihan, et al.
Veröffentlicht: (2025)
An Analysis and Mitigation of the Reversal Curse
von: Lv, Ang, et al.
Veröffentlicht: (2023)
von: Lv, Ang, et al.
Veröffentlicht: (2023)
MOA: Multi-Objective Alignment for Role-Playing Agents
von: Liao, Chonghua, et al.
Veröffentlicht: (2025)
von: Liao, Chonghua, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Models
von: Chen, Changyu, et al.
Veröffentlicht: (2024) -
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
von: Lin, Hongzhan, et al.
Veröffentlicht: (2024) -
Batch-ICL: Effective, Efficient, and Order-Agnostic In-Context Learning
von: Zhang, Kaiyi, et al.
Veröffentlicht: (2024) -
A Survey on Self-Evolution of Large Language Models
von: Tao, Zhengwei, et al.
Veröffentlicht: (2024) -
HoPE: A Novel Positional Encoding Without Long-Term Decay for Enhanced Context Awareness and Extrapolation
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)