Distilling Tool Knowledge into Language Models via Back-Translated Traces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Xingyue, Hu, Xianglong, Ding, Zifeng, He, Yuan, Rishabh, Alzarooni, Waleed, Ye, Ziyu, Fan, Wendong, He, Bailan, Bo, Haige, Hu, Changran, Li, Guohao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EigentSearch-Q+: Enhancing Deep Research Agents with Structured Reasoning Tools
von: Zhang, Boer, et al.
Veröffentlicht: (2026)
von: Zhang, Boer, et al.
Veröffentlicht: (2026)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
Reasoning Compression with Mixed-Policy Distillation
von: Yang, Han, et al.
Veröffentlicht: (2026)
von: Yang, Han, et al.
Veröffentlicht: (2026)
PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
TCP: a Benchmark for Temporal Constraint-Based Planning
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
A Canonical Internal Model for Disturbance Rejection for a Class of Nonlinear Systems Subject to Trigonometric-Polynomial Disturbances
von: He, Changran, et al.
Veröffentlicht: (2026)
von: He, Changran, et al.
Veröffentlicht: (2026)
Exponential Tracking and Disturbance Rejection for Euler–Lagrange Systems With High‐Order Actuator Dynamics
von: Changran He, et al.
Veröffentlicht: (2024)
von: Changran He, et al.
Veröffentlicht: (2024)
Can Knowledge Graphs Make Large Language Models More Trustworthy? An Empirical Study Over Open-ended Question Answering
von: Sui, Yuan, et al.
Veröffentlicht: (2024)
von: Sui, Yuan, et al.
Veröffentlicht: (2024)
Self-Exploring Language Models for Explainable Link Forecasting on Temporal Graphs via Reinforcement Learning
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
Red Teaming GPT-4V: Are GPT-4V Safe Against Uni/Multi-Modal Jailbreak Attacks?
von: Chen, Shuo, et al.
Veröffentlicht: (2024)
von: Chen, Shuo, et al.
Veröffentlicht: (2024)
Predicate-Conditional Conformalized Answer Sets for Knowledge Graph Embeddings
von: Zhu, Yuqicheng, et al.
Veröffentlicht: (2025)
von: Zhu, Yuqicheng, et al.
Veröffentlicht: (2025)
TraceTrans: Translation and Spatial Tracing for Surgical Prediction
von: Luo, Xiyu, et al.
Veröffentlicht: (2025)
von: Luo, Xiyu, et al.
Veröffentlicht: (2025)
Supposedly Equivalent Facts That Aren't? Entity Frequency in Pre-training Induces Asymmetry in LLMs
von: He, Yuan, et al.
Veröffentlicht: (2025)
von: He, Yuan, et al.
Veröffentlicht: (2025)
Magnet: Multi-turn Tool-use Data Synthesis and Distillation via Graph Translation
von: Yin, Fan, et al.
Veröffentlicht: (2025)
von: Yin, Fan, et al.
Veröffentlicht: (2025)
Remote Sensing Image Classification with Decoupled Knowledge Distillation
von: He, Yaping, et al.
Veröffentlicht: (2025)
von: He, Yaping, et al.
Veröffentlicht: (2025)
Distillation-Enabled Knowledge Alignment for Generative Semantic Communications of AIGC Images
von: Hu, Jingzhi, et al.
Veröffentlicht: (2025)
von: Hu, Jingzhi, et al.
Veröffentlicht: (2025)
Generative AI as a Tool for Enhancing Reflective Learning in Students
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
Distillation-Enabled Knowledge Alignment Protocol for Semantic Communication in AI Agent Networks
von: Hu, Jingzhi, et al.
Veröffentlicht: (2025)
von: Hu, Jingzhi, et al.
Veröffentlicht: (2025)
Self-Evolution Knowledge Distillation for LLM-based Machine Translation
von: Song, Yuncheng, et al.
Veröffentlicht: (2024)
von: Song, Yuncheng, et al.
Veröffentlicht: (2024)
Modeling and Simulation of 2D Transducers Based on Suspended Graphene-Based Heterostructures in Nanoelectromechanical Pressure Sensors
von: Liu, Quan, et al.
Veröffentlicht: (2024)
von: Liu, Quan, et al.
Veröffentlicht: (2024)
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved Sampling
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
The influence of supplier financial violations on clients' environmental, social, and governance performance: Evidence from China's manufacturing industry
von: Jianhao Hu, et al.
Veröffentlicht: (2024)
von: Jianhao Hu, et al.
Veröffentlicht: (2024)
Unified Knowledge Distillation Framework: Fine-Grained Alignment and Geometric Relationship Preservation for Deep Face Recognition
von: Mishra, Durgesh, et al.
Veröffentlicht: (2025)
von: Mishra, Durgesh, et al.
Veröffentlicht: (2025)
Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers
von: Huang, Xingyue, et al.
Veröffentlicht: (2025)
von: Huang, Xingyue, et al.
Veröffentlicht: (2025)
Effective Knowledge Transfer for Multi-Task Recommendation Models
von: Cai, Guohao, et al.
Veröffentlicht: (2026)
von: Cai, Guohao, et al.
Veröffentlicht: (2026)
KDMOS:Knowledge Distillation for Motion Segmentation
von: Cao, Chunyu, et al.
Veröffentlicht: (2025)
von: Cao, Chunyu, et al.
Veröffentlicht: (2025)
Negative correlation between the Bern score and opening pressure in myelography positive spontaneous intracranial hypotension
von: Dan Zhang, et al.
Veröffentlicht: (2024)
von: Dan Zhang, et al.
Veröffentlicht: (2024)
WeMusic-Agent: Efficient Conversational Music Recommendation via Knowledge Internalization and Agentic Boundary Learning
von: Bi, Wendong, et al.
Veröffentlicht: (2025)
von: Bi, Wendong, et al.
Veröffentlicht: (2025)
TraceCoder: Towards Traceable ICD Coding via Multi-Source Knowledge Integration
von: Ren, Mucheng, et al.
Veröffentlicht: (2025)
von: Ren, Mucheng, et al.
Veröffentlicht: (2025)
On the Generalization of Knowledge Distillation: An Information-Theoretic View
von: Li, Bingying, et al.
Veröffentlicht: (2026)
von: Li, Bingying, et al.
Veröffentlicht: (2026)
Distilling Knowledge from Heterogeneous Architectures for Semantic Segmentation
von: Huang, Yanglin, et al.
Veröffentlicht: (2025)
von: Huang, Yanglin, et al.
Veröffentlicht: (2025)
LoRAP: Low-Rank Aggregation Prompting for Quantized Graph Neural Networks Training
von: Liu, Chenyu, et al.
Veröffentlicht: (2026)
von: Liu, Chenyu, et al.
Veröffentlicht: (2026)
Autonomous Chain-of-Thought Distillation for Graph-Based Fraud Detection
von: Li, Yuan, et al.
Veröffentlicht: (2026)
von: Li, Yuan, et al.
Veröffentlicht: (2026)
Creep Fracture Life Prediction of 9%Cr Heat‐Resistant Steel Based on the Back‐Propagation Artificial Neural Network Considering Microstructure Characteristics
von: Liwei Zhao, et al.
Veröffentlicht: (2024)
von: Liwei Zhao, et al.
Veröffentlicht: (2024)
Four ribbons of double-layer graphene suspending masses for NEMS applications
von: Fan, Xuge, et al.
Veröffentlicht: (2024)
von: Fan, Xuge, et al.
Veröffentlicht: (2024)
Neural Collapse Inspired Knowledge Distillation
von: Zhang, Shuoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Shuoxi, et al.
Veröffentlicht: (2024)
Trace forms on the cyclotomic Hecke algebras and cocenters of the cyclotomic Schur algebras
von: He, Zhekun, et al.
Veröffentlicht: (2022)
von: He, Zhekun, et al.
Veröffentlicht: (2022)
Efficient UAV Swarm-Based Multi-Task Federated Learning with Dynamic Task Knowledge Sharing
von: Yang, Yubo, et al.
Veröffentlicht: (2025)
von: Yang, Yubo, et al.
Veröffentlicht: (2025)
Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
Dynamic Frequency-Adaptive Knowledge Distillation for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EigentSearch-Q+: Enhancing Deep Research Agents with Structured Reasoning Tools
von: Zhang, Boer, et al.
Veröffentlicht: (2026) -
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024) -
Reasoning Compression with Mixed-Policy Distillation
von: Yang, Han, et al.
Veröffentlicht: (2026) -
PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model
von: Liu, Yilun, et al.
Veröffentlicht: (2024) -
TCP: a Benchmark for Temporal Constraint-Based Planning
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)