DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Jung, Sunghee, Lee, Donghun, Lee, Shinbok, Seo, Gaeun, Lee, Daniel, Ko, Byeongil, Cho, Junrae, Kim, Kihyun, Kim, Eunggyun, Shin, Myeongcheol |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FunctionChat-Bench: Comprehensive Evaluation of Language Models' Generative Capabilities in Korean Tool-use Dialogs
di: Lee, Shinbok, et al.
Pubblicazione: (2024)
di: Lee, Shinbok, et al.
Pubblicazione: (2024)
Kanana: Compute-efficient Bilingual Language Models
di: Kanana LLM Team, et al.
Pubblicazione: (2025)
di: Kanana LLM Team, et al.
Pubblicazione: (2025)
SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety
di: Kim, Geon-Hyeong, et al.
Pubblicazione: (2025)
di: Kim, Geon-Hyeong, et al.
Pubblicazione: (2025)
Learning Design Preferences through Design Feature Extraction and Weighted Ensemble
di: Shin, Dongju, et al.
Pubblicazione: (2024)
di: Shin, Dongju, et al.
Pubblicazione: (2024)
Position Optimization of Passive Patch Based on Mode Contribution Factor for Vibration Attenuation of Asymmetric 1D Structure
di: Dongwoo Hong, et al.
Pubblicazione: (2024)
di: Dongwoo Hong, et al.
Pubblicazione: (2024)
Rethinking DPO: The Role of Rejected Responses in Preference Misalignment
di: Cho, Jay Hyeon, et al.
Pubblicazione: (2025)
di: Cho, Jay Hyeon, et al.
Pubblicazione: (2025)
Captioning for Text-Video Retrieval via Dual-Group Direct Preference Optimization
di: Lee, Ji Soo, et al.
Pubblicazione: (2025)
di: Lee, Ji Soo, et al.
Pubblicazione: (2025)
Coarse-grained quantum state tomography with optimal POVM construction
di: Jung, Donghun, et al.
Pubblicazione: (2024)
di: Jung, Donghun, et al.
Pubblicazione: (2024)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
Catalyst: a Novel Regularizer for Structured Pruning with Auxiliary Extension of Parameter Space
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
ASAP: Attention Sink Anchored Pruning
di: Lee, Jaehyuk, et al.
Pubblicazione: (2026)
di: Lee, Jaehyuk, et al.
Pubblicazione: (2026)
Immiscible Binary Organic Phase‐Change Composites with Segregated h‐BN Networks for Advanced Thermal Management
di: Donghun Lee, et al.
Pubblicazione: (2025)
di: Donghun Lee, et al.
Pubblicazione: (2025)
Deep Learning Analysis of Localized Interlayer Stacking Displacement and Dynamics in Bilayer Phosphorene
di: Kihyun Lee, et al.
Pubblicazione: (2025)
di: Kihyun Lee, et al.
Pubblicazione: (2025)
Broadband Ground Motion Synthesis by Diffusion Model with Minimal Condition
di: Jung, Jaeheun, et al.
Pubblicazione: (2024)
di: Jung, Jaeheun, et al.
Pubblicazione: (2024)
Versatile Biosensing Tool: CRISPR‐Cas12a System‐Integrated Electrochemical Biosensor for Severe Fever with Thrombocytopenia Syndrome Virus Detection in Clinical and Environmental Conditions
di: Daehyeon Yoo, et al.
Pubblicazione: (2025)
di: Daehyeon Yoo, et al.
Pubblicazione: (2025)
Margin Matching Preference Optimization: Enhanced Model Alignment with Granular Feedback
di: Kim, Kyuyoung, et al.
Pubblicazione: (2024)
di: Kim, Kyuyoung, et al.
Pubblicazione: (2024)
IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
di: Kim, Wonjoong, et al.
Pubblicazione: (2025)
di: Kim, Wonjoong, et al.
Pubblicazione: (2025)
A Core-Structure-Based Automated Analysis Tool for Commercial Virtualization Obfuscation Deobfuscation
di: Kim, Wanju, et al.
Pubblicazione: (2026)
di: Kim, Wanju, et al.
Pubblicazione: (2026)
E-Book Usability in Educational Technology Classes: Teachers and Teacher Candidates' Perception toward E-Book for Teaching and Learning
di: Shin, Sunghee
Pubblicazione: (2014)
di: Shin, Sunghee
Pubblicazione: (2014)
Can seasonal prediction models capture the Arctic mid‐latitude teleconnection on monthly time scales?
di: Gaeun Kim, et al.
Pubblicazione: (2024)
di: Gaeun Kim, et al.
Pubblicazione: (2024)
Gold nanoshells with varying morphologies through templated surfactant‐assisted seed‐growth method
di: Sunghee Lee, et al.
Pubblicazione: (2024)
di: Sunghee Lee, et al.
Pubblicazione: (2024)
Hierarchical Summary Statistics Encoding Across Primary Visual and Posterior Parietal Cortices
di: Young‐Beom Lee, et al.
Pubblicazione: (2026)
di: Young‐Beom Lee, et al.
Pubblicazione: (2026)
Boost Your Human Image Generation Model via Direct Preference Optimization
di: Na, Sanghyeon, et al.
Pubblicazione: (2024)
di: Na, Sanghyeon, et al.
Pubblicazione: (2024)
Retrieval-Augmented Fine-Tuning With Preference Optimization For Visual Program Generation
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
2025 update on $\varepsilon_K$ in the Standard Model with lattice QCD inputs
di: Jwa, Seungyeob, et al.
Pubblicazione: (2025)
di: Jwa, Seungyeob, et al.
Pubblicazione: (2025)
$β$-DPO: Direct Preference Optimization with Dynamic $β$
di: Wu, Junkang, et al.
Pubblicazione: (2024)
di: Wu, Junkang, et al.
Pubblicazione: (2024)
DPO-Shift: Shifting the Distribution of Direct Preference Optimization
di: Yang, Xiliang, et al.
Pubblicazione: (2025)
di: Yang, Xiliang, et al.
Pubblicazione: (2025)
Crystalline-to-Crystalline Phase Transition between Germanium Selenide Polymorphs with High Resistance Contrast
di: Kim, Joonho, et al.
Pubblicazione: (2025)
di: Kim, Joonho, et al.
Pubblicazione: (2025)
Data-Driven Dimensional Synthesis of Diverse Planar Four-bar Function Generation Mechanisms via Direct Parameterization
di: Kim, Woon Ryong, et al.
Pubblicazione: (2025)
di: Kim, Woon Ryong, et al.
Pubblicazione: (2025)
Differential Information Distribution: A Bayesian Perspective on Direct Preference Optimization
di: Won, Yunjae, et al.
Pubblicazione: (2025)
di: Won, Yunjae, et al.
Pubblicazione: (2025)
Towards Scalable Human-aligned Benchmark for Text-guided Image Editing
di: Ryu, Suho, et al.
Pubblicazione: (2025)
di: Ryu, Suho, et al.
Pubblicazione: (2025)
KTRL+F: Knowledge-Augmented In-Document Search
di: Oh, Hanseok, et al.
Pubblicazione: (2023)
di: Oh, Hanseok, et al.
Pubblicazione: (2023)
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
di: Liu, Ziyu, et al.
Pubblicazione: (2024)
di: Liu, Ziyu, et al.
Pubblicazione: (2024)
StablePrompt: Automatic Prompt Tuning using Reinforcement Learning for Large Language Models
di: Kwon, Minchan, et al.
Pubblicazione: (2024)
di: Kwon, Minchan, et al.
Pubblicazione: (2024)
Autonomous Bayesian Optimization‐Based Control System for Droplet Generation
di: Seongsu Cho, et al.
Pubblicazione: (2025)
di: Seongsu Cho, et al.
Pubblicazione: (2025)
Direct Preference Optimization-Enhanced Multi-Guided Diffusion Model for Traffic Scenario Generation
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
Enhancing Temporal Action Localization: Advanced S6 Modeling with Recurrent Mechanism
di: Lee, Sangyoun, et al.
Pubblicazione: (2024)
di: Lee, Sangyoun, et al.
Pubblicazione: (2024)
SEE-DPO: Self Entropy Enhanced Direct Preference Optimization
di: Shekhar, Shivanshu, et al.
Pubblicazione: (2024)
di: Shekhar, Shivanshu, et al.
Pubblicazione: (2024)
TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization
di: Abdullah, Abdulhady Abas, et al.
Pubblicazione: (2026)
di: Abdullah, Abdulhady Abas, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FunctionChat-Bench: Comprehensive Evaluation of Language Models' Generative Capabilities in Korean Tool-use Dialogs
di: Lee, Shinbok, et al.
Pubblicazione: (2024) -
Kanana: Compute-efficient Bilingual Language Models
di: Kanana LLM Team, et al.
Pubblicazione: (2025) -
SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety
di: Kim, Geon-Hyeong, et al.
Pubblicazione: (2025) -
Learning Design Preferences through Design Feature Extraction and Weighted Ensemble
di: Shin, Dongju, et al.
Pubblicazione: (2024) -
Position Optimization of Passive Patch Based on Mode Contribution Factor for Vibration Attenuation of Asymmetric 1D Structure
di: Dongwoo Hong, et al.
Pubblicazione: (2024)