Dynamic Noise Preference Optimization: Self-Improvement of Large Language Models with Self-Synthetic Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Haoyan, Le, Khiem, Hua, Ting, Gao, Shangqian, Xu, Binfeng, Tang, Zheng, Xu, Jie, Chawla, Nitesh V., Jin, Hongxia, Srinivasan, Vijay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AgentDrug: Utilizing Large Language Models in An Agentic Workflow for Zero-Shot Molecular Editing
von: Le, Khiem, et al.
Veröffentlicht: (2024)
von: Le, Khiem, et al.
Veröffentlicht: (2024)
Transformation-Augmented GRPO for Enhancing Exploration in Reasoning of Large Language Models
von: Le, Khiem, et al.
Veröffentlicht: (2026)
von: Le, Khiem, et al.
Veröffentlicht: (2026)
FLAME: Towards Federated Fine-Tuning Large Language Models Through Adaptive SMoE
von: Le, Khiem, et al.
Veröffentlicht: (2025)
von: Le, Khiem, et al.
Veröffentlicht: (2025)
Capability Self-Assessment: Teaching LLMs to Know Their Limits
von: Yang, Haoyan, et al.
Veröffentlicht: (2026)
von: Yang, Haoyan, et al.
Veröffentlicht: (2026)
Automated Benchmark Generation from Domain Guidelines Informed by Bloom's Taxonomy
von: Chen, Si, et al.
Veröffentlicht: (2026)
von: Chen, Si, et al.
Veröffentlicht: (2026)
Can Decision Trees Teach Large Language Models? Distilling Verbalized Knowledge for Molecular Property Prediction
von: Le, Khiem, et al.
Veröffentlicht: (2026)
von: Le, Khiem, et al.
Veröffentlicht: (2026)
DISP-LLM: Dimension-Independent Structural Pruning for Large Language Models
von: Gao, Shangqian, et al.
Veröffentlicht: (2024)
von: Gao, Shangqian, et al.
Veröffentlicht: (2024)
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation
von: Chen, Qiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Qiyuan, et al.
Veröffentlicht: (2025)
Context Attribution with Multi-Armed Bandit Optimization
von: Pan, Deng, et al.
Veröffentlicht: (2025)
von: Pan, Deng, et al.
Veröffentlicht: (2025)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
MolX: Enhancing Large Language Models for Molecular Understanding With A Multi-Modal Extension
von: Le, Khiem, et al.
Veröffentlicht: (2024)
von: Le, Khiem, et al.
Veröffentlicht: (2024)
CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
Building Scaffolding Dialogue Data with LLM-Simulated Novices
von: Chen, Si, et al.
Veröffentlicht: (2025)
von: Chen, Si, et al.
Veröffentlicht: (2025)
Self-Improvement of Large Language Models: A Technical Overview and Future Outlook
von: Yang, Haoyan, et al.
Veröffentlicht: (2026)
von: Yang, Haoyan, et al.
Veröffentlicht: (2026)
Fast Explanations via Policy Gradient-Optimized Explainer
von: Pan, Deng, et al.
Veröffentlicht: (2024)
von: Pan, Deng, et al.
Veröffentlicht: (2024)
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
von: Chen, Renjie, et al.
Veröffentlicht: (2025)
von: Chen, Renjie, et al.
Veröffentlicht: (2025)
Self-Consistency Preference Optimization
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
Self-Boosting Large Language Models with Synthetic Preference Data
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
Self-Steering Optimization: Autonomous Preference Optimization for Large Language Models
von: Xiang, Hao, et al.
Veröffentlicht: (2024)
von: Xiang, Hao, et al.
Veröffentlicht: (2024)
pFedDSH: Enabling Knowledge Transfer in Personalized Federated Learning through Data-free Sub-Hypernetwork
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
A Survey of Large Language Models for Graphs
von: Ren, Xubin, et al.
Veröffentlicht: (2024)
von: Ren, Xubin, et al.
Veröffentlicht: (2024)
Dynamic Bayesian Optimization Framework for Instruction Tuning in Partial Differential Equation Discovery
von: Qu, Junqi, et al.
Veröffentlicht: (2025)
von: Qu, Junqi, et al.
Veröffentlicht: (2025)
Any Large Language Model Can Be a Reliable Judge: Debiasing with a Reasoning-based Bias Detector
von: Yang, Haoyan, et al.
Veröffentlicht: (2025)
von: Yang, Haoyan, et al.
Veröffentlicht: (2025)
Bridging the AI Adoption Gap: Designing an Interactive Pedagogical Agent for Higher Education Instructors
von: Chen, Si, et al.
Veröffentlicht: (2025)
von: Chen, Si, et al.
Veröffentlicht: (2025)
The Variational Approach in Filtering and Correlated Noise
von: Srinivasan, Sharan, et al.
Veröffentlicht: (2026)
von: Srinivasan, Sharan, et al.
Veröffentlicht: (2026)
Self-Improvement for Audio Large Language Model using Unlabeled Speech
von: Wang, Shaowen, et al.
Veröffentlicht: (2025)
von: Wang, Shaowen, et al.
Veröffentlicht: (2025)
ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruning
von: Gao, Shangqian, et al.
Veröffentlicht: (2025)
von: Gao, Shangqian, et al.
Veröffentlicht: (2025)
Self-supervised Preference Optimization: Enhance Your Language Model with Preference Degree Awareness
von: Li, Jian, et al.
Veröffentlicht: (2024)
von: Li, Jian, et al.
Veröffentlicht: (2024)
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
Adaptive Testing for LLM Evaluation: A Psychometric Alternative to Static Benchmarks
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
Dynamic Controlled Variables Based Dynamic Self-Optimizing Control
von: Zhou, Chenchen, et al.
Veröffentlicht: (2026)
von: Zhou, Chenchen, et al.
Veröffentlicht: (2026)
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
von: Yan, Jun, et al.
Veröffentlicht: (2023)
von: Yan, Jun, et al.
Veröffentlicht: (2023)
LatentPrompt: Optimizing Promts in Latent Space
von: Bystroński, Mateusz, et al.
Veröffentlicht: (2025)
von: Bystroński, Mateusz, et al.
Veröffentlicht: (2025)
Instruction-following Evaluation through Verbalizer Manipulation
von: Li, Shiyang, et al.
Veröffentlicht: (2023)
von: Li, Shiyang, et al.
Veröffentlicht: (2023)
Class-Aware Contrastive Optimization for Imbalanced Text Classification
von: Khvatskii, Grigorii, et al.
Veröffentlicht: (2024)
von: Khvatskii, Grigorii, et al.
Veröffentlicht: (2024)
TeachingCoach: A Fine-Tuned Scaffolding Chatbot for Instructional Guidance to Instructors
von: Molnar, Isabel, et al.
Veröffentlicht: (2026)
von: Molnar, Isabel, et al.
Veröffentlicht: (2026)
Graph Synthetic Out-of-Distribution Exposure with Large Language Models
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
Geometry of Knowledge Allows Extending Diversity Boundaries of Large Language Models
von: Bystroński, Mateusz, et al.
Veröffentlicht: (2025)
von: Bystroński, Mateusz, et al.
Veröffentlicht: (2025)
Conformalized Selective Regression
von: Sokol, Anna, et al.
Veröffentlicht: (2024)
von: Sokol, Anna, et al.
Veröffentlicht: (2024)
The Hidden Influence of Latent Feature Magnitude When Learning with Imbalanced Data
von: Dablain, Damien A., et al.
Veröffentlicht: (2024)
von: Dablain, Damien A., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AgentDrug: Utilizing Large Language Models in An Agentic Workflow for Zero-Shot Molecular Editing
von: Le, Khiem, et al.
Veröffentlicht: (2024) -
Transformation-Augmented GRPO for Enhancing Exploration in Reasoning of Large Language Models
von: Le, Khiem, et al.
Veröffentlicht: (2026) -
FLAME: Towards Federated Fine-Tuning Large Language Models Through Adaptive SMoE
von: Le, Khiem, et al.
Veröffentlicht: (2025) -
Capability Self-Assessment: Teaching LLMs to Know Their Limits
von: Yang, Haoyan, et al.
Veröffentlicht: (2026) -
Automated Benchmark Generation from Domain Guidelines Informed by Bloom's Taxonomy
von: Chen, Si, et al.
Veröffentlicht: (2026)