Uncertainty Aware Learning for Language Model Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yikun, Zheng, Rui, Ding, Liang, Zhang, Qi, Lin, Dahua, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning Large Language Models from Self-Reference AI Feedback with one General Principle
von: Bao, Rong, et al.
Veröffentlicht: (2024)
von: Bao, Rong, et al.
Veröffentlicht: (2024)
The Bitter Lesson of Diffusion Language Models for Agentic Workflows: A Comprehensive Reality Check
von: Lu, Qingyu, et al.
Veröffentlicht: (2026)
von: Lu, Qingyu, et al.
Veröffentlicht: (2026)
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
von: Rao, Jun, et al.
Veröffentlicht: (2024)
von: Rao, Jun, et al.
Veröffentlicht: (2024)
UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models
von: Xue, Boyang, et al.
Veröffentlicht: (2024)
von: Xue, Boyang, et al.
Veröffentlicht: (2024)
OOP: Object-Oriented Programming Evaluation Benchmark for Large Language Models
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Building Accurate Translation-Tailored LLMs with Language Aware Instruction Tuning
von: Zan, Changtong, et al.
Veröffentlicht: (2024)
von: Zan, Changtong, et al.
Veröffentlicht: (2024)
VisuoThink: Empowering LVLM Reasoning with Multimodal Tree Search
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
Error Analysis Prompting Enables Human-Like Translation Evaluation in Large Language Models
von: Lu, Qingyu, et al.
Veröffentlicht: (2023)
von: Lu, Qingyu, et al.
Veröffentlicht: (2023)
Entropy-Guided Watermarking for LLMs: A Test-Time Framework for Robust and Traceable Text Generation
von: Cai, Shizhan, et al.
Veröffentlicht: (2025)
von: Cai, Shizhan, et al.
Veröffentlicht: (2025)
ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
Rescue: Ranking LLM Responses with Partial Ordering to Improve Response Generation
von: Wang, Yikun, et al.
Veröffentlicht: (2023)
von: Wang, Yikun, et al.
Veröffentlicht: (2023)
Revisiting Knowledge Distillation for Autoregressive Language Models
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
Learning from Imperfect Data: Towards Efficient Knowledge Distillation of Autoregressive Language Models for Text-to-SQL
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
Intention Analysis Makes LLMs A Good Jailbreak Defender
von: Zhang, Yuqi, et al.
Veröffentlicht: (2024)
von: Zhang, Yuqi, et al.
Veröffentlicht: (2024)
Unveiling the Misuse Potential of Base Large Language Models via In-Context Learning
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Improving Large Language Models with Concept-Aware Fine-Tuning
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
NoVo: Norm Voting off Hallucinations with Attention Heads in Large Language Models
von: Ho, Zheng Yi, et al.
Veröffentlicht: (2024)
von: Ho, Zheng Yi, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Large Language Models for Explainable Disease Diagnosis
von: Zhou, Shuang, et al.
Veröffentlicht: (2025)
von: Zhou, Shuang, et al.
Veröffentlicht: (2025)
E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and Generation
von: Zhong, Qihuang, et al.
Veröffentlicht: (2022)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2022)
Improving Complex Reasoning over Knowledge Graph with Logic-Aware Curriculum Tuning
von: Xia, Tianle, et al.
Veröffentlicht: (2024)
von: Xia, Tianle, et al.
Veröffentlicht: (2024)
Adding Alignment Control to Language Models
von: Zhu, Wenhong, et al.
Veröffentlicht: (2025)
von: Zhu, Wenhong, et al.
Veröffentlicht: (2025)
PANDA: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation
von: Zhong, Qihuang, et al.
Veröffentlicht: (2022)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2022)
Reason-KE++: Aligning the Process, Not Just the Outcome, for Faithful LLM Knowledge Editing
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
Robust Knowledge Editing via Explicit Reasoning Chains for Distractor-Resilient Multi-Hop QA
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
MQM-APE: Toward High-Quality Error Annotation Predictors with Automatic Post-Editing in LLM Translation Evaluators
von: Lu, Qingyu, et al.
Veröffentlicht: (2024)
von: Lu, Qingyu, et al.
Veröffentlicht: (2024)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
Runaway is Ashamed, But Helpful: On the Early-Exit Behavior of Large Language Model-based Agents in Embodied Environments
von: Lu, Qingyu, et al.
Veröffentlicht: (2025)
von: Lu, Qingyu, et al.
Veröffentlicht: (2025)
LLM-DA: Data Augmentation via Large Language Models for Few-Shot Named Entity Recognition
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
Linear Alignment: A Closed-form Solution for Aligning Human Preferences without Tuning and Feedback
von: Gao, Songyang, et al.
Veröffentlicht: (2024)
von: Gao, Songyang, et al.
Veröffentlicht: (2024)
Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
von: Gu, Yuzhe, et al.
Veröffentlicht: (2025)
von: Gu, Yuzhe, et al.
Veröffentlicht: (2025)
Scaling Behavior for Large Language Models regarding Numeral Systems: An Example using Pythia
von: Zhou, Zhejian, et al.
Veröffentlicht: (2024)
von: Zhou, Zhejian, et al.
Veröffentlicht: (2024)
Revisiting Demonstration Selection Strategies in In-Context Learning
von: Peng, Keqin, et al.
Veröffentlicht: (2024)
von: Peng, Keqin, et al.
Veröffentlicht: (2024)
Diversifying the Mixture-of-Experts Representation for Language Models with Orthogonal Optimizer
von: Liu, Boan, et al.
Veröffentlicht: (2023)
von: Liu, Boan, et al.
Veröffentlicht: (2023)
Enhancing Input-Label Mapping in In-Context Learning with Contrastive Decoding
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
FERA: Uncertainty-Aware Federated Reasoning for Large Language Models
von: Wang, Ruhan, et al.
Veröffentlicht: (2026)
von: Wang, Ruhan, et al.
Veröffentlicht: (2026)
Towards Harmonized Uncertainty Estimation for Large Language Models
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Better, Faster: Harnessing Self-Improvement in Large Reasoning Models
von: Zhong, Qihuang, et al.
Veröffentlicht: (2026)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2026)
Edit Once, Update Everywhere: A Simple Framework for Cross-Lingual Knowledge Synchronization in LLMs
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Aligning Large Language Models from Self-Reference AI Feedback with one General Principle
von: Bao, Rong, et al.
Veröffentlicht: (2024) -
The Bitter Lesson of Diffusion Language Models for Agentic Workflows: A Comprehensive Reality Check
von: Lu, Qingyu, et al.
Veröffentlicht: (2026) -
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024) -
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
von: Rao, Jun, et al.
Veröffentlicht: (2024) -
UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models
von: Xue, Boyang, et al.
Veröffentlicht: (2024)