Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gu, Yuzhe, Zhang, Wenwei, Lyu, Chengqi, Lin, Dahua, Chen, Kai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
von: Hua, Zhouqi, et al.
Veröffentlicht: (2025)
von: Hua, Zhouqi, et al.
Veröffentlicht: (2025)
ANAH: Analytical Annotation of Hallucinations in Large Language Models
von: Ji, Ziwei, et al.
Veröffentlicht: (2024)
von: Ji, Ziwei, et al.
Veröffentlicht: (2024)
ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
von: Gu, Yuzhe, et al.
Veröffentlicht: (2024)
von: Gu, Yuzhe, et al.
Veröffentlicht: (2024)
CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
von: Liu, Shudong, et al.
Veröffentlicht: (2025)
von: Liu, Shudong, et al.
Veröffentlicht: (2025)
Training Language Models to Critique With Multi-agent Feedback
von: Lan, Tian, et al.
Veröffentlicht: (2024)
von: Lan, Tian, et al.
Veröffentlicht: (2024)
Fake Alignment: Are LLMs Really Aligned Well?
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
AlchemistCoder: Harmonizing and Eliciting Code Capability by Hindsight Tuning on Multi-source Data
von: Song, Zifan, et al.
Veröffentlicht: (2024)
von: Song, Zifan, et al.
Veröffentlicht: (2024)
Fine-grained and Explainable Factuality Evaluation for Multimodal Summarization
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
Improving Model Factuality with Fine-grained Critique-based Evaluator
von: Xie, Yiqing, et al.
Veröffentlicht: (2024)
von: Xie, Yiqing, et al.
Veröffentlicht: (2024)
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning
von: Lyu, Chengqi, et al.
Veröffentlicht: (2025)
von: Lyu, Chengqi, et al.
Veröffentlicht: (2025)
MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
CIBench: Evaluating Your LLMs with a Code Interpreter Plugin
von: Zhang, Chuyu, et al.
Veröffentlicht: (2024)
von: Zhang, Chuyu, et al.
Veröffentlicht: (2024)
DiVA: Fine-grained Factuality Verification with Agentic-Discriminative Verifier
von: Huang, Hui, et al.
Veröffentlicht: (2026)
von: Huang, Hui, et al.
Veröffentlicht: (2026)
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
von: Gao, Songyang, et al.
Veröffentlicht: (2025)
von: Gao, Songyang, et al.
Veröffentlicht: (2025)
Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
Recovering Diversity Without Losing Alignment: A DPO Recipe for Post-Trained LLMs
von: Samuel, Vinay, et al.
Veröffentlicht: (2026)
von: Samuel, Vinay, et al.
Veröffentlicht: (2026)
ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs
von: Zhuo, Jingming, et al.
Veröffentlicht: (2024)
von: Zhuo, Jingming, et al.
Veröffentlicht: (2024)
Ada-LEval: Evaluating long-context LLMs with length-adaptable benchmarks
von: Wang, Chonghua, et al.
Veröffentlicht: (2024)
von: Wang, Chonghua, et al.
Veröffentlicht: (2024)
Capability Salience Vector: Fine-grained Alignment of Loss and Capabilities for Downstream Task Scaling Law
von: Ge, Qiming, et al.
Veröffentlicht: (2025)
von: Ge, Qiming, et al.
Veröffentlicht: (2025)
CriticEval: Evaluating Large Language Model as Critic
von: Lan, Tian, et al.
Veröffentlicht: (2024)
von: Lan, Tian, et al.
Veröffentlicht: (2024)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling
von: Li, Peiji, et al.
Veröffentlicht: (2025)
von: Li, Peiji, et al.
Veröffentlicht: (2025)
FLAME: Factuality-Aware Alignment for Large Language Models
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
UAQFact: Evaluating Factual Knowledge Utilization of LLMs on Unanswerable Questions
von: Tan, Chuanyuan, et al.
Veröffentlicht: (2025)
von: Tan, Chuanyuan, et al.
Veröffentlicht: (2025)
Exploring the Generalizability of Factual Hallucination Mitigation via Enhancing Precise Knowledge Utilization
von: Zhang, Siyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Siyuan, et al.
Veröffentlicht: (2025)
LLM Factoscope: Uncovering LLMs' Factual Discernment through Inner States Analysis
von: He, Jinwen, et al.
Veröffentlicht: (2023)
von: He, Jinwen, et al.
Veröffentlicht: (2023)
FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness
von: Dong, Xiaoning, et al.
Veröffentlicht: (2026)
von: Dong, Xiaoning, et al.
Veröffentlicht: (2026)
ART: Attention Replacement Technique to Improve Factuality in LLMs
von: Luo, Ziqin, et al.
Veröffentlicht: (2026)
von: Luo, Ziqin, et al.
Veröffentlicht: (2026)
InternLM2.5-StepProver: Advancing Automated Theorem Proving via Critic-Guided Search
von: Wu, Zijian, et al.
Veröffentlicht: (2024)
von: Wu, Zijian, et al.
Veröffentlicht: (2024)
From Parameters to Prompts: Understanding and Mitigating the Factuality Gap between Fine-Tuned LLMs
von: Gong, Xuan, et al.
Veröffentlicht: (2025)
von: Gong, Xuan, et al.
Veröffentlicht: (2025)
Cat-DPO: Category-Adaptive Safety Alignment
von: Yang, Tiankai, et al.
Veröffentlicht: (2026)
von: Yang, Tiankai, et al.
Veröffentlicht: (2026)
GIFT: Games as Informal Training for Generalizable LLMs
von: Lyu, Nuoyan, et al.
Veröffentlicht: (2026)
von: Lyu, Nuoyan, et al.
Veröffentlicht: (2026)
Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
von: Xu, Shusheng, et al.
Veröffentlicht: (2024)
von: Xu, Shusheng, et al.
Veröffentlicht: (2024)
Code Needs Comments: Enhancing Code LLMs with Comment Augmentation
von: Song, Demin, et al.
Veröffentlicht: (2024)
von: Song, Demin, et al.
Veröffentlicht: (2024)
LongCite: Enabling LLMs to Generate Fine-grained Citations in Long-context QA
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
MR-Align: Meta-Reasoning Informed Factuality Alignment for Large Reasoning Models
von: Wang, Xinming, et al.
Veröffentlicht: (2025)
von: Wang, Xinming, et al.
Veröffentlicht: (2025)
Beyond Imitation: Leveraging Fine-grained Quality Signals for Alignment
von: Guo, Geyang, et al.
Veröffentlicht: (2023)
von: Guo, Geyang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
von: Hua, Zhouqi, et al.
Veröffentlicht: (2025) -
ANAH: Analytical Annotation of Hallucinations in Large Language Models
von: Ji, Ziwei, et al.
Veröffentlicht: (2024) -
ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
von: Gu, Yuzhe, et al.
Veröffentlicht: (2024) -
CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
von: Liu, Shudong, et al.
Veröffentlicht: (2025) -
Training Language Models to Critique With Multi-agent Feedback
von: Lan, Tian, et al.
Veröffentlicht: (2024)