LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ping, Bowen, Zeng, Jiali, Meng, Fandong, Wang, Shuo, Zhou, Jie, Zhang, Shanghang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TIM: Teaching Large Language Models to Translate with Comparison
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
Improving Machine Translation with Large Language Models: A Preliminary Study with Cooperative Decoding
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
Understanding and Addressing the Under-Translation Problem from the Perspective of Decoding Objective
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
Generative Multi-Modal Knowledge Retrieval with Large Language Models
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
Offline Exploration-Aware Fine-Tuning for Long-Chain Mathematical Reasoning
von: Mu, Yongyu, et al.
Veröffentlicht: (2026)
von: Mu, Yongyu, et al.
Veröffentlicht: (2026)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
von: Mu, Yongyu, et al.
Veröffentlicht: (2025)
von: Mu, Yongyu, et al.
Veröffentlicht: (2025)
Less, but Better: Efficient Multilingual Expansion for LLMs via Layer-wise Mixture-of-Experts
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
TasTe: Teaching Large Language Models to Translate through Self-Reflection
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
SlangDIT: Benchmarking LLMs in Interpretative Slang Translation
von: Liang, Yunlong, et al.
Veröffentlicht: (2025)
von: Liang, Yunlong, et al.
Veröffentlicht: (2025)
Groundedness in Retrieval-augmented Long-form Generation: An Empirical Study
von: Stolfo, Alessandro
Veröffentlicht: (2024)
von: Stolfo, Alessandro
Veröffentlicht: (2024)
DeepTrans: Deep Reasoning Translation via Reinforcement Learning
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
ExTrans: Multilingual Deep Reasoning Translation via Exemplar-Enhanced Reinforcement Learning
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
LexMatcher: Dictionary-centric Data Collection for LLM-based Machine Translation
von: Yin, Yongjing, et al.
Veröffentlicht: (2024)
von: Yin, Yongjing, et al.
Veröffentlicht: (2024)
LongLLaDA: Unlocking Long Context Capabilities in Diffusion LLMs
von: Liu, Xiaoran, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoran, et al.
Veröffentlicht: (2025)
Dancing with Critiques: Enhancing LLM Reasoning with Stepwise Natural Language Self-Critique
von: Li, Yansi, et al.
Veröffentlicht: (2025)
von: Li, Yansi, et al.
Veröffentlicht: (2025)
Enhancing Mathematical Reasoning in LLMs by Stepwise Correction
von: Wu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2024)
Accelerating Inference in Large Language Models with a Unified Layer Skipping Strategy
von: Liu, Yijin, et al.
Veröffentlicht: (2024)
von: Liu, Yijin, et al.
Veröffentlicht: (2024)
THOR-MoE: Hierarchical Task-Guided and Context-Responsive Routing for Neural Machine Translation
von: Liang, Yunlong, et al.
Veröffentlicht: (2025)
von: Liang, Yunlong, et al.
Veröffentlicht: (2025)
Figure It Out: Improve the Frontier of Reasoning with Executable Visual States
von: Chen, Meiqi, et al.
Veröffentlicht: (2025)
von: Chen, Meiqi, et al.
Veröffentlicht: (2025)
Towards Codable Watermarking for Injecting Multi-bits Information to LLMs
von: Wang, Lean, et al.
Veröffentlicht: (2023)
von: Wang, Lean, et al.
Veröffentlicht: (2023)
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
Language Generation with Strictly Proper Scoring Rules
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
DelTA: An Online Document-Level Translation Agent Based on Multi-Level Memory
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
von: Lai, Xin, et al.
Veröffentlicht: (2024)
von: Lai, Xin, et al.
Veröffentlicht: (2024)
EVA-Score: Evaluating Abstractive Long-form Summarization on Informativeness through Extraction and Validation
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
CSCD-NS: a Chinese Spelling Check Dataset for Native Speakers
von: Hu, Yong, et al.
Veröffentlicht: (2022)
von: Hu, Yong, et al.
Veröffentlicht: (2022)
HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational Modeling
von: Zhou, Chulun, et al.
Veröffentlicht: (2025)
von: Zhou, Chulun, et al.
Veröffentlicht: (2025)
EAG: Extract and Generate Multi-way Aligned Corpus for Complete Multi-lingual Neural Machine Translation
von: Xu, Yulin, et al.
Veröffentlicht: (2022)
von: Xu, Yulin, et al.
Veröffentlicht: (2022)
LFQA-E: Carefully Benchmarking Long-form QA Evaluation
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization
von: Sun, Huashan, et al.
Veröffentlicht: (2025)
von: Sun, Huashan, et al.
Veröffentlicht: (2025)
Comments as Natural Logic Pivots: Improve Code Generation via Comment Perspective
von: Chen, Yijie, et al.
Veröffentlicht: (2024)
von: Chen, Yijie, et al.
Veröffentlicht: (2024)
Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
Beyond Next Token Prediction: Patch-Level Training for Large Language Models
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
XAL: EXplainable Active Learning Makes Classifiers Better Low-resource Learners
von: Luo, Yun, et al.
Veröffentlicht: (2023)
von: Luo, Yun, et al.
Veröffentlicht: (2023)
Self-Evolving Critique Abilities in Large Language Models
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
Long-form RewardBench: Evaluating Reward Models for Long-form Generation
von: Huang, Hui, et al.
Veröffentlicht: (2026)
von: Huang, Hui, et al.
Veröffentlicht: (2026)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
von: Ke, Pei, et al.
Veröffentlicht: (2023)
von: Ke, Pei, et al.
Veröffentlicht: (2023)
LongR: Unleashing Long-Context Reasoning via Reinforcement Learning with Dense Utility Rewards
von: Ping, Bowen, et al.
Veröffentlicht: (2026)
von: Ping, Bowen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
TIM: Teaching Large Language Models to Translate with Comparison
von: Zeng, Jiali, et al.
Veröffentlicht: (2023) -
Improving Machine Translation with Large Language Models: A Preliminary Study with Cooperative Decoding
von: Zeng, Jiali, et al.
Veröffentlicht: (2023) -
Understanding and Addressing the Under-Translation Problem from the Perspective of Decoding Objective
von: Shao, Chenze, et al.
Veröffentlicht: (2024) -
Generative Multi-Modal Knowledge Retrieval with Large Language Models
von: Long, Xinwei, et al.
Veröffentlicht: (2024) -
DRT: Deep Reasoning Translation via Long Chain-of-Thought
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)