DIVE: Diversified Iterative Self-Improvement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qin, Yiwei, Liu, Yixiu, Liu, Pengfei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Progress or Regress? Self-Improvement Reversal in Post-training
von: Wu, Ting, et al.
Veröffentlicht: (2024)
von: Wu, Ting, et al.
Veröffentlicht: (2024)
O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
von: Huang, Zhen, et al.
Veröffentlicht: (2024)
von: Huang, Zhen, et al.
Veröffentlicht: (2024)
DIVE: Embedding Compression via Self-Limiting Gradient Updates
von: Zhao, Dongfang
Veröffentlicht: (2026)
von: Zhao, Dongfang
Veröffentlicht: (2026)
O1 Replication Journey: A Strategic Progress Report -- Part 1
von: Qin, Yiwei, et al.
Veröffentlicht: (2024)
von: Qin, Yiwei, et al.
Veröffentlicht: (2024)
Diversify and Conquer: Diversity-Centric Data Selection with Iterative Refinement
von: Yu, Simon, et al.
Veröffentlicht: (2024)
von: Yu, Simon, et al.
Veröffentlicht: (2024)
SAFETY-J: Evaluating Safety with Critique
von: Liu, Yixiu, et al.
Veröffentlicht: (2024)
von: Liu, Yixiu, et al.
Veröffentlicht: (2024)
Iterative Self-Improvement of Vision Language Models for Image Scoring and Self-Explanation
von: Tanji, Naoto, et al.
Veröffentlicht: (2025)
von: Tanji, Naoto, et al.
Veröffentlicht: (2025)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
von: Feng, Yuchen, et al.
Veröffentlicht: (2025)
von: Feng, Yuchen, et al.
Veröffentlicht: (2025)
daVinci-LLM:Towards the Science of Pretraining
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
Are LLMs Stable Formal Logic Translators in Logical Reasoning Across Linguistically Diversified Texts?
von: Li, Qingchuan, et al.
Veröffentlicht: (2025)
von: Li, Qingchuan, et al.
Veröffentlicht: (2025)
Generative AI Act II: Test Time Scaling Drives Cognition Engineering
von: Xia, Shijie, et al.
Veröffentlicht: (2025)
von: Xia, Shijie, et al.
Veröffentlicht: (2025)
Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering
von: Chu, Zheng, et al.
Veröffentlicht: (2025)
von: Chu, Zheng, et al.
Veröffentlicht: (2025)
KCS: Diversify Multi-hop Question Generation with Knowledge Composition Sampling
von: Wang, Yangfan, et al.
Veröffentlicht: (2025)
von: Wang, Yangfan, et al.
Veröffentlicht: (2025)
Benchmarking Large Language Models on Controllable Generation under Diversified Instructions
von: Chen, Yihan, et al.
Veröffentlicht: (2024)
von: Chen, Yihan, et al.
Veröffentlicht: (2024)
ISSR: Iterative Selection with Self-Review for Vocabulary Test Distractor Generation
von: Liu, Yu-Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Yu-Cheng, et al.
Veröffentlicht: (2025)
Better, Faster: Harnessing Self-Improvement in Large Reasoning Models
von: Zhong, Qihuang, et al.
Veröffentlicht: (2026)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2026)
SGIC: A Self-Guided Iterative Calibration Framework for RAG
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
Iterative Improvement of an Additively Regularized Topic Model
von: Gorbulev, Alex, et al.
Veröffentlicht: (2024)
von: Gorbulev, Alex, et al.
Veröffentlicht: (2024)
RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation
von: Brehme, Lorenz, et al.
Veröffentlicht: (2026)
von: Brehme, Lorenz, et al.
Veröffentlicht: (2026)
Diversifying the Mixture-of-Experts Representation for Language Models with Orthogonal Optimizer
von: Liu, Boan, et al.
Veröffentlicht: (2023)
von: Liu, Boan, et al.
Veröffentlicht: (2023)
Diversifying the Expert Knowledge for Task-Agnostic Pruning in Sparse Mixture-of-Experts
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
Test-time Recursive Thinking: Self-Improvement without External Feedback
von: Zhuang, Yufan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yufan, et al.
Veröffentlicht: (2026)
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement
von: Wang, Xiaobo, et al.
Veröffentlicht: (2026)
von: Wang, Xiaobo, et al.
Veröffentlicht: (2026)
Presenting a Paper is an Art: Self-Improvement Aesthetic Agents for Academic Presentations
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
Data Darwinism Part I: Unlocking the Value of Scientific Data for Pre-training
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
von: Liang, Yiming, et al.
Veröffentlicht: (2024)
von: Liang, Yiming, et al.
Veröffentlicht: (2024)
PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2024)
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2024)
History-Guided Iterative Visual Reasoning with Self-Correction
von: Yang, Xinglong, et al.
Veröffentlicht: (2026)
von: Yang, Xinglong, et al.
Veröffentlicht: (2026)
Strength Lies in Differences! Improving Strategy Planning for Non-collaborative Dialogues via Diversified User Simulation
von: Zhang, Tong, et al.
Veröffentlicht: (2024)
von: Zhang, Tong, et al.
Veröffentlicht: (2024)
Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
Not All Metrics Are Guilty: Improving NLG Evaluation by Diversifying References
von: Tang, Tianyi, et al.
Veröffentlicht: (2023)
von: Tang, Tianyi, et al.
Veröffentlicht: (2023)
Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
ReMiT: RL-Guided Mid-Training for Iterative LLM Evolution
von: Huang, Junjie, et al.
Veröffentlicht: (2026)
von: Huang, Junjie, et al.
Veröffentlicht: (2026)
Contextual Experience Replay for Self-Improvement of Language Agents
von: Liu, Yitao, et al.
Veröffentlicht: (2025)
von: Liu, Yitao, et al.
Veröffentlicht: (2025)
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
von: Moradi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
von: Moradi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
A Survey on LLM Inference-Time Self-Improvement
von: Dong, Xiangjue, et al.
Veröffentlicht: (2024)
von: Dong, Xiangjue, et al.
Veröffentlicht: (2024)
Enabling Language Models to Implicitly Learn Self-Improvement
von: Wang, Ziqi, et al.
Veröffentlicht: (2023)
von: Wang, Ziqi, et al.
Veröffentlicht: (2023)
InFoBench: Evaluating Instruction Following Ability in Large Language Models
von: Qin, Yiwei, et al.
Veröffentlicht: (2024)
von: Qin, Yiwei, et al.
Veröffentlicht: (2024)
TypedThinker: Diversify Large Language Model Reasoning with Typed Thinking
von: Wang, Danqing, et al.
Veröffentlicht: (2024)
von: Wang, Danqing, et al.
Veröffentlicht: (2024)
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Progress or Regress? Self-Improvement Reversal in Post-training
von: Wu, Ting, et al.
Veröffentlicht: (2024) -
O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
von: Huang, Zhen, et al.
Veröffentlicht: (2024) -
DIVE: Embedding Compression via Self-Limiting Gradient Updates
von: Zhao, Dongfang
Veröffentlicht: (2026) -
O1 Replication Journey: A Strategic Progress Report -- Part 1
von: Qin, Yiwei, et al.
Veröffentlicht: (2024) -
Diversify and Conquer: Diversity-Centric Data Selection with Iterative Refinement
von: Yu, Simon, et al.
Veröffentlicht: (2024)