I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
Fuente:
arXiv
Salvato in:
| Autori principali: | Liang, Yiming, Zhang, Ge, Qu, Xingwei, Zheng, Tianyu, Guo, Jiawei, Du, Xinrun, Yang, Zhenzhu, Liu, Jiaheng, Lin, Chenghua, Ma, Lei, Huang, Wenhao, Zhang, Jiajun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
di: Zheng, Tianyu, et al.
Pubblicazione: (2024)
di: Zheng, Tianyu, et al.
Pubblicazione: (2024)
Aligning Instruction Tuning with Pre-training
di: Liang, Yiming, et al.
Pubblicazione: (2025)
di: Liang, Yiming, et al.
Pubblicazione: (2025)
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
di: Ma, Kaijing, et al.
Pubblicazione: (2024)
di: Ma, Kaijing, et al.
Pubblicazione: (2024)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
di: Du, Xinrun, et al.
Pubblicazione: (2024)
di: Du, Xinrun, et al.
Pubblicazione: (2024)
Enhancing Knowledge Distillation of Large Language Models through Efficient Multi-Modal Distribution Alignment
di: Peng, Tianyu, et al.
Pubblicazione: (2024)
di: Peng, Tianyu, et al.
Pubblicazione: (2024)
TEGEE: Task dEfinition Guided Expert Ensembling for Generalizable and Few-shot Learning
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
Human-Instruction-Free LLM Self-Alignment with Limited Samples
di: Guo, Hongyi, et al.
Pubblicazione: (2024)
di: Guo, Hongyi, et al.
Pubblicazione: (2024)
Online Iterative Self-Alignment for Radiology Report Generation
di: Xiao, Ting, et al.
Pubblicazione: (2025)
di: Xiao, Ting, et al.
Pubblicazione: (2025)
Self-Sovereign Agent
di: Qu, Wenjie, et al.
Pubblicazione: (2026)
di: Qu, Wenjie, et al.
Pubblicazione: (2026)
RLoop: An Self-Improving Framework for Reinforcement Learning with Iterative Policy Initialization
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
Can MLLMs Understand the Deep Implication Behind Chinese Images?
di: Zhang, Chenhao, et al.
Pubblicazione: (2024)
di: Zhang, Chenhao, et al.
Pubblicazione: (2024)
Stackelberg Self-Annotation: A Robust Approach to Data-Efficient LLM Alignment
di: Chu, Xu, et al.
Pubblicazione: (2025)
di: Chu, Xu, et al.
Pubblicazione: (2025)
ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation
di: Ren, Weiming, et al.
Pubblicazione: (2024)
di: Ren, Weiming, et al.
Pubblicazione: (2024)
GIEBench: Towards Holistic Evaluation of Group Identity-based Empathy for Large Language Models
di: Wang, Leyan, et al.
Pubblicazione: (2024)
di: Wang, Leyan, et al.
Pubblicazione: (2024)
Encyclo-K: Evaluating LLMs with Dynamically Composed Knowledge Statements
di: Liang, Yiming, et al.
Pubblicazione: (2025)
di: Liang, Yiming, et al.
Pubblicazione: (2025)
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
di: Zhang, Ge, et al.
Pubblicazione: (2024)
di: Zhang, Ge, et al.
Pubblicazione: (2024)
COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values
di: P Team, et al.
Pubblicazione: (2025)
di: P Team, et al.
Pubblicazione: (2025)
Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting Code
di: Lin, Haobo, et al.
Pubblicazione: (2026)
di: Lin, Haobo, et al.
Pubblicazione: (2026)
OmniBench: Towards The Future of Universal Omni-Language Models
di: Li, Yizhi, et al.
Pubblicazione: (2024)
di: Li, Yizhi, et al.
Pubblicazione: (2024)
Steel-LLM:From Scratch to Open Source -- A Personal Journey in Building a Chinese-Centric LLM
di: Gu, Qingshui, et al.
Pubblicazione: (2025)
di: Gu, Qingshui, et al.
Pubblicazione: (2025)
Self-Supervised Visual Preference Alignment
di: Zhu, Ke, et al.
Pubblicazione: (2024)
di: Zhu, Ke, et al.
Pubblicazione: (2024)
First Return, Entropy-Eliciting Explore
di: Zheng, Tianyu, et al.
Pubblicazione: (2025)
di: Zheng, Tianyu, et al.
Pubblicazione: (2025)
Anchored Alignment for Self-Explanations Enhancement
di: Villa-Arenas, Luis Felipe, et al.
Pubblicazione: (2024)
di: Villa-Arenas, Luis Felipe, et al.
Pubblicazione: (2024)
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
di: Zhang, Bingqing, et al.
Pubblicazione: (2024)
di: Zhang, Bingqing, et al.
Pubblicazione: (2024)
Haemonchus contortus-SHEEP RELATIONSHIP: A REVIEW
di: Francisco J. Angulo-Cubillán
Pubblicazione: (2007)
di: Francisco J. Angulo-Cubillán
Pubblicazione: (2007)
The Fine Line: Navigating Large Language Model Pretraining with Down-streaming Capability Analysis
di: Yang, Chen, et al.
Pubblicazione: (2024)
di: Yang, Chen, et al.
Pubblicazione: (2024)
RA-ISF: Learning to Answer and Understand from Retrieval Augmentation via Iterative Self-Feedback
di: Liu, Yanming, et al.
Pubblicazione: (2024)
di: Liu, Yanming, et al.
Pubblicazione: (2024)
Just a Scratch: Enhancing LLM Capabilities for Self-harm Detection through Intent Differentiation and Emoji Interpretation
di: Ghosh, Soumitra, et al.
Pubblicazione: (2025)
di: Ghosh, Soumitra, et al.
Pubblicazione: (2025)
MORE-3S:Multimodal-based Offline Reinforcement Learning with Shared Semantic Spaces
di: Zheng, Tianyu, et al.
Pubblicazione: (2024)
di: Zheng, Tianyu, et al.
Pubblicazione: (2024)
MMRA: A Benchmark for Evaluating Multi-Granularity and Multi-Image Relational Association Capabilities in Large Visual Language Models
di: Wu, Siwei, et al.
Pubblicazione: (2024)
di: Wu, Siwei, et al.
Pubblicazione: (2024)
Vibe AIGC: A New Paradigm for Content Generation via Agentic Orchestration
di: Liu, Jiaheng, et al.
Pubblicazione: (2026)
di: Liu, Jiaheng, et al.
Pubblicazione: (2026)
CMDAG: A Chinese Metaphor Dataset with Annotated Grounds as CoT for Boosting Metaphor Generation
di: Shao, Yujie, et al.
Pubblicazione: (2024)
di: Shao, Yujie, et al.
Pubblicazione: (2024)
Observing Micromotives and Macrobehavior of Large Language Models
di: Cheng, Yuyang, et al.
Pubblicazione: (2024)
di: Cheng, Yuyang, et al.
Pubblicazione: (2024)
SelfCodeAlign: Self-Alignment for Code Generation
di: Wei, Yuxiang, et al.
Pubblicazione: (2024)
di: Wei, Yuxiang, et al.
Pubblicazione: (2024)
KARPA: A Training-free Method of Adapting Knowledge Graph as References for Large Language Model's Reasoning Path Aggregation
di: Fang, Siyuan, et al.
Pubblicazione: (2024)
di: Fang, Siyuan, et al.
Pubblicazione: (2024)
SHAPE : Self-Improved Visual Preference Alignment by Iteratively Generating Holistic Winner
di: Chen, Kejia, et al.
Pubblicazione: (2025)
di: Chen, Kejia, et al.
Pubblicazione: (2025)
LongEval: A Comprehensive Analysis of Long-Text Generation Through a Plan-based Paradigm
di: Wu, Siwei, et al.
Pubblicazione: (2025)
di: Wu, Siwei, et al.
Pubblicazione: (2025)
MuPT: A Generative Symbolic Music Pretrained Transformer
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
di: Qu, Xingwei, et al.
Pubblicazione: (2024)
SALMON: Self-Alignment with Instructable Reward Models
di: Sun, Zhiqing, et al.
Pubblicazione: (2023)
di: Sun, Zhiqing, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
di: Zheng, Tianyu, et al.
Pubblicazione: (2024) -
Aligning Instruction Tuning with Pre-training
di: Liang, Yiming, et al.
Pubblicazione: (2025) -
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
di: Qu, Xingwei, et al.
Pubblicazione: (2024) -
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
di: Ma, Kaijing, et al.
Pubblicazione: (2024) -
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
di: Du, Xinrun, et al.
Pubblicazione: (2024)