Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zheng, Tianyu, Guo, Shuyue, Qu, Xingwei, Guo, Jiawei, Du, Xinrun, Jia, Qi, Lin, Chenghua, Huang, Wenhao, Fu, Jie, Zhang, Ge |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
par: Liang, Yiming, et autres
Publié: (2024)
par: Liang, Yiming, et autres
Publié: (2024)
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
par: Zhang, Ge, et autres
Publié: (2024)
par: Zhang, Ge, et autres
Publié: (2024)
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
par: Qu, Xingwei, et autres
Publié: (2024)
par: Qu, Xingwei, et autres
Publié: (2024)
Aligning Instruction Tuning with Pre-training
par: Liang, Yiming, et autres
Publié: (2025)
par: Liang, Yiming, et autres
Publié: (2025)
CMDAG: A Chinese Metaphor Dataset with Annotated Grounds as CoT for Boosting Metaphor Generation
par: Shao, Yujie, et autres
Publié: (2024)
par: Shao, Yujie, et autres
Publié: (2024)
COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values
par: P Team, et autres
Publié: (2025)
par: P Team, et autres
Publié: (2025)
COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning
par: Bai, Yuelin, et autres
Publié: (2024)
par: Bai, Yuelin, et autres
Publié: (2024)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
par: Du, Xinrun, et autres
Publié: (2024)
par: Du, Xinrun, et autres
Publié: (2024)
Better Alignment with Instruction Back-and-Forth Translation
par: Nguyen, Thao, et autres
Publié: (2024)
par: Nguyen, Thao, et autres
Publié: (2024)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
par: Guo, Jiawei, et autres
Publié: (2024)
par: Guo, Jiawei, et autres
Publié: (2024)
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
par: Li, Yizhi, et autres
Publié: (2025)
par: Li, Yizhi, et autres
Publié: (2025)
KunPeng: A Global Ocean Environmental Model
par: Zhao, Yi, et autres
Publié: (2025)
par: Zhao, Yi, et autres
Publié: (2025)
Can MLLMs Understand the Deep Implication Behind Chinese Images?
par: Zhang, Chenhao, et autres
Publié: (2024)
par: Zhang, Chenhao, et autres
Publié: (2024)
CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models
par: LI, Yizhi, et autres
Publié: (2024)
par: LI, Yizhi, et autres
Publié: (2024)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
par: Wang, Shun, et autres
Publié: (2024)
par: Wang, Shun, et autres
Publié: (2024)
First Return, Entropy-Eliciting Explore
par: Zheng, Tianyu, et autres
Publié: (2025)
par: Zheng, Tianyu, et autres
Publié: (2025)
MuPT: A Generative Symbolic Music Pretrained Transformer
par: Qu, Xingwei, et autres
Publié: (2024)
par: Qu, Xingwei, et autres
Publié: (2024)
TEGEE: Task dEfinition Guided Expert Ensembling for Generalizable and Few-shot Learning
par: Qu, Xingwei, et autres
Publié: (2024)
par: Qu, Xingwei, et autres
Publié: (2024)
MORE-3S:Multimodal-based Offline Reinforcement Learning with Shared Semantic Spaces
par: Zheng, Tianyu, et autres
Publié: (2024)
par: Zheng, Tianyu, et autres
Publié: (2024)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
par: Zhang, Haoran, et autres
Publié: (2024)
par: Zhang, Haoran, et autres
Publié: (2024)
QiMeng-NeuComBack: Self-Evolving Translation from IR to Assembly Code
par: Fang, Hainan, et autres
Publié: (2025)
par: Fang, Hainan, et autres
Publié: (2025)
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
par: Ma, Kaijing, et autres
Publié: (2024)
par: Ma, Kaijing, et autres
Publié: (2024)
A Cross‐Cultural Comparison Between Chinese and Russian Social Media Self‐Praise
par: Yaping Guo, et autres
Publié: (2025)
par: Yaping Guo, et autres
Publié: (2025)
Kün: Edebiyat ve Kültür Araştırmaları Dergisi
Publié: (2025)
Publié: (2025)
Handbook of probiotics and prebiotics / Yuan Kun Lee
par: Lee, Yuan Kun
Publié: (2009)
par: Lee, Yuan Kun
Publié: (2009)
Observing Micromotives and Macrobehavior of Large Language Models
par: Cheng, Yuyang, et autres
Publié: (2024)
par: Cheng, Yuyang, et autres
Publié: (2024)
KARPA: A Training-free Method of Adapting Knowledge Graph as References for Large Language Model's Reasoning Path Aggregation
par: Fang, Siyuan, et autres
Publié: (2024)
par: Fang, Siyuan, et autres
Publié: (2024)
Anchor then Polish for Low-light Enhancement
par: Du, Tianle, et autres
Publié: (2026)
par: Du, Tianle, et autres
Publié: (2026)
StructLM: Towards Building Generalist Models for Structured Knowledge Grounding
par: Zhuang, Alex, et autres
Publié: (2024)
par: Zhuang, Alex, et autres
Publié: (2024)
Read to Play (R2-Play): Decision Transformer with Multimodal Game Instruction
par: Jin, Yonggang, et autres
Publié: (2024)
par: Jin, Yonggang, et autres
Publié: (2024)
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
par: Qu, Xingwei, et autres
Publié: (2025)
par: Qu, Xingwei, et autres
Publié: (2025)
Long Context Alignment with Short Instructions and Synthesized Positions
par: Wu, Wenhao, et autres
Publié: (2024)
par: Wu, Wenhao, et autres
Publié: (2024)
BatCoder: Self-Supervised Bidirectional Code-Documentation Learning via Back-Translation
par: Xu, Jingwen, et autres
Publié: (2026)
par: Xu, Jingwen, et autres
Publié: (2026)
COIG-Writer: A High-Quality Dataset for Chinese Creative Writing with Thought Processes
par: Li, Yunwen, et autres
Publié: (2025)
par: Li, Yunwen, et autres
Publié: (2025)
Human-Instruction-Free LLM Self-Alignment with Limited Samples
par: Guo, Hongyi, et autres
Publié: (2024)
par: Guo, Hongyi, et autres
Publié: (2024)
Local Success Does Not Compose: Benchmarking Large Language Models for Compositional Formal Verification
par: Xu, Xu, et autres
Publié: (2025)
par: Xu, Xu, et autres
Publié: (2025)
ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation
par: Ren, Weiming, et autres
Publié: (2024)
par: Ren, Weiming, et autres
Publié: (2024)
Polish Translation Studies in Action
Publié: (2021)
Publié: (2021)
The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents
par: Wang, Xinrun, et autres
Publié: (2026)
par: Wang, Xinrun, et autres
Publié: (2026)
Slow-Fast Policy Optimization: Reposition-Before-Update for LLM Reasoning
par: Wang, Ziyan, et autres
Publié: (2025)
par: Wang, Ziyan, et autres
Publié: (2025)
Documents similaires
-
I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
par: Liang, Yiming, et autres
Publié: (2024) -
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
par: Zhang, Ge, et autres
Publié: (2024) -
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
par: Qu, Xingwei, et autres
Publié: (2024) -
Aligning Instruction Tuning with Pre-training
par: Liang, Yiming, et autres
Publié: (2025) -
CMDAG: A Chinese Metaphor Dataset with Annotated Grounds as CoT for Boosting Metaphor Generation
par: Shao, Yujie, et autres
Publié: (2024)