Understanding Layer Significance in LLM Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Shi, Guangyuan, Lu, Zexin, Dong, Xiaoyu, Zhang, Wenlong, Zhang, Xuanyu, Feng, Yujie, Wu, Xiao-Ming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2025)
di: Feng, Yujie, et al.
Pubblicazione: (2025)
KIF: Knowledge Identification and Fusion for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2024)
di: Feng, Yujie, et al.
Pubblicazione: (2024)
AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2025)
di: Feng, Yujie, et al.
Pubblicazione: (2025)
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
di: Liu, Bo, et al.
Pubblicazione: (2024)
di: Liu, Bo, et al.
Pubblicazione: (2024)
LANID: LLM-assisted New Intent Discovery
di: Fan, Lu, et al.
Pubblicazione: (2025)
di: Fan, Lu, et al.
Pubblicazione: (2025)
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment
di: Chen, Nuo, et al.
Pubblicazione: (2024)
di: Chen, Nuo, et al.
Pubblicazione: (2024)
CSCE: Boosting LLM Reasoning by Simultaneous Enhancing of Causal Significance and Consistency
di: Wang, Kangsheng, et al.
Pubblicazione: (2024)
di: Wang, Kangsheng, et al.
Pubblicazione: (2024)
Hierarchical Alignment: Surgical Fine-Tuning via Functional Layer Specialization in Large Language Models
di: Zhang, Yukun, et al.
Pubblicazione: (2025)
di: Zhang, Yukun, et al.
Pubblicazione: (2025)
EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
di: Zhao, Xiangyu, et al.
Pubblicazione: (2023)
di: Zhao, Xiangyu, et al.
Pubblicazione: (2023)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models
di: Feng, Yujie, et al.
Pubblicazione: (2026)
di: Feng, Yujie, et al.
Pubblicazione: (2026)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
di: Taniguchi, Rei, et al.
Pubblicazione: (2026)
di: Taniguchi, Rei, et al.
Pubblicazione: (2026)
Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction
di: Hota, Asutosh, et al.
Pubblicazione: (2025)
di: Hota, Asutosh, et al.
Pubblicazione: (2025)
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
di: Xiao, Jianfei, et al.
Pubblicazione: (2026)
di: Xiao, Jianfei, et al.
Pubblicazione: (2026)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
di: Sun, Shuqiao, et al.
Pubblicazione: (2024)
di: Sun, Shuqiao, et al.
Pubblicazione: (2024)
TODO: Enhancing LLM Alignment with Ternary Preferences
di: Guo, Yuxiang, et al.
Pubblicazione: (2024)
di: Guo, Yuxiang, et al.
Pubblicazione: (2024)
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge
di: Chen, Luyu, et al.
Pubblicazione: (2025)
di: Chen, Luyu, et al.
Pubblicazione: (2025)
R2-Write: Reflection and Revision for Open-Ended Writing with Deep Reasoning
di: Liu, Wanlong, et al.
Pubblicazione: (2026)
di: Liu, Wanlong, et al.
Pubblicazione: (2026)
Preference Alignment Improves Language Model-Based TTS
di: Tian, Jinchuan, et al.
Pubblicazione: (2024)
di: Tian, Jinchuan, et al.
Pubblicazione: (2024)
Evaluating and Improving Cultural Awareness of Reward Models for LLM Alignment
di: Zhang, Hongbin, et al.
Pubblicazione: (2025)
di: Zhang, Hongbin, et al.
Pubblicazione: (2025)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
di: Ke, Pei, et al.
Pubblicazione: (2023)
di: Ke, Pei, et al.
Pubblicazione: (2023)
Teaching Your Models to Understand Code via Focal Preference Alignment
di: Wu, Jie, et al.
Pubblicazione: (2025)
di: Wu, Jie, et al.
Pubblicazione: (2025)
PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations
di: Wu, Yuhe, et al.
Pubblicazione: (2026)
di: Wu, Yuhe, et al.
Pubblicazione: (2026)
Condor: Enhance LLM Alignment with Knowledge-Driven Data Synthesis and Refinement
di: Cao, Maosong, et al.
Pubblicazione: (2025)
di: Cao, Maosong, et al.
Pubblicazione: (2025)
Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning
di: Huan, Maggie, et al.
Pubblicazione: (2025)
di: Huan, Maggie, et al.
Pubblicazione: (2025)
SpiroLLM: Finetuning Pretrained LLMs to Understand Spirogram Time Series with Clinical Validation in COPD Reporting
di: Mei, Shuhao, et al.
Pubblicazione: (2025)
di: Mei, Shuhao, et al.
Pubblicazione: (2025)
Entity Alignment with Noisy Annotations from Large Language Models
di: Chen, Shengyuan, et al.
Pubblicazione: (2024)
di: Chen, Shengyuan, et al.
Pubblicazione: (2024)
LightRetriever: A LLM-based Text Retrieval Architecture with Extremely Faster Query Inference
di: Ma, Guangyuan, et al.
Pubblicazione: (2025)
di: Ma, Guangyuan, et al.
Pubblicazione: (2025)
A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities
di: Chen, Jiaqi, et al.
Pubblicazione: (2026)
di: Chen, Jiaqi, et al.
Pubblicazione: (2026)
Rethinking Text-based Protein Understanding: Retrieval or LLM?
di: Wu, Juntong, et al.
Pubblicazione: (2025)
di: Wu, Juntong, et al.
Pubblicazione: (2025)
DPRF: A Generalizable Dynamic Persona Refinement Framework for Optimizing Behavior Alignment Between Personalized LLM Role-Playing Agents and Humans
di: Yao, Bingsheng, et al.
Pubblicazione: (2025)
di: Yao, Bingsheng, et al.
Pubblicazione: (2025)
AlignBench: Benchmarking Chinese Alignment of Large Language Models
di: Liu, Xiao, et al.
Pubblicazione: (2023)
di: Liu, Xiao, et al.
Pubblicazione: (2023)
Small Language Models Offer Significant Potential for Science Community
di: Zhang, Jian
Pubblicazione: (2025)
di: Zhang, Jian
Pubblicazione: (2025)
Efficient Knowledge Infusion via KG-LLM Alignment
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
SALMONN-omni: A Standalone Speech LLM without Codec Injection for Full-duplex Conversation
di: Yu, Wenyi, et al.
Pubblicazione: (2025)
di: Yu, Wenyi, et al.
Pubblicazione: (2025)
LLM-Align: Utilizing Large Language Models for Entity Alignment in Knowledge Graphs
di: Chen, Xuan, et al.
Pubblicazione: (2024)
di: Chen, Xuan, et al.
Pubblicazione: (2024)
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
di: Yu, Wenyi, et al.
Pubblicazione: (2024)
di: Yu, Wenyi, et al.
Pubblicazione: (2024)
Inference-time Alignment via Sparse Junction Steering
di: Hu, Runyi, et al.
Pubblicazione: (2026)
di: Hu, Runyi, et al.
Pubblicazione: (2026)
Multi-Scale Manifold Alignment for Interpreting Large Language Models: A Unified Information-Geometric Framework
di: Zhang, Yukun, et al.
Pubblicazione: (2025)
di: Zhang, Yukun, et al.
Pubblicazione: (2025)
Extracting and Understanding the Superficial Knowledge in Alignment
di: Chen, Runjin, et al.
Pubblicazione: (2025)
di: Chen, Runjin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2025) -
KIF: Knowledge Identification and Fusion for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2024) -
AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning
di: Feng, Yujie, et al.
Pubblicazione: (2025) -
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
di: Liu, Bo, et al.
Pubblicazione: (2024) -
LANID: LLM-assisted New Intent Discovery
di: Fan, Lu, et al.
Pubblicazione: (2025)