Forgetting as a Feature: Cognitive Alignment of Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Christoforos, Alexandros |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025)
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025)
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025)
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2026)
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2026)
Exploring Forgetting in Large Language Model Pre-Training
von: Liao, Chonghua, et al.
Veröffentlicht: (2024)
von: Liao, Chonghua, et al.
Veröffentlicht: (2024)
A Survey on Hardware Accelerators for Large Language Models
von: Kachris, Christoforos
Veröffentlicht: (2024)
von: Kachris, Christoforos
Veröffentlicht: (2024)
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
von: Wang, Shuo, et al.
Veröffentlicht: (2024)
von: Wang, Shuo, et al.
Veröffentlicht: (2024)
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
Talking to Yourself: Defying Forgetting in Large Language Models
von: Sun, Yutao, et al.
Veröffentlicht: (2026)
von: Sun, Yutao, et al.
Veröffentlicht: (2026)
Learning and Forgetting Unsafe Examples in Large Language Models
von: Zhao, Jiachen, et al.
Veröffentlicht: (2023)
von: Zhao, Jiachen, et al.
Veröffentlicht: (2023)
Examining Forgetting in Continual Pre-training of Aligned Large Language Models
von: Li, Chen-An, et al.
Veröffentlicht: (2024)
von: Li, Chen-An, et al.
Veröffentlicht: (2024)
Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models
von: Zhu, Didi, et al.
Veröffentlicht: (2024)
von: Zhu, Didi, et al.
Veröffentlicht: (2024)
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
Offline Learning and Forgetting for Reasoning with Large Language Models
von: Ni, Tianwei, et al.
Veröffentlicht: (2025)
von: Ni, Tianwei, et al.
Veröffentlicht: (2025)
Unveiling and Addressing Pseudo Forgetting in Large Language Models
von: Sun, Huashan, et al.
Veröffentlicht: (2024)
von: Sun, Huashan, et al.
Veröffentlicht: (2024)
Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration
von: Mao, Xin, et al.
Veröffentlicht: (2024)
von: Mao, Xin, et al.
Veröffentlicht: (2024)
Personality Alignment of Large Language Models
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
Pedagogical Alignment of Large Language Models
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
von: Luo, Yun, et al.
Veröffentlicht: (2023)
von: Luo, Yun, et al.
Veröffentlicht: (2023)
Erasing Without Remembering: Implicit Knowledge Forgetting in Large Language Models
von: Wang, Huazheng, et al.
Veröffentlicht: (2025)
von: Wang, Huazheng, et al.
Veröffentlicht: (2025)
Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
von: Huang, Jianheng, et al.
Veröffentlicht: (2024)
von: Huang, Jianheng, et al.
Veröffentlicht: (2024)
Model-Dowser: Data-Free Importance Probing to Mitigate Catastrophic Forgetting in Multimodal Large Language Models
von: Hwang, Hyeontaek, et al.
Veröffentlicht: (2026)
von: Hwang, Hyeontaek, et al.
Veröffentlicht: (2026)
Seeing Eye to AI: Human Alignment via Gaze-Based Response Rewards for Large Language Models
von: Lopez-Cardona, Angela, et al.
Veröffentlicht: (2024)
von: Lopez-Cardona, Angela, et al.
Veröffentlicht: (2024)
Hybrid Alignment Training for Large Language Models
von: Wang, Chenglong, et al.
Veröffentlicht: (2024)
von: Wang, Chenglong, et al.
Veröffentlicht: (2024)
Answer When Needed, Forget When Not: Language Models Pretend to Forget via In-Context Knowledge Unlearning
von: Takashiro, Shota, et al.
Veröffentlicht: (2024)
von: Takashiro, Shota, et al.
Veröffentlicht: (2024)
Investigating Cultural Alignment of Large Language Models
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting
von: Ding, Fei, et al.
Veröffentlicht: (2025)
von: Ding, Fei, et al.
Veröffentlicht: (2025)
Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding
von: Zhong, Lin, et al.
Veröffentlicht: (2026)
von: Zhong, Lin, et al.
Veröffentlicht: (2026)
Not Every Token Needs Forgetting: Selective Unlearning to Limit Change in Utility in Large Language Model Unlearning
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
Balancing Speciality and Versatility: A Coarse to Fine Framework for Mitigating Catastrophic Forgetting in Large Language Models
von: Zhang, Hengyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Hengyuan, et al.
Veröffentlicht: (2024)
Towards Efficient and Effective Alignment of Large Language Models
von: Jiang, Yuxin
Veröffentlicht: (2025)
von: Jiang, Yuxin
Veröffentlicht: (2025)
On the Alignment of Large Language Models with Global Human Opinion
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Alignment for Efficient Tool Calling of Large Language Models
von: Xu, Hongshen, et al.
Veröffentlicht: (2025)
von: Xu, Hongshen, et al.
Veröffentlicht: (2025)
Self-Pluralising Culture Alignment for Large Language Models
von: Xu, Shaoyang, et al.
Veröffentlicht: (2024)
von: Xu, Shaoyang, et al.
Veröffentlicht: (2024)
NILE: Internal Consistency Alignment in Large Language Models
von: Hu, Minda, et al.
Veröffentlicht: (2024)
von: Hu, Minda, et al.
Veröffentlicht: (2024)
BloomWise: Enhancing Problem-Solving capabilities of Large Language Models using Bloom's-Taxonomy-Inspired Prompts
von: Zoumpoulidi, Maria-Eleni, et al.
Veröffentlicht: (2024)
von: Zoumpoulidi, Maria-Eleni, et al.
Veröffentlicht: (2024)
Cognitive Memory in Large Language Models
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
Cognitive Effects in Large Language Models
von: Shaki, Jonathan, et al.
Veröffentlicht: (2023)
von: Shaki, Jonathan, et al.
Veröffentlicht: (2023)
To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
von: Tian, Bozhong, et al.
Veröffentlicht: (2024)
FIT to Forget: Robust Continual Unlearning for Large Language Models
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Xu, Xiaoyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025) -
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
von: Christoforos, Alexandros, et al.
Veröffentlicht: (2025) -
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2026) -
Exploring Forgetting in Large Language Model Pre-Training
von: Liao, Chonghua, et al.
Veröffentlicht: (2024) -
A Survey on Hardware Accelerators for Large Language Models
von: Kachris, Christoforos
Veröffentlicht: (2024)