Continual Learning for Large Language Models: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Tongtong, Luo, Linhao, Li, Yuan-Fang, Pan, Shirui, Vu, Thuy-Trang, Haffari, Gholamreza |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Active Continual Learning: On Balancing Knowledge Retention and Learnability
von: Vu, Thuy-Trang, et al.
Veröffentlicht: (2023)
von: Vu, Thuy-Trang, et al.
Veröffentlicht: (2023)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
Beyond Imitation: Recovering Dense Rewards from Demonstrations
von: Li, Jiangnan, et al.
Veröffentlicht: (2025)
von: Li, Jiangnan, et al.
Veröffentlicht: (2025)
CARD: Towards Conditional Design of Multi-agent Topological Structures
von: Wu, Tongtong, et al.
Veröffentlicht: (2026)
von: Wu, Tongtong, et al.
Veröffentlicht: (2026)
AIPO: Learning to Reason from Active Interaction
von: Liu, Junnan, et al.
Veröffentlicht: (2026)
von: Liu, Junnan, et al.
Veröffentlicht: (2026)
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
von: Liu, Junnan, et al.
Veröffentlicht: (2025)
von: Liu, Junnan, et al.
Veröffentlicht: (2025)
Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Sampler
von: Wang, Minghan, et al.
Veröffentlicht: (2026)
von: Wang, Minghan, et al.
Veröffentlicht: (2026)
Graph-constrained Reasoning: Faithful Reasoning on Knowledge Graphs with Large Language Models
von: Luo, Linhao, et al.
Veröffentlicht: (2024)
von: Luo, Linhao, et al.
Veröffentlicht: (2024)
ChatRule: Mining Logical Rules with Large Language Models for Knowledge Graph Reasoning
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
Adapting Large Language Models for Document-Level Machine Translation
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
Towards Inference-time Scaling for Continuous Space Reasoning
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
Discrete Minds in a Continuous World: Do Language Models Know Time Passes?
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
Conversational SimulMT: Efficient Simultaneous Translation with Large Language Models
von: Wang, Minghan, et al.
Veröffentlicht: (2024)
von: Wang, Minghan, et al.
Veröffentlicht: (2024)
The Best of Both Worlds: Bridging Quality and Diversity in Data Selection with Bipartite Graph
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
Direct Evaluation of Chain-of-Thought in Multi-hop Reasoning with Knowledge Graphs
von: Nguyen, Minh-Vuong, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh-Vuong, et al.
Veröffentlicht: (2024)
Simultaneous Machine Translation with Large Language Models
von: Wang, Minghan, et al.
Veröffentlicht: (2023)
von: Wang, Minghan, et al.
Veröffentlicht: (2023)
Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models
von: Jin, Rihui, et al.
Veröffentlicht: (2025)
von: Jin, Rihui, et al.
Veröffentlicht: (2025)
SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Models
von: Li, Zhuang, et al.
Veröffentlicht: (2024)
von: Li, Zhuang, et al.
Veröffentlicht: (2024)
MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
von: Liu, Akide, et al.
Veröffentlicht: (2024)
von: Liu, Akide, et al.
Veröffentlicht: (2024)
Extending LLMs to New Languages: A Case Study of Llama and Persian Adaptation
von: Sani, Samin Mahdizadeh, et al.
Veröffentlicht: (2024)
von: Sani, Samin Mahdizadeh, et al.
Veröffentlicht: (2024)
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation
von: Luo, Linhao, et al.
Veröffentlicht: (2025)
von: Luo, Linhao, et al.
Veröffentlicht: (2025)
Exploring the Potential of Multimodal LLM with Knowledge-Intensive Multimodal ASR
von: Wang, Minghan, et al.
Veröffentlicht: (2024)
von: Wang, Minghan, et al.
Veröffentlicht: (2024)
Scalable Frame-based Construction of Sociocultural NormBases for Socially-Aware Dialogues
von: Qu, Shilin, et al.
Veröffentlicht: (2024)
von: Qu, Shilin, et al.
Veröffentlicht: (2024)
IntroLM: Introspective Language Models via Prefilling-Time Self-Evaluation
von: Kasnavieh, Hossein Hosseini, et al.
Veröffentlicht: (2026)
von: Kasnavieh, Hossein Hosseini, et al.
Veröffentlicht: (2026)
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
von: Feng, Tao, et al.
Veröffentlicht: (2025)
von: Feng, Tao, et al.
Veröffentlicht: (2025)
CONGRAD:Conflicting Gradient Filtering for Multilingual Preference Alignment
von: Li, Jiangnan, et al.
Veröffentlicht: (2025)
von: Li, Jiangnan, et al.
Veröffentlicht: (2025)
PromptDSI: Prompt-based Rehearsal-free Continual Learning for Document Retrieval
von: Huynh, Tuan-Luc, et al.
Veröffentlicht: (2024)
von: Huynh, Tuan-Luc, et al.
Veröffentlicht: (2024)
SpeechDialogueFactory: Generating High-Quality Speech Dialogue Data to Accelerate Your Speech-LLM Development
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
Continual Learning of Large Language Models: A Comprehensive Survey
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
$\text{M}^{2}$LLM: Multi-view Molecular Representation Learning with Large Language Models
von: Ju, Jiaxin, et al.
Veröffentlicht: (2025)
von: Ju, Jiaxin, et al.
Veröffentlicht: (2025)
Resurfacing Paralinguistic Awareness in Large Audio Language Models
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
Continual Speech Learning with Fused Speech Features
von: Wang, Guitao, et al.
Veröffentlicht: (2025)
von: Wang, Guitao, et al.
Veröffentlicht: (2025)
A Survey on Training-free Alignment of Large Language Models
von: Pan, Birong, et al.
Veröffentlicht: (2025)
von: Pan, Birong, et al.
Veröffentlicht: (2025)
Unifying Large Language Models and Knowledge Graphs: A Roadmap
von: Pan, Shirui, et al.
Veröffentlicht: (2023)
von: Pan, Shirui, et al.
Veröffentlicht: (2023)
Towards Lifelong Learning of Large Language Models: A Survey
von: Zheng, Junhao, et al.
Veröffentlicht: (2024)
von: Zheng, Junhao, et al.
Veröffentlicht: (2024)
MixLoRA-DSI: Dynamically Expandable Mixture-of-LoRA Experts for Rehearsal-Free Generative Retrieval over Dynamic Corpora
von: Huynh, Tuan-Luc, et al.
Veröffentlicht: (2025)
von: Huynh, Tuan-Luc, et al.
Veröffentlicht: (2025)
Beyond a Single Perspective: Text Anomaly Detection with Multi-View Language Representations
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
Modelling Political Coalition Negotiations Using LLM-based Agents
von: Moghimifar, Farhad, et al.
Veröffentlicht: (2024)
von: Moghimifar, Farhad, et al.
Veröffentlicht: (2024)
How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey
von: Qi, Yayun, et al.
Veröffentlicht: (2024)
von: Qi, Yayun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Active Continual Learning: On Balancing Knowledge Retention and Learnability
von: Vu, Thuy-Trang, et al.
Veröffentlicht: (2023) -
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
von: Luo, Linhao, et al.
Veröffentlicht: (2023) -
Beyond Imitation: Recovering Dense Rewards from Demonstrations
von: Li, Jiangnan, et al.
Veröffentlicht: (2025) -
CARD: Towards Conditional Design of Multi-agent Topological Structures
von: Wu, Tongtong, et al.
Veröffentlicht: (2026) -
AIPO: Learning to Reason from Active Interaction
von: Liu, Junnan, et al.
Veröffentlicht: (2026)