LSEBMCL: A Latent Space Energy-Based Model for Continual Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Xiaodi, Li, Dingcheng, Gao, Rujun, Zamani, Mahmoud, Khan, Latifur |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PPSEBM: An Energy-Based Model with Progressive Parameter Selection for Continual Learning
von: Li, Xiaodi, et al.
Veröffentlicht: (2025)
von: Li, Xiaodi, et al.
Veröffentlicht: (2025)
Adversarial Reinforcement Learning for Large Language Model Agent Safety
von: Wang, Zizhao, et al.
Veröffentlicht: (2025)
von: Wang, Zizhao, et al.
Veröffentlicht: (2025)
Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacks
von: Liu, Haoyu, et al.
Veröffentlicht: (2026)
von: Liu, Haoyu, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Backtracking Feedback
von: Sel, Bilgehan, et al.
Veröffentlicht: (2026)
von: Sel, Bilgehan, et al.
Veröffentlicht: (2026)
LatentBreak: Jailbreaking Large Language Models through Latent Space Feedback
von: Mura, Raffaele, et al.
Veröffentlicht: (2025)
von: Mura, Raffaele, et al.
Veröffentlicht: (2025)
Lightweight Retrieval-Augmented Generation and Large Language Model-Based Modeling for Scalable Patient-Trial Matching
von: Li, Xiaodi, et al.
Veröffentlicht: (2026)
von: Li, Xiaodi, et al.
Veröffentlicht: (2026)
Understanding Jailbreak Success: A Study of Latent Space Dynamics in Large Language Models
von: Ball, Sarah, et al.
Veröffentlicht: (2024)
von: Ball, Sarah, et al.
Veröffentlicht: (2024)
Large Language Models Explore by Latent Distilling
von: Zeng, Yuanhao, et al.
Veröffentlicht: (2026)
von: Zeng, Yuanhao, et al.
Veröffentlicht: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Leveraging Codebook Knowledge with NLI and ChatGPT for Zero-Shot Political Relation Classification
von: Hu, Yibo, et al.
Veröffentlicht: (2023)
von: Hu, Yibo, et al.
Veröffentlicht: (2023)
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
von: Li, Hengli, et al.
Veröffentlicht: (2025)
von: Li, Hengli, et al.
Veröffentlicht: (2025)
LLM-Match: An Open-Sourced Patient Matching Model Based on Large Language Models and Retrieval-Augmented Generation
von: Li, Xiaodi, et al.
Veröffentlicht: (2025)
von: Li, Xiaodi, et al.
Veröffentlicht: (2025)
In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
Debiasing Multilingual LLMs in Cross-lingual Latent Space
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
Deliberation in Latent Space via Differentiable Cache Augmentation
von: Liu, Luyang, et al.
Veröffentlicht: (2024)
von: Liu, Luyang, et al.
Veröffentlicht: (2024)
Learning Dynamics in Continual Pre-Training for Large Language Models
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
Local Topology Measures of Contextual Language Model Latent Spaces With Applications to Dialogue Term Extraction
von: Ruppik, Benjamin Matthias, et al.
Veröffentlicht: (2024)
von: Ruppik, Benjamin Matthias, et al.
Veröffentlicht: (2024)
Model-diff: A Tool for Comparative Study of Language Models in the Input Space
von: Liu, Weitang, et al.
Veröffentlicht: (2024)
von: Liu, Weitang, et al.
Veröffentlicht: (2024)
Self-Improving World Modelling with Latent Actions
von: Qiu, Yifu, et al.
Veröffentlicht: (2026)
von: Qiu, Yifu, et al.
Veröffentlicht: (2026)
Unifying Continuous and Discrete Text Diffusion with Non-simultaneous Diffusion Processes
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
Continuous Autoregressive Language Models
von: Shao, Chenze, et al.
Veröffentlicht: (2025)
von: Shao, Chenze, et al.
Veröffentlicht: (2025)
Masked Diffusion Models as Energy Minimization
von: Chen, Sitong, et al.
Veröffentlicht: (2025)
von: Chen, Sitong, et al.
Veröffentlicht: (2025)
Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions
von: Li, Yongqi, et al.
Veröffentlicht: (2026)
von: Li, Yongqi, et al.
Veröffentlicht: (2026)
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
von: Feng, Yujie, et al.
Veröffentlicht: (2026)
von: Feng, Yujie, et al.
Veröffentlicht: (2026)
Reasoning to Learn from Latent Thoughts
von: Ruan, Yangjun, et al.
Veröffentlicht: (2025)
von: Ruan, Yangjun, et al.
Veröffentlicht: (2025)
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning
von: Wang, Xinyi, et al.
Veröffentlicht: (2023)
von: Wang, Xinyi, et al.
Veröffentlicht: (2023)
Continual Learning of Large Language Models: A Comprehensive Survey
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
The Transfer Neurons Hypothesis: An Underlying Mechanism for Language Latent Space Transitions in Multilingual LLMs
von: Tezuka, Hinata, et al.
Veröffentlicht: (2025)
von: Tezuka, Hinata, et al.
Veröffentlicht: (2025)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Towards Compute-Optimal Many-Shot In-Context Learning
von: Golchin, Shahriar, et al.
Veröffentlicht: (2025)
von: Golchin, Shahriar, et al.
Veröffentlicht: (2025)
Unlocking Continual Learning Abilities in Language Models
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Supervised Reinforcement Learning: From Expert Trajectories to Step-wise Reasoning
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
Reasoning with Latent Thoughts: On the Power of Looped Transformers
von: Saunshi, Nikunj, et al.
Veröffentlicht: (2025)
von: Saunshi, Nikunj, et al.
Veröffentlicht: (2025)
Projected Autoregression: Autoregressive Language Generation in Continuous State Space
von: Naparstek, Oshri
Veröffentlicht: (2026)
von: Naparstek, Oshri
Veröffentlicht: (2026)
Empowering Diffusion Models on the Embedding Space for Text Generation
von: Gao, Zhujin, et al.
Veröffentlicht: (2022)
von: Gao, Zhujin, et al.
Veröffentlicht: (2022)
Latent Context Compilation: Distilling Long Context into Compact Portable Memory
von: Li, Zeju, et al.
Veröffentlicht: (2026)
von: Li, Zeju, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PPSEBM: An Energy-Based Model with Progressive Parameter Selection for Continual Learning
von: Li, Xiaodi, et al.
Veröffentlicht: (2025) -
Adversarial Reinforcement Learning for Large Language Model Agent Safety
von: Wang, Zizhao, et al.
Veröffentlicht: (2025) -
Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacks
von: Liu, Haoyu, et al.
Veröffentlicht: (2026) -
Reinforcement Learning with Backtracking Feedback
von: Sel, Bilgehan, et al.
Veröffentlicht: (2026) -
LatentBreak: Jailbreaking Large Language Models through Latent Space Feedback
von: Mura, Raffaele, et al.
Veröffentlicht: (2025)