Continual Learning with Global Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Xueying, Shang, Jinghuan, Sun, Yifan, Balasubramanian, Niranjan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does RoBERTa Perform Better than BERT in Continual Learning: An Attention Sink Perspective
von: Bai, Xueying, et al.
Veröffentlicht: (2024)
von: Bai, Xueying, et al.
Veröffentlicht: (2024)
Comparing Pre-trained Human Language Models: Is it Better with Human Context as Groups, Individual Traits, or Both?
von: Soni, Nikita, et al.
Veröffentlicht: (2024)
von: Soni, Nikita, et al.
Veröffentlicht: (2024)
Large Human Language Models: A Need and the Challenges
von: Soni, Nikita, et al.
Veröffentlicht: (2023)
von: Soni, Nikita, et al.
Veröffentlicht: (2023)
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
von: Soni, Nikita, et al.
Veröffentlicht: (2025)
von: Soni, Nikita, et al.
Veröffentlicht: (2025)
Addressing the Ecological Fallacy in Larger LMs with Human Context
von: Soni, Nikita, et al.
Veröffentlicht: (2026)
von: Soni, Nikita, et al.
Veröffentlicht: (2026)
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
von: Aakanksha, et al.
Veröffentlicht: (2024)
von: Aakanksha, et al.
Veröffentlicht: (2024)
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
A Common Pitfall of Margin-based Language Model Alignment: Gradient Entanglement
von: Yuan, Hui, et al.
Veröffentlicht: (2024)
von: Yuan, Hui, et al.
Veröffentlicht: (2024)
Aligner: Efficient Alignment by Learning to Correct
von: Ji, Jiaming, et al.
Veröffentlicht: (2024)
von: Ji, Jiaming, et al.
Veröffentlicht: (2024)
Sample-Efficient Alignment for LLMs
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Aligners: Decoupling LLMs and Alignment
von: Ngweta, Lilian, et al.
Veröffentlicht: (2024)
von: Ngweta, Lilian, et al.
Veröffentlicht: (2024)
Interpretable Safety Alignment via SAE-Constructed Low-Rank Subspace Adaptation
von: Wang, Dianyun, et al.
Veröffentlicht: (2025)
von: Wang, Dianyun, et al.
Veröffentlicht: (2025)
AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
von: Trivedi, Harsh, et al.
Veröffentlicht: (2024)
von: Trivedi, Harsh, et al.
Veröffentlicht: (2024)
From Instructions to Constraints: Language Model Alignment with Automatic Constraint Verification
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
Adversarial Preference Optimization: Enhancing Your Alignment via RM-LLM Game
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models
von: Liu, Qin, et al.
Veröffentlicht: (2024)
von: Liu, Qin, et al.
Veröffentlicht: (2024)
AlignBench: Benchmarking Chinese Alignment of Large Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
In-context Continual Learning Assisted by an External Continual Learner
von: Momeni, Saleh, et al.
Veröffentlicht: (2024)
von: Momeni, Saleh, et al.
Veröffentlicht: (2024)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
Learning to Align, Aligning to Learn: A Unified Approach for Self-Optimized Alignment
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment
von: Shen, Yunyi, et al.
Veröffentlicht: (2025)
von: Shen, Yunyi, et al.
Veröffentlicht: (2025)
Enhancing Time Series Forecasting via Multi-Level Text Alignment with LLMs
von: Zhao, Taibiao, et al.
Veröffentlicht: (2025)
von: Zhao, Taibiao, et al.
Veröffentlicht: (2025)
Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations
von: Luo, Haozheng, et al.
Veröffentlicht: (2026)
von: Luo, Haozheng, et al.
Veröffentlicht: (2026)
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
Self-Play Preference Optimization for Language Model Alignment
von: Wu, Yue, et al.
Veröffentlicht: (2024)
von: Wu, Yue, et al.
Veröffentlicht: (2024)
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
SALMON: Self-Alignment with Instructable Reward Models
von: Sun, Zhiqing, et al.
Veröffentlicht: (2023)
von: Sun, Zhiqing, et al.
Veröffentlicht: (2023)
Continuous QA Learning with Structured Prompts
von: Zheng, Yinhe
Veröffentlicht: (2022)
von: Zheng, Yinhe
Veröffentlicht: (2022)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
von: Pattnaik, Pulkit, et al.
Veröffentlicht: (2024)
von: Pattnaik, Pulkit, et al.
Veröffentlicht: (2024)
Human Inspired Progressive Alignment and Comparative Learning for Grounded Word Acquisition
von: Bao, Yuwei, et al.
Veröffentlicht: (2023)
von: Bao, Yuwei, et al.
Veröffentlicht: (2023)
DHA: Learning Decoupled-Head Attention from Transformer Checkpoints via Adaptive Heads Fusion
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
Reformatted Alignment
von: Fan, Run-Ze, et al.
Veröffentlicht: (2024)
von: Fan, Run-Ze, et al.
Veröffentlicht: (2024)
Detection Is Cheap, Routing Is Learned: Why Refusal-Based Alignment Evaluation Fails
von: Frank, Gregory N.
Veröffentlicht: (2026)
von: Frank, Gregory N.
Veröffentlicht: (2026)
Efficient Text-Attributed Graph Learning through Selective Annotation and Graph Alignment
von: Xie, Huanyi, et al.
Veröffentlicht: (2025)
von: Xie, Huanyi, et al.
Veröffentlicht: (2025)
TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling
von: Qiu, Jiahao, et al.
Veröffentlicht: (2024)
von: Qiu, Jiahao, et al.
Veröffentlicht: (2024)
Unlocking Continual Learning Abilities in Language Models
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Does RoBERTa Perform Better than BERT in Continual Learning: An Attention Sink Perspective
von: Bai, Xueying, et al.
Veröffentlicht: (2024) -
Comparing Pre-trained Human Language Models: Is it Better with Human Context as Groups, Individual Traits, or Both?
von: Soni, Nikita, et al.
Veröffentlicht: (2024) -
Large Human Language Models: A Need and the Challenges
von: Soni, Nikita, et al.
Veröffentlicht: (2023) -
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
von: Soni, Nikita, et al.
Veröffentlicht: (2025) -
Addressing the Ecological Fallacy in Larger LMs with Human Context
von: Soni, Nikita, et al.
Veröffentlicht: (2026)