Geometry Conflict: Explaining and Controlling Forgetting in LLM Continual Post-Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yuanyi, Yang, Yifan, Lu, Su, Gu, Yanggan, Wang, Pengkai, Wang, Wenjun, Yan, Zhaoyi, Xie, Congkai, Wu, Jianmin, Cao, Jialun, Cheung, Shing-Chi, Yang, Hongxia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
E-PMQ: Expert-Guided Post-Merge Quantization with Merged-Weight Anchoring
von: Wang, Wenjun, et al.
Veröffentlicht: (2026)
von: Wang, Wenjun, et al.
Veröffentlicht: (2026)
Access Sets Matter: Budgeting Expert Reads for Scalable Weight-Space Model Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
MergePipe: A Budget-Aware Parameter Management System for Scalable LLM Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
FeatCal: Feature Calibration for Post-Merging Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2026)
von: Gu, Yanggan, et al.
Veröffentlicht: (2026)
Model Merging Scaling Laws in Large Language Models
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
PMC-InterCPT: Rethinking Biomedical Interleaved Data for Multimodal Continued Pretraining
von: Zhu, Guanghao, et al.
Veröffentlicht: (2026)
von: Zhu, Guanghao, et al.
Veröffentlicht: (2026)
InfiFPO: Implicit Model Fusion via Preference Optimization in Large Language Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training
von: Wang, Pengkai, et al.
Veröffentlicht: (2025)
von: Wang, Pengkai, et al.
Veröffentlicht: (2025)
Concerned with Data Contamination? Assessing Countermeasures in Code Language Model
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
InfiCoEvalChain: A Blockchain-Based Decentralized Framework for Collaborative LLM Evaluation
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
ReuseDroid: A VLM-empowered Android UI Test Migrator Boosted by Active Feedback
von: Li, Xiaolei, et al.
Veröffentlicht: (2025)
von: Li, Xiaolei, et al.
Veröffentlicht: (2025)
Quantization Meets Reasoning: Exploring LLM Low-Bit Quantization Degradation for Mathematical Reasoning
von: Li, Zhen, et al.
Veröffentlicht: (2025)
von: Li, Zhen, et al.
Veröffentlicht: (2025)
What Builds Effective In-Context Examples for Code Generation?
von: Li, Dongze, et al.
Veröffentlicht: (2025)
von: Li, Dongze, et al.
Veröffentlicht: (2025)
ModelWisdom: An Integrated Toolkit for TLA+ Model Visualization, Digest and Repair
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
Amidated Pectin/Gelatin‐Based Films Containing Black Rice Anthocyanins and Their Coating Preservation Effect on Seriola aureovittata Fillets
von: Yi Yang, et al.
Veröffentlicht: (2025)
von: Yi Yang, et al.
Veröffentlicht: (2025)
InfiR2: A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
The Use of Digital Watermarking for Intelligence Multimedia Document Distribution
von: Shing-Chi Cheung
Veröffentlicht: (2008)
von: Shing-Chi Cheung
Veröffentlicht: (2008)
Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
When LLMs Meet API Documentation: Can Retrieval Augmentation Aid Code Generation Just as It Helps Developers?
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
Can Large Language Models Model Programs Formally?
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
InfiMed-Foundation: Pioneering Advanced Multimodal Medical Models with Compute-Efficient Pre-Training and Multi-Stage Fine-Tuning
von: Zhu, Guanghao, et al.
Veröffentlicht: (2025)
von: Zhu, Guanghao, et al.
Veröffentlicht: (2025)
High-order Joint Constituency and Dependency Parsing
von: Gu, Yanggan, et al.
Veröffentlicht: (2023)
von: Gu, Yanggan, et al.
Veröffentlicht: (2023)
CODECLEANER: Elevating Standards with A Robust Data Contamination Mitigation Toolkit
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
Isolating Language-Coding from Problem-Solving: Benchmarking LLMs with PseudoEval
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
Rethinking Local Learning: A Cheaper and Faster Recipe for LLM Post-Training
von: Shi, Hengyu, et al.
Veröffentlicht: (2026)
von: Shi, Hengyu, et al.
Veröffentlicht: (2026)
CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training
von: Thede, Lukas, et al.
Veröffentlicht: (2026)
von: Thede, Lukas, et al.
Veröffentlicht: (2026)
Can Emulating Semantic Translation Help LLMs with Code Translation? A Study Based on Pseudocode
von: Chen, Songqiang, et al.
Veröffentlicht: (2025)
von: Chen, Songqiang, et al.
Veröffentlicht: (2025)
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
von: Li, Meiziniu, et al.
Veröffentlicht: (2024)
von: Li, Meiziniu, et al.
Veröffentlicht: (2024)
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
von: Li, Meiziniu, et al.
Veröffentlicht: (2022)
von: Li, Meiziniu, et al.
Veröffentlicht: (2022)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
von: Lai, Song, et al.
Veröffentlicht: (2025)
von: Lai, Song, et al.
Veröffentlicht: (2025)
Matching Accuracy, Different Geometry: Evolution Strategies vs GRPO in LLM Post-Training
von: Hoy, William, et al.
Veröffentlicht: (2026)
von: Hoy, William, et al.
Veröffentlicht: (2026)
Do Not Forget About Politics! Explaining the Selective Breach of Bilateral Investment Treaties
von: Zhiyuan Wang
Veröffentlicht: (2025)
von: Zhiyuan Wang
Veröffentlicht: (2025)
InfiMed: Low-Resource Medical MLLMs with Advancing Understanding and Reasoning
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training
von: Pan, Chengjun, et al.
Veröffentlicht: (2026)
von: Pan, Chengjun, et al.
Veröffentlicht: (2026)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
JavaBench: A Benchmark of Object-Oriented Code Generation for Evaluating Large Language Models
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
PTQTP: Post-Training Quantization to Trit-Planes for Large Language Models
von: Xiao, He, et al.
Veröffentlicht: (2025)
von: Xiao, He, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
E-PMQ: Expert-Guided Post-Merge Quantization with Merged-Weight Anchoring
von: Wang, Wenjun, et al.
Veröffentlicht: (2026) -
Access Sets Matter: Budgeting Expert Reads for Scalable Weight-Space Model Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026) -
MergePipe: A Budget-Aware Parameter Management System for Scalable LLM Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026) -
FeatCal: Feature Calibration for Post-Merging Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2026) -
Model Merging Scaling Laws in Large Language Models
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)