MUSCLE: A Model Update Strategy for Compatible LLM Evolution
Fuente:
arXiv
Salvato in:
| Autori principali: | Echterhoff, Jessica, Faghri, Fartash, Vemulapalli, Raviteja, Hu, Ting-Yao, Li, Chun-Liang, Tuzel, Oncel, Pouransari, Hadi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MobileCLIP: Fast Image-Text Models through Multi-Modal Reinforced Training
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2023)
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2023)
Knowledge Transfer from Vision Foundation Models for Efficient Training of Small Task-specific Models
di: Vemulapalli, Raviteja, et al.
Pubblicazione: (2023)
di: Vemulapalli, Raviteja, et al.
Pubblicazione: (2023)
TiC-CLIP: Continual Training of CLIP Models
di: Garg, Saurabh, et al.
Pubblicazione: (2023)
di: Garg, Saurabh, et al.
Pubblicazione: (2023)
CLIP with Quality Captions: A Strong Pretraining for Vision Tasks
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2024)
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2024)
AMUSE: Audio-Visual Benchmark and Alignment Framework for Agentic Multi-Speaker Understanding
di: Chowdhury, Sanjoy, et al.
Pubblicazione: (2025)
di: Chowdhury, Sanjoy, et al.
Pubblicazione: (2025)
MobileCLIP2: Improving Multi-Modal Reinforced Training
di: Faghri, Fartash, et al.
Pubblicazione: (2025)
di: Faghri, Fartash, et al.
Pubblicazione: (2025)
FocalLens: Instruction Tuning Enables Zero-Shot Conditional Image Representations
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2025)
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2025)
Learning from Self Critique and Refinement for Faithful LLM Summarization
di: Hu, Ting-Yao, et al.
Pubblicazione: (2025)
di: Hu, Ting-Yao, et al.
Pubblicazione: (2025)
Proxy-FDA: Proxy-based Feature Distribution Alignment for Fine-tuning Vision Foundation Models without Forgetting
di: Huang, Chen, et al.
Pubblicazione: (2025)
di: Huang, Chen, et al.
Pubblicazione: (2025)
SAM-CLIP: Merging Vision Foundation Models towards Semantic and Spatial Understanding
di: Wang, Haoxiang, et al.
Pubblicazione: (2023)
di: Wang, Haoxiang, et al.
Pubblicazione: (2023)
TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining
di: Li, Jeffrey, et al.
Pubblicazione: (2025)
di: Li, Jeffrey, et al.
Pubblicazione: (2025)
FastVLM: Efficient Vision Encoding for Vision Language Models
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2024)
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2024)
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization
di: Lu, Yen-Ju, et al.
Pubblicazione: (2025)
di: Lu, Yen-Ju, et al.
Pubblicazione: (2025)
Learning to Reason for Hallucination Span Detection
di: Su, Hsuan, et al.
Pubblicazione: (2025)
di: Su, Hsuan, et al.
Pubblicazione: (2025)
Pretraining with hierarchical memories: separating long-tail and common knowledge
di: Pouransari, Hadi, et al.
Pubblicazione: (2025)
di: Pouransari, Hadi, et al.
Pubblicazione: (2025)
VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2026)
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2026)
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
di: Mehta, Sachin, et al.
Pubblicazione: (2024)
di: Mehta, Sachin, et al.
Pubblicazione: (2024)
ASTRA-bench: Evaluating Tool-Use Agent Reasoning and Action Planning with Personal User Context
di: Xiu, Zidi, et al.
Pubblicazione: (2026)
di: Xiu, Zidi, et al.
Pubblicazione: (2026)
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
di: Mirzadeh, Iman, et al.
Pubblicazione: (2024)
di: Mirzadeh, Iman, et al.
Pubblicazione: (2024)
LiTo: Surface Light Field Tokenization
di: Chang, Jen-Hao Rick, et al.
Pubblicazione: (2026)
di: Chang, Jen-Hao Rick, et al.
Pubblicazione: (2026)
Dataset Decomposition: Faster LLM Training with Variable Sequence Length Curriculum
di: Pouransari, Hadi, et al.
Pubblicazione: (2024)
di: Pouransari, Hadi, et al.
Pubblicazione: (2024)
Beyond a Single Extractor: Re-thinking HTML-to-Text Extraction for LLM Pretraining
di: Li, Jeffrey, et al.
Pubblicazione: (2026)
di: Li, Jeffrey, et al.
Pubblicazione: (2026)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
di: Samragh, Mohammad, et al.
Pubblicazione: (2024)
di: Samragh, Mohammad, et al.
Pubblicazione: (2024)
Cognitive Bias in Decision-Making with LLMs
di: Echterhoff, Jessica, et al.
Pubblicazione: (2024)
di: Echterhoff, Jessica, et al.
Pubblicazione: (2024)
Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity
di: Joudaki, Amir, et al.
Pubblicazione: (2025)
di: Joudaki, Amir, et al.
Pubblicazione: (2025)
LLM-Driven Kernel Evolution: Automating Driver Updates in Linux
di: Kharlamova, Arina, et al.
Pubblicazione: (2025)
di: Kharlamova, Arina, et al.
Pubblicazione: (2025)
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
Evaluating Large Language Models as Generative User Simulators for Conversational Recommendation
di: Yoon, Se-eun, et al.
Pubblicazione: (2024)
di: Yoon, Se-eun, et al.
Pubblicazione: (2024)
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
di: Armandpour, Mohammadreza, et al.
Pubblicazione: (2026)
di: Armandpour, Mohammadreza, et al.
Pubblicazione: (2026)
Quantifying Cognitive Bias Induction in LLM-Generated Content
di: Alessa, Abeer, et al.
Pubblicazione: (2025)
di: Alessa, Abeer, et al.
Pubblicazione: (2025)
Abstract Concept Modelling in Conceptual Spaces: A Study on Chess Strategies
di: Banaee, Hadi, et al.
Pubblicazione: (2026)
di: Banaee, Hadi, et al.
Pubblicazione: (2026)
Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents
di: Lin, Minhua, et al.
Pubblicazione: (2026)
di: Lin, Minhua, et al.
Pubblicazione: (2026)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
di: Schweighofer, Kajetan, et al.
Pubblicazione: (2026)
di: Schweighofer, Kajetan, et al.
Pubblicazione: (2026)
Learning Agent-Compatible Context Management for Long-Horizon Tasks
di: Yi, Lu, et al.
Pubblicazione: (2026)
di: Yi, Lu, et al.
Pubblicazione: (2026)
GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations
di: Carrasco, Alejandro, et al.
Pubblicazione: (2026)
di: Carrasco, Alejandro, et al.
Pubblicazione: (2026)
BoostLoRA: Growing Effective Rank by Boosting Adapters
di: Anantha, Raviteja, et al.
Pubblicazione: (2026)
di: Anantha, Raviteja, et al.
Pubblicazione: (2026)
Frequency-Aware Masked Autoencoders for Multimodal Pretraining on Biosignals
di: Liu, Ran, et al.
Pubblicazione: (2023)
di: Liu, Ran, et al.
Pubblicazione: (2023)
CAMPHOR: Collaborative Agents for Multi-input Planning and High-Order Reasoning On Device
di: Fu, Yicheng, et al.
Pubblicazione: (2024)
di: Fu, Yicheng, et al.
Pubblicazione: (2024)
Synth4Seg -- Learning Defect Data Synthesis for Defect Segmentation using Bi-level Optimization
di: Mou, Shancong, et al.
Pubblicazione: (2024)
di: Mou, Shancong, et al.
Pubblicazione: (2024)
Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning
di: Qiu, Xin, et al.
Pubblicazione: (2025)
di: Qiu, Xin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MobileCLIP: Fast Image-Text Models through Multi-Modal Reinforced Training
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2023) -
Knowledge Transfer from Vision Foundation Models for Efficient Training of Small Task-specific Models
di: Vemulapalli, Raviteja, et al.
Pubblicazione: (2023) -
TiC-CLIP: Continual Training of CLIP Models
di: Garg, Saurabh, et al.
Pubblicazione: (2023) -
CLIP with Quality Captions: A Strong Pretraining for Vision Tasks
di: Vasu, Pavan Kumar Anasosalu, et al.
Pubblicazione: (2024) -
AMUSE: Audio-Visual Benchmark and Alignment Framework for Agentic Multi-Speaker Understanding
di: Chowdhury, Sanjoy, et al.
Pubblicazione: (2025)