AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control
Fuente:
arXiv
Salvato in:
| Autori principali: | Bui, Quang-Hung, Ta, Anh Son |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AdaZeta: Adaptive Zeroth-Order Tensor-Train Adaption for Memory-Efficient Large Language Models Fine-Tuning
di: Yang, Yifan, et al.
Pubblicazione: (2024)
di: Yang, Yifan, et al.
Pubblicazione: (2024)
AdaSTaR: Adaptive Data Sampling for Training Self-Taught Reasoners
di: Koh, Woosung, et al.
Pubblicazione: (2025)
di: Koh, Woosung, et al.
Pubblicazione: (2025)
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
di: Son, Hyegang, et al.
Pubblicazione: (2024)
di: Son, Hyegang, et al.
Pubblicazione: (2024)
AdaBoN: Adaptive Best-of-N Alignment
di: Raman, Vinod, et al.
Pubblicazione: (2025)
di: Raman, Vinod, et al.
Pubblicazione: (2025)
AdaSPEC: Selective Knowledge Distillation for Efficient Speculative Decoders
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning
di: Dinh, Quang Minh, et al.
Pubblicazione: (2024)
di: Dinh, Quang Minh, et al.
Pubblicazione: (2024)
Reinforce-Ada: An Adaptive Sampling Framework under Non-linear RL Objectives
di: Xiong, Wei, et al.
Pubblicazione: (2025)
di: Xiong, Wei, et al.
Pubblicazione: (2025)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
di: Zhou, Hongyi, et al.
Pubblicazione: (2025)
di: Zhou, Hongyi, et al.
Pubblicazione: (2025)
Memory-Efficient LLM Training with Online Subspace Descent
di: Liang, Kaizhao, et al.
Pubblicazione: (2024)
di: Liang, Kaizhao, et al.
Pubblicazione: (2024)
NeuroAda: Activating Each Neuron's Potential for Parameter-Efficient Fine-Tuning
di: Zhang, Zhi, et al.
Pubblicazione: (2025)
di: Zhang, Zhi, et al.
Pubblicazione: (2025)
AdaCuRL: Adaptive Curriculum Reinforcement Learning with Invalid Sample Mitigation and Historical Revisiting
di: Li, Renda, et al.
Pubblicazione: (2025)
di: Li, Renda, et al.
Pubblicazione: (2025)
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models
di: Zhang, Jun, et al.
Pubblicazione: (2025)
di: Zhang, Jun, et al.
Pubblicazione: (2025)
A Collision-Free Hot-Tier Extension for Engram-Style Conditional Memory: A Controlled Study of Training Dynamics
di: Lin, Tao
Pubblicazione: (2026)
di: Lin, Tao
Pubblicazione: (2026)
AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation
di: Du, Weihua, et al.
Pubblicazione: (2026)
di: Du, Weihua, et al.
Pubblicazione: (2026)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
TableDART: Dynamic Adaptive Multi-Modal Routing for Table Understanding
di: Xing, Xiaobo, et al.
Pubblicazione: (2025)
di: Xing, Xiaobo, et al.
Pubblicazione: (2025)
Adapt-Pruner: Adaptive Structural Pruning for Efficient Small Language Model Training
di: Pan, Rui, et al.
Pubblicazione: (2025)
di: Pan, Rui, et al.
Pubblicazione: (2025)
Entropy Adaptive Decoding: Dynamic Model Switching for Efficient Inference
di: Simonds, Toby
Pubblicazione: (2025)
di: Simonds, Toby
Pubblicazione: (2025)
Branch-and-Browse: Efficient and Controllable Web Exploration with Tree-Structured Reasoning and Action Memory
di: He, Shiqi, et al.
Pubblicazione: (2025)
di: He, Shiqi, et al.
Pubblicazione: (2025)
Adaptive Soft Rolling KV Freeze with Entropy-Guided Recovery: Sublinear Memory Growth for Efficient LLM Inference
di: Metinov, Adilet, et al.
Pubblicazione: (2025)
di: Metinov, Adilet, et al.
Pubblicazione: (2025)
AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptation
di: Liu, Ziyun, et al.
Pubblicazione: (2026)
di: Liu, Ziyun, et al.
Pubblicazione: (2026)
NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions
di: Hou, Shizheng, et al.
Pubblicazione: (2026)
di: Hou, Shizheng, et al.
Pubblicazione: (2026)
Manual Verbalizer Enrichment for Few-Shot Text Classification
di: Nguyen, Quang Anh, et al.
Pubblicazione: (2024)
di: Nguyen, Quang Anh, et al.
Pubblicazione: (2024)
VietMix: A Naturally-Occurring Parallel Corpus and Augmentation Framework for Vietnamese-English Code-Mixed Machine Translation
di: Tran, Hieu, et al.
Pubblicazione: (2025)
di: Tran, Hieu, et al.
Pubblicazione: (2025)
Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models
di: Vendrell, Victor Conchello, et al.
Pubblicazione: (2026)
di: Vendrell, Victor Conchello, et al.
Pubblicazione: (2026)
Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control
di: Rezazadeh, Alireza, et al.
Pubblicazione: (2025)
di: Rezazadeh, Alireza, et al.
Pubblicazione: (2025)
POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation
di: Qiu, Zeju, et al.
Pubblicazione: (2026)
di: Qiu, Zeju, et al.
Pubblicazione: (2026)
MSSR: Memory-Aware Adaptive Replay for Continual LLM Fine-Tuning
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
An Adaptive Placement and Parallelism Framework for Accelerating RLHF Training
di: Xiao, Youshao, et al.
Pubblicazione: (2023)
di: Xiao, Youshao, et al.
Pubblicazione: (2023)
FlashSampling: Fast and Memory-Efficient Exact Sampling
di: Ruiz, Tomas, et al.
Pubblicazione: (2026)
di: Ruiz, Tomas, et al.
Pubblicazione: (2026)
Efficient Agent Training for Computer Use
di: He, Yanheng, et al.
Pubblicazione: (2025)
di: He, Yanheng, et al.
Pubblicazione: (2025)
How to Train Data-Efficient LLMs
di: Sachdeva, Noveen, et al.
Pubblicazione: (2024)
di: Sachdeva, Noveen, et al.
Pubblicazione: (2024)
UltraEdit: Training-, Subject-, and Memory-Free Lifelong Editing in Language Models
di: Gu, Xiaojie, et al.
Pubblicazione: (2025)
di: Gu, Xiaojie, et al.
Pubblicazione: (2025)
COAP: Memory-Efficient Training with Correlation-Aware Gradient Projection
di: Xiao, Jinqi, et al.
Pubblicazione: (2024)
di: Xiao, Jinqi, et al.
Pubblicazione: (2024)
LLMem: Estimating GPU Memory Usage for Fine-Tuning Pre-Trained LLMs
di: Kim, Taeho, et al.
Pubblicazione: (2024)
di: Kim, Taeho, et al.
Pubblicazione: (2024)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
di: Nguyen, Dang, et al.
Pubblicazione: (2024)
di: Nguyen, Dang, et al.
Pubblicazione: (2024)
The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
SibylSense: Adaptive Rubric Learning via Memory Tuning and Adversarial Probing
di: Xu, Yifei, et al.
Pubblicazione: (2026)
di: Xu, Yifei, et al.
Pubblicazione: (2026)
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
Entropy Meets Importance: A Unified Head Importance-Entropy Score for Stable and Efficient Transformer Pruning
di: Choi, Minsik, et al.
Pubblicazione: (2025)
di: Choi, Minsik, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AdaZeta: Adaptive Zeroth-Order Tensor-Train Adaption for Memory-Efficient Large Language Models Fine-Tuning
di: Yang, Yifan, et al.
Pubblicazione: (2024) -
AdaSTaR: Adaptive Data Sampling for Training Self-Taught Reasoners
di: Koh, Woosung, et al.
Pubblicazione: (2025) -
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
di: Son, Hyegang, et al.
Pubblicazione: (2024) -
AdaBoN: Adaptive Best-of-N Alignment
di: Raman, Vinod, et al.
Pubblicazione: (2025) -
AdaSPEC: Selective Knowledge Distillation for Efficient Speculative Decoders
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)