Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Hyunwoo, Lee, Junha, Choi, Mincheol, Lee, Jeonghwan, Cho, Jaeshin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
di: Lee, Hyunwoo, et al.
Pubblicazione: (2025)
di: Lee, Hyunwoo, et al.
Pubblicazione: (2025)
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
di: Lee, Hyunwoo, et al.
Pubblicazione: (2024)
di: Lee, Hyunwoo, et al.
Pubblicazione: (2024)
Memory-Efficient Acceleration of Block Low-Rank Foundation Models on Resource Constrained GPUs
di: Abillama, Pierre, et al.
Pubblicazione: (2025)
di: Abillama, Pierre, et al.
Pubblicazione: (2025)
Frictional Q-Learning
di: Kim, Hyunwoo, et al.
Pubblicazione: (2025)
di: Kim, Hyunwoo, et al.
Pubblicazione: (2025)
Learning to Transfer Human Hand Skills for Robot Manipulations
di: Park, Sungjae, et al.
Pubblicazione: (2025)
di: Park, Sungjae, et al.
Pubblicazione: (2025)
Temporal Dynamic Embedding for Irregularly Sampled Time Series
di: Kim, Mincheol, et al.
Pubblicazione: (2025)
di: Kim, Mincheol, et al.
Pubblicazione: (2025)
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
di: Sander, Jacob, et al.
Pubblicazione: (2025)
di: Sander, Jacob, et al.
Pubblicazione: (2025)
Constant Acceleration Flow
di: Park, Dogyun, et al.
Pubblicazione: (2024)
di: Park, Dogyun, et al.
Pubblicazione: (2024)
Revisiting Weight Averaging for Model Merging
di: Choi, Jiho, et al.
Pubblicazione: (2024)
di: Choi, Jiho, et al.
Pubblicazione: (2024)
LiteVLM: A Low-Latency Vision-Language Model Inference Pipeline for Resource-Constrained Environments
di: Huang, Jin, et al.
Pubblicazione: (2025)
di: Huang, Jin, et al.
Pubblicazione: (2025)
Sparse Weight Averaging with Multiple Particles for Iterative Magnitude Pruning
di: Choi, Moonseok, et al.
Pubblicazione: (2023)
di: Choi, Moonseok, et al.
Pubblicazione: (2023)
TRACED: Transition-aware Regret Approximation with Co-learnability for Environment Design
di: Cho, Geonwoo, et al.
Pubblicazione: (2025)
di: Cho, Geonwoo, et al.
Pubblicazione: (2025)
Wafer-Level Etch Spatial Profiling for Process Monitoring from Time-Series with Time-LLM
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
Inversion-based Latent Bayesian Optimization
di: Chu, Jaewon, et al.
Pubblicazione: (2024)
di: Chu, Jaewon, et al.
Pubblicazione: (2024)
Assigning Distinct Roles to Quantized and Low-Rank Matrices Toward Optimal Weight Decomposition
di: Cho, Yoonjun, et al.
Pubblicazione: (2025)
di: Cho, Yoonjun, et al.
Pubblicazione: (2025)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
di: Lee, Hosung, et al.
Pubblicazione: (2024)
di: Lee, Hosung, et al.
Pubblicazione: (2024)
Learning Equi-angular Representations for Online Continual Learning
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
Learning to Act Robustly with View-Invariant Latent Actions
di: Jeong, Youngjoon, et al.
Pubblicazione: (2026)
di: Jeong, Youngjoon, et al.
Pubblicazione: (2026)
MONAQ: Multi-Objective Neural Architecture Querying for Time-Series Analysis on Resource-Constrained Devices
di: Trirat, Patara, et al.
Pubblicazione: (2025)
di: Trirat, Patara, et al.
Pubblicazione: (2025)
RILQ: Rank-Insensitive LoRA-based Quantization Error Compensation for Boosting 2-bit Large Language Model Accuracy
di: Lee, Geonho, et al.
Pubblicazione: (2024)
di: Lee, Geonho, et al.
Pubblicazione: (2024)
Rethinking Multimodal Fusion for Time Series: Auxiliary Modalities Need Constrained Fusion
di: Lee, Seunghan, et al.
Pubblicazione: (2026)
di: Lee, Seunghan, et al.
Pubblicazione: (2026)
TraM : Enhancing User Sleep Prediction with Transformer-based Multivariate Time Series Modeling and Machine Learning Ensembles
di: Kim, Jinjae, et al.
Pubblicazione: (2024)
di: Kim, Jinjae, et al.
Pubblicazione: (2024)
Generating Multi-Table Time Series EHR from Latent Space with Minimal Preprocessing
di: Cho, Eunbyeol, et al.
Pubblicazione: (2025)
di: Cho, Eunbyeol, et al.
Pubblicazione: (2025)
Self-Supervised Pre-Training for Precipitation Post-Processor
di: An, Sojung, et al.
Pubblicazione: (2023)
di: An, Sojung, et al.
Pubblicazione: (2023)
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
di: Chen, Lei, et al.
Pubblicazione: (2026)
di: Chen, Lei, et al.
Pubblicazione: (2026)
Adaptive Sparsified Graph Learning Framework for Vessel Behavior Anomalies
di: Kim, Jeehong, et al.
Pubblicazione: (2025)
di: Kim, Jeehong, et al.
Pubblicazione: (2025)
Fast Training of Sinusoidal Neural Fields via Scaling Initialization
di: Yeom, Taesun, et al.
Pubblicazione: (2024)
di: Yeom, Taesun, et al.
Pubblicazione: (2024)
Multidimensional Adaptive Coefficient for Inference Trajectory Optimization in Flow and Diffusion
di: Lee, Dohoon, et al.
Pubblicazione: (2024)
di: Lee, Dohoon, et al.
Pubblicazione: (2024)
Grokfast: Accelerated Grokking by Amplifying Slow Gradients
di: Lee, Jaerin, et al.
Pubblicazione: (2024)
di: Lee, Jaerin, et al.
Pubblicazione: (2024)
WINA: Weight Informed Neuron Activation for Accelerating Large Language Model Inference
di: Chen, Sihan, et al.
Pubblicazione: (2025)
di: Chen, Sihan, et al.
Pubblicazione: (2025)
MetaCLBench: Meta Continual Learning Benchmark on Resource-Constrained Edge Devices
di: Li, Sijia, et al.
Pubblicazione: (2025)
di: Li, Sijia, et al.
Pubblicazione: (2025)
Leveraging Knowledge Distillation for Efficient Deep Reinforcement Learning in Resource-Constrained Environments
di: Meng, Guanlin
Pubblicazione: (2023)
di: Meng, Guanlin
Pubblicazione: (2023)
FALQON: Accelerating LoRA Fine-tuning with Low-Bit Floating-Point Arithmetic
di: Choi, Kanghyun, et al.
Pubblicazione: (2025)
di: Choi, Kanghyun, et al.
Pubblicazione: (2025)
Perturb-and-Compare Approach for Detecting Out-of-Distribution Samples in Constrained Access Environments
di: Lee, Heeyoung, et al.
Pubblicazione: (2024)
di: Lee, Heeyoung, et al.
Pubblicazione: (2024)
Latent Bayesian Optimization via Autoregressive Normalizing Flows
di: Lee, Seunghun, et al.
Pubblicazione: (2025)
di: Lee, Seunghun, et al.
Pubblicazione: (2025)
STEER: Inference-Time Risk Control via Constrained Quality-Diversity Search
di: Yang, Eric, et al.
Pubblicazione: (2026)
di: Yang, Eric, et al.
Pubblicazione: (2026)
NNiT: Width-Agnostic Neural Network Generation with Structurally Aligned Weight Spaces
di: Kim, Jiwoo, et al.
Pubblicazione: (2026)
di: Kim, Jiwoo, et al.
Pubblicazione: (2026)
LION-DG: Layer-Informed Initialization with Deep Gradient Protocols for Accelerated Neural Network Training
di: Kim, Hyunjun
Pubblicazione: (2026)
di: Kim, Hyunjun
Pubblicazione: (2026)
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference
di: Lee, Jaewoo, et al.
Pubblicazione: (2026)
di: Lee, Jaewoo, et al.
Pubblicazione: (2026)
Protein Language Models Diverge from Natural Language: Comparative Analysis and Improved Inference
di: Hart, Anna, et al.
Pubblicazione: (2026)
di: Hart, Anna, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
di: Lee, Hyunwoo, et al.
Pubblicazione: (2025) -
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
di: Lee, Hyunwoo, et al.
Pubblicazione: (2024) -
Memory-Efficient Acceleration of Block Low-Rank Foundation Models on Resource Constrained GPUs
di: Abillama, Pierre, et al.
Pubblicazione: (2025) -
Frictional Q-Learning
di: Kim, Hyunwoo, et al.
Pubblicazione: (2025) -
Learning to Transfer Human Hand Skills for Robot Manipulations
di: Park, Sungjae, et al.
Pubblicazione: (2025)