Saved in:
| Main Authors: | Kim, Youngeun, Lee, Seunghwan, Jung, Aecheon, Ryu, Bogon, Hong, Sungeun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.06921 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SyMerge: From Non-Interference to Synergistic Merging via Single-Layer Adaptation
by: Jung, Aecheon, et al.
Published: (2024)
by: Jung, Aecheon, et al.
Published: (2024)
Dynamic Rank Adjustment for Accurate and Efficient Neural Network Training
by: Shin, Hyuntak, et al.
Published: (2025)
by: Shin, Hyuntak, et al.
Published: (2025)
ZOO-Prune: Training-Free Token Pruning via Zeroth-Order Gradient Estimation in Vision-Language Models
by: Kim, Youngeun, et al.
Published: (2025)
by: Kim, Youngeun, et al.
Published: (2025)
Backpropagation-Free Test-Time Adaptation via Probabilistic Gaussian Alignment
by: Zhang, Youjia, et al.
Published: (2025)
by: Zhang, Youjia, et al.
Published: (2025)
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks
by: Jung, Aecheon, et al.
Published: (2025)
by: Jung, Aecheon, et al.
Published: (2025)
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens
by: Kim, Jaehyeon, et al.
Published: (2024)
by: Kim, Jaehyeon, et al.
Published: (2024)
MC-GRPO: Median-Centered Group Relative Policy Optimization for Small-Rollout Reinforcement Learning
by: Kim, Youngeun
Published: (2026)
by: Kim, Youngeun
Published: (2026)
CodeGEMM: A Codebook-Centric Approach to Efficient GEMM in Quantized LLMs
by: Park, Gunho, et al.
Published: (2025)
by: Park, Gunho, et al.
Published: (2025)
Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression
by: Gao, Junqi, et al.
Published: (2026)
by: Gao, Junqi, et al.
Published: (2026)
Beta-Sigma VAE: Separating beta and decoder variance in Gaussian variational autoencoder
by: Kim, Seunghwan, et al.
Published: (2024)
by: Kim, Seunghwan, et al.
Published: (2024)
Task Singular Vectors: Reducing Task Interference in Model Merging
by: Gargiulo, Antonio Andrea, et al.
Published: (2024)
by: Gargiulo, Antonio Andrea, et al.
Published: (2024)
LongVQ: Long Sequence Modeling with Vector Quantization on Structured Memory
by: Liu, Zicheng, et al.
Published: (2024)
by: Liu, Zicheng, et al.
Published: (2024)
Adaptive Task Vectors for Large Language Models
by: Kang, Joonseong, et al.
Published: (2025)
by: Kang, Joonseong, et al.
Published: (2025)
Revisiting Weight Averaging for Model Merging
by: Choi, Jiho, et al.
Published: (2024)
by: Choi, Jiho, et al.
Published: (2024)
ReSpike: Residual Frames-based Hybrid Spiking Neural Networks for Efficient Action Recognition
by: Xiao, Shiting, et al.
Published: (2024)
by: Xiao, Shiting, et al.
Published: (2024)
DisTaC: Conditioning Task Vectors via Distillation for Robust Model Merging
by: Yoshida, Kotaro, et al.
Published: (2025)
by: Yoshida, Kotaro, et al.
Published: (2025)
SBVR: Summation of BitVector Representation for Efficient LLM Quantization
by: Bang, Wonjun, et al.
Published: (2025)
by: Bang, Wonjun, et al.
Published: (2025)
MSQ: Memory-Efficient Bit Sparsification Quantization
by: Han, Seokho, et al.
Published: (2025)
by: Han, Seokho, et al.
Published: (2025)
Less is More: Efficient Model Merging with Binary Task Switch
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Diet Your LLM: Dimension-wise Global Pruning of LLMs via Merging Task-specific Importance Score
by: Hong, Jimyung, et al.
Published: (2026)
by: Hong, Jimyung, et al.
Published: (2026)
Rethinking Channel Dimensions to Isolate Outliers for Low-bit Weight Quantization of Large Language Models
by: Heo, Jung Hwan, et al.
Published: (2023)
by: Heo, Jung Hwan, et al.
Published: (2023)
Whoever Started the Interference Should End It: Guiding Data-Free Model Merging via Task Vectors
by: Cheng, Runxi, et al.
Published: (2025)
by: Cheng, Runxi, et al.
Published: (2025)
Leech Lattice Vector Quantization for Efficient LLM Compression
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
FibQuant: Universal Vector Quantization for Random-Access KV-Cache Compression
by: Lee, Namyoon, et al.
Published: (2026)
by: Lee, Namyoon, et al.
Published: (2026)
Instance-Aware Test-Time Segmentation for Continual Domain Shifts
by: Lee, Seunghwan, et al.
Published: (2025)
by: Lee, Seunghwan, et al.
Published: (2025)
FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
by: Lee, Jung Hyun, et al.
Published: (2023)
by: Lee, Jung Hyun, et al.
Published: (2023)
AdaRank: Adaptive Rank Pruning for Enhanced Model Merging
by: Lee, Chanhyuk, et al.
Published: (2025)
by: Lee, Chanhyuk, et al.
Published: (2025)
CASHG: Context-Aware Stylized Online Handwriting Generation
by: Shin, Jinsu, et al.
Published: (2026)
by: Shin, Jinsu, et al.
Published: (2026)
Memory-Efficient Structured Backpropagation for On-Device LLM Fine-Tuning
by: Park, Juneyoung, et al.
Published: (2026)
by: Park, Juneyoung, et al.
Published: (2026)
Towards Minimizing Feature Drift in Model Merging: Layer-wise Task Vector Fusion for Adaptive Knowledge Integration
by: Sun, Wenju, et al.
Published: (2025)
by: Sun, Wenju, et al.
Published: (2025)
Estimating Subgraph Importance with Structural Prior Domain Knowledge
by: Kim, Changhyun, et al.
Published: (2026)
by: Kim, Changhyun, et al.
Published: (2026)
Memorize Early, Then Query: Inlier-Memorization-Guided Active Outlier Detection
by: Kang, Minseo, et al.
Published: (2026)
by: Kang, Minseo, et al.
Published: (2026)
Hi-SAFE: Hierarchical Secure Aggregation for Lightweight Federated Learning
by: Joo, Hyeong-Gun, et al.
Published: (2025)
by: Joo, Hyeong-Gun, et al.
Published: (2025)
AnyBCQ: Hardware Efficient Flexible Binary-Coded Quantization for Multi-Precision LLMs
by: Park, Gunho, et al.
Published: (2025)
by: Park, Gunho, et al.
Published: (2025)
Bilinear Coordinate Alignment for Training-Free Task-Vector Transfer
by: Son, Jungyong, et al.
Published: (2026)
by: Son, Jungyong, et al.
Published: (2026)
Online Vector Quantized Attention
by: Alonso, Nick, et al.
Published: (2026)
by: Alonso, Nick, et al.
Published: (2026)
Pyramid Vector Quantization for LLMs
by: van der Ouderaa, Tycho F. A., et al.
Published: (2024)
by: van der Ouderaa, Tycho F. A., et al.
Published: (2024)
AdaMerging: Adaptive Model Merging for Multi-Task Learning
by: Yang, Enneng, et al.
Published: (2023)
by: Yang, Enneng, et al.
Published: (2023)
ICaRus: Identical Cache Reuse for Efficient Multi Model Inference
by: Woo, Sunghyeon, et al.
Published: (2026)
by: Woo, Sunghyeon, et al.
Published: (2026)
Explaining How Quantization Disparately Skews a Model
by: Bellam, Abhimanyu, et al.
Published: (2025)
by: Bellam, Abhimanyu, et al.
Published: (2025)
Similar Items
-
SyMerge: From Non-Interference to Synergistic Merging via Single-Layer Adaptation
by: Jung, Aecheon, et al.
Published: (2024) -
Dynamic Rank Adjustment for Accurate and Efficient Neural Network Training
by: Shin, Hyuntak, et al.
Published: (2025) -
ZOO-Prune: Training-Free Token Pruning via Zeroth-Order Gradient Estimation in Vision-Language Models
by: Kim, Youngeun, et al.
Published: (2025) -
Backpropagation-Free Test-Time Adaptation via Probabilistic Gaussian Alignment
by: Zhang, Youjia, et al.
Published: (2025) -
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks
by: Jung, Aecheon, et al.
Published: (2025)