Model-Preserving Adaptive Rounding
Fuente:
arXiv
Saved in:
| Main Authors: | Tseng, Albert, Sun, Zhaofeng, De Sa, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
L$^3$: Large Lookup Layers
by: Tseng, Albert, et al.
Published: (2026)
by: Tseng, Albert, et al.
Published: (2026)
QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks
by: Tseng, Albert, et al.
Published: (2024)
by: Tseng, Albert, et al.
Published: (2024)
FAAR: Format-Aware Adaptive Rounding for NVFP4
by: Li, Hanglin, et al.
Published: (2026)
by: Li, Hanglin, et al.
Published: (2026)
FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
by: Lee, Jung Hyun, et al.
Published: (2023)
by: Lee, Jung Hyun, et al.
Published: (2023)
Rounding-Guided Backdoor Injection in Deep Learning Model Quantization
by: Chen, Xiangxiang, et al.
Published: (2025)
by: Chen, Xiangxiang, et al.
Published: (2025)
Deep Reinforcement Learning with Task-Adaptive Retrieval via Hypernetwork
by: Jin, Yonggang, et al.
Published: (2023)
by: Jin, Yonggang, et al.
Published: (2023)
From Bits to Rounds: Parallel Decoding with Exploration for Diffusion Language Models
by: Fu, Hengyu, et al.
Published: (2025)
by: Fu, Hengyu, et al.
Published: (2025)
Cluster-Aware Multi-Round Update for Wireless Federated Learning in Heterogeneous Environments
by: Sun, Pengcheng, et al.
Published: (2025)
by: Sun, Pengcheng, et al.
Published: (2025)
Preserving Plasticity in Continual Learning with Adaptive Linearity Injection
by: Rohani, Seyed Roozbeh Razavi, et al.
Published: (2025)
by: Rohani, Seyed Roozbeh Razavi, et al.
Published: (2025)
ModuLoRA: Finetuning 2-Bit LLMs on Consumer GPUs by Integrating with Modular Quantizers
by: Yin, Junjie, et al.
Published: (2023)
by: Yin, Junjie, et al.
Published: (2023)
Single-Round Scalable Analytic Federated Learning
by: Bacellar, Alan T. L., et al.
Published: (2025)
by: Bacellar, Alan T. L., et al.
Published: (2025)
Preserve-Then-Quantize: Balancing Rank Budgets for Quantization Error Reconstruction in LLMs
by: Cho, Yoonjun, et al.
Published: (2026)
by: Cho, Yoonjun, et al.
Published: (2026)
Preserving Diversity in Supervised Fine-Tuning of Large Language Models
by: Li, Ziniu, et al.
Published: (2024)
by: Li, Ziniu, et al.
Published: (2024)
QTIP: Quantization with Trellises and Incoherence Processing
by: Tseng, Albert, et al.
Published: (2024)
by: Tseng, Albert, et al.
Published: (2024)
CoreQ: Learning-Free Mismatch Correction and Successive Rounding for Quantization
by: Cha, Seohyeon, et al.
Published: (2026)
by: Cha, Seohyeon, et al.
Published: (2026)
Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming
by: He, Zhiqiang, et al.
Published: (2025)
by: He, Zhiqiang, et al.
Published: (2025)
Adaptive Clipping for Privacy-Preserving Few-Shot Learning: Enhancing Generalization with Limited Data
by: Ranaweera, Kanishka, et al.
Published: (2025)
by: Ranaweera, Kanishka, et al.
Published: (2025)
Adaptive Learning of Design Strategies over Non-Hierarchical Multi-Fidelity Models via Policy Alignment
by: Agrawal, Akash, et al.
Published: (2024)
by: Agrawal, Akash, et al.
Published: (2024)
Policy Gradient with Adaptive Entropy Annealing for Continual Fine-Tuning
by: Zhang, Yaqian, et al.
Published: (2026)
by: Zhang, Yaqian, et al.
Published: (2026)
PAC Privacy Preserving Diffusion Models
by: Xu, Qipan, et al.
Published: (2023)
by: Xu, Qipan, et al.
Published: (2023)
One Permutation Is All You Need: Fast, Reliable Variable Importance and Model Stress-Testing
by: Dorador, Albert
Published: (2025)
by: Dorador, Albert
Published: (2025)
Adaptive Non-local Observable on Quantum Neural Networks
by: Lin, Hsin-Yi, et al.
Published: (2025)
by: Lin, Hsin-Yi, et al.
Published: (2025)
Quantum Reinforcement Learning by Adaptive Non-local Observables
by: Lin, Hsin-Yi, et al.
Published: (2025)
by: Lin, Hsin-Yi, et al.
Published: (2025)
Quantum Super-resolution by Adaptive Non-local Observables
by: Lin, Hsin-Yi, et al.
Published: (2026)
by: Lin, Hsin-Yi, et al.
Published: (2026)
Compute-Optimal LLMs Provably Generalize Better With Scale
by: Finzi, Marc, et al.
Published: (2025)
by: Finzi, Marc, et al.
Published: (2025)
CoAst: Validation-Free Contribution Assessment for Federated Learning based on Cross-Round Valuation
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
SG-XDEAT: Sparsity-Guided Cross-Dimensional and Cross-Encoding Attention with Target-Aware Conditioning in Tabular Learning
by: Cheng, Chih-Chuan, et al.
Published: (2025)
by: Cheng, Chih-Chuan, et al.
Published: (2025)
RCCDA: Adaptive Model Updates in the Presence of Concept Drift under a Constrained Resource Budget
by: Piaseczny, Adam, et al.
Published: (2025)
by: Piaseczny, Adam, et al.
Published: (2025)
Diagonal Adaptive Non-local Observables on Quantum Neural Networks
by: Tseng, Huan-Hsin, et al.
Published: (2026)
by: Tseng, Huan-Hsin, et al.
Published: (2026)
HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts
by: Zhao, Hao, et al.
Published: (2024)
by: Zhao, Hao, et al.
Published: (2024)
RTMol: Rethinking Molecule-text Alignment in a Round-trip View
by: Chen, Letian, et al.
Published: (2025)
by: Chen, Letian, et al.
Published: (2025)
Towards Understanding Multi-Round Large Language Model Reasoning: Approximability, Learnability and Generalizability
by: Xu, Chenhui, et al.
Published: (2025)
by: Xu, Chenhui, et al.
Published: (2025)
Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUs
by: Zhao, Kang, et al.
Published: (2024)
by: Zhao, Kang, et al.
Published: (2024)
Adaptive Multi-Fidelity Reinforcement Learning for Variance Reduction in Engineering Design Optimization
by: Agrawal, Akash, et al.
Published: (2025)
by: Agrawal, Akash, et al.
Published: (2025)
Mamba: Linear-Time Sequence Modeling with Selective State Spaces
by: Gu, Albert, et al.
Published: (2023)
by: Gu, Albert, et al.
Published: (2023)
Knowledge-Aware Modeling with Frequency Adaptive Learning for Battery Health Prognostics
by: Pamshetti, Vijay Babu, et al.
Published: (2025)
by: Pamshetti, Vijay Babu, et al.
Published: (2025)
TAP: Two-Stage Adaptive Personalization of Multi-Task and Multi-Modal Foundation Models in Federated Learning
by: Lee, Seohyun, et al.
Published: (2025)
by: Lee, Seohyun, et al.
Published: (2025)
STAT: Shrinking Transformers After Training
by: Flynn, Megan, et al.
Published: (2024)
by: Flynn, Megan, et al.
Published: (2024)
SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models
by: Sun, Wenhao, et al.
Published: (2026)
by: Sun, Wenhao, et al.
Published: (2026)
Diffusion Adaptive Text Embedding for Text-to-Image Diffusion Models
by: Na, Byeonghu, et al.
Published: (2025)
by: Na, Byeonghu, et al.
Published: (2025)
Similar Items
-
L$^3$: Large Lookup Layers
by: Tseng, Albert, et al.
Published: (2026) -
QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks
by: Tseng, Albert, et al.
Published: (2024) -
FAAR: Format-Aware Adaptive Rounding for NVFP4
by: Li, Hanglin, et al.
Published: (2026) -
FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
by: Lee, Jung Hyun, et al.
Published: (2023) -
Rounding-Guided Backdoor Injection in Deep Learning Model Quantization
by: Chen, Xiangxiang, et al.
Published: (2025)