Quantization with Unified Adaptive Distillation to enable multi-LoRA based one-for-all Generative Vision Models on edge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vajrala, Sowmya, Parmar, Aakash, R, Prasanna, Kodavanti, Sravanth, Arveti, Manjunath, Miriyala, Srinivas Soumitri, Senapati, Ashok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
Towards Efficient Image Deblurring for Edge Deployment
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026)
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026)
TOC-SR: Task-Optimal Compact diffusion for Image Super Resolution
von: Vajrala, Sowmya, et al.
Veröffentlicht: (2026)
von: Vajrala, Sowmya, et al.
Veröffentlicht: (2026)
Edge-Efficient Image Restoration: Transformer Distillation into State-Space Models
von: Miriyala, Srinivas Soumitri, et al.
Veröffentlicht: (2026)
von: Miriyala, Srinivas Soumitri, et al.
Veröffentlicht: (2026)
Mobile-friendly Image de-noising: Hardware Conscious Optimization for Edge Application
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026)
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026)
Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
NanoSD: Edge Efficient Foundation Model for Real Time Image Restoration
von: Sanyal, Subhajit, et al.
Veröffentlicht: (2026)
von: Sanyal, Subhajit, et al.
Veröffentlicht: (2026)
QuAILoRA: Quantization-Aware Initialization for LoRA
von: Lawton, Neal, et al.
Veröffentlicht: (2024)
von: Lawton, Neal, et al.
Veröffentlicht: (2024)
Structure-Guided Histopathology Synthesis via Dual-LoRA Diffusion
von: Xu, Xuan, et al.
Veröffentlicht: (2026)
von: Xu, Xuan, et al.
Veröffentlicht: (2026)
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation
von: Azimi, Rambod, et al.
Veröffentlicht: (2024)
von: Azimi, Rambod, et al.
Veröffentlicht: (2024)
S-LoRA: Serving Thousands of Concurrent LoRA Adapters
von: Sheng, Ying, et al.
Veröffentlicht: (2023)
von: Sheng, Ying, et al.
Veröffentlicht: (2023)
LiON-LoRA: Rethinking LoRA Fusion to Unify Controllable Spatial and Temporal Generation for Video Diffusion
von: Zhang, Yisu, et al.
Veröffentlicht: (2025)
von: Zhang, Yisu, et al.
Veröffentlicht: (2025)
Run LoRA Run: Faster and Lighter LoRA Implementations
von: Cherniuk, Daria, et al.
Veröffentlicht: (2023)
von: Cherniuk, Daria, et al.
Veröffentlicht: (2023)
Dual LoRA: Enhancing LoRA with Magnitude and Direction Updates
von: Xu, Yixing, et al.
Veröffentlicht: (2025)
von: Xu, Yixing, et al.
Veröffentlicht: (2025)
Video2LoRA: Unified Semantic-Controlled Video Generation via Per-Reference-Video LoRA
von: Wu, Zexi, et al.
Veröffentlicht: (2026)
von: Wu, Zexi, et al.
Veröffentlicht: (2026)
LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits
von: Mirzaei, Amir Reza, et al.
Veröffentlicht: (2025)
von: Mirzaei, Amir Reza, et al.
Veröffentlicht: (2025)
Beyond LoRA: Exploring Efficient Fine-Tuning Techniques for Time Series Foundational Models
von: Gupta, Divij, et al.
Veröffentlicht: (2024)
von: Gupta, Divij, et al.
Veröffentlicht: (2024)
LoRA as Oracle
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
Vision as LoRA
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
SLAD : Shared LoRA Adapters for Task Specific Distillation
von: Bensaid, Reda, et al.
Veröffentlicht: (2026)
von: Bensaid, Reda, et al.
Veröffentlicht: (2026)
FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA
von: Zhang, Haoran, et al.
Veröffentlicht: (2026)
von: Zhang, Haoran, et al.
Veröffentlicht: (2026)
LoRA-Drop: Temporal LoRA Decoding for Efficient LLM Inference
von: Rajabzadeh, Hossein, et al.
Veröffentlicht: (2026)
von: Rajabzadeh, Hossein, et al.
Veröffentlicht: (2026)
TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts
von: Jeon, Jeimin, et al.
Veröffentlicht: (2026)
von: Jeon, Jeimin, et al.
Veröffentlicht: (2026)
LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging
von: Lee, Seungeon, et al.
Veröffentlicht: (2025)
von: Lee, Seungeon, et al.
Veröffentlicht: (2025)
Accurate LoRA-Finetuning Quantization of LLMs via Information Retention
von: Qin, Haotong, et al.
Veröffentlicht: (2024)
von: Qin, Haotong, et al.
Veröffentlicht: (2024)
LoRA Meets Dropout under a Unified Framework
von: Wang, Sheng, et al.
Veröffentlicht: (2024)
von: Wang, Sheng, et al.
Veröffentlicht: (2024)
ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads
von: Zuo, Jingwei, et al.
Veröffentlicht: (2026)
von: Zuo, Jingwei, et al.
Veröffentlicht: (2026)
LoRA-FAIR: Federated LoRA Fine-Tuning with Aggregation and Initialization Refinement
von: Bian, Jieming, et al.
Veröffentlicht: (2024)
von: Bian, Jieming, et al.
Veröffentlicht: (2024)
LoRA-drop: Efficient LoRA Parameter Pruning based on Output Evaluation
von: Zhou, Hongyun, et al.
Veröffentlicht: (2024)
von: Zhou, Hongyun, et al.
Veröffentlicht: (2024)
LoRA Done RITE: Robust Invariant Transformation Equilibration for LoRA Optimization
von: Yen, Jui-Nan, et al.
Veröffentlicht: (2024)
von: Yen, Jui-Nan, et al.
Veröffentlicht: (2024)
CE-LoRA: Computation-Efficient LoRA Fine-Tuning for Language Models
von: Chen, Guanduo, et al.
Veröffentlicht: (2025)
von: Chen, Guanduo, et al.
Veröffentlicht: (2025)
LoRA Diffusion: Zero-Shot LoRA Synthesis for Diffusion Model Personalization
von: Smith, Ethan, et al.
Veröffentlicht: (2024)
von: Smith, Ethan, et al.
Veröffentlicht: (2024)
SHE-LoRA: Selective Homomorphic Encryption for Federated Tuning with Heterogeneous LoRA
von: Liu, Jianmin, et al.
Veröffentlicht: (2025)
von: Liu, Jianmin, et al.
Veröffentlicht: (2025)
TC-LoRA: Temporally Modulated Conditional LoRA for Adaptive Diffusion Control
von: Cho, Minkyoung, et al.
Veröffentlicht: (2025)
von: Cho, Minkyoung, et al.
Veröffentlicht: (2025)
NP-LoRA: Null Space Projection for Subject-Style LoRA Fusion
von: Chen, Chuheng, et al.
Veröffentlicht: (2025)
von: Chen, Chuheng, et al.
Veröffentlicht: (2025)
ID-LoRA: Identity-Driven Audio-Video Personalization with In-Context LoRA
von: Dahan, Aviad, et al.
Veröffentlicht: (2026)
von: Dahan, Aviad, et al.
Veröffentlicht: (2026)
Convergence Analysis of Aggregation-Broadcast in LoRA-enabled Distributed Fine-Tuning
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Budgeted LoRA: Distillation as Structured Compute Allocation for Efficient Inference
von: Sabry, Mohammed, et al.
Veröffentlicht: (2026)
von: Sabry, Mohammed, et al.
Veröffentlicht: (2026)
A Note on LoRA
von: Fomenko, Vlad, et al.
Veröffentlicht: (2024)
von: Fomenko, Vlad, et al.
Veröffentlicht: (2024)
Mixture of LoRA Experts
von: Wu, Xun, et al.
Veröffentlicht: (2024)
von: Wu, Xun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026) -
Towards Efficient Image Deblurring for Edge Deployment
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026) -
TOC-SR: Task-Optimal Compact diffusion for Image Super Resolution
von: Vajrala, Sowmya, et al.
Veröffentlicht: (2026) -
Edge-Efficient Image Restoration: Transformer Distillation into State-Space Models
von: Miriyala, Srinivas Soumitri, et al.
Veröffentlicht: (2026) -
Mobile-friendly Image de-noising: Hardware Conscious Optimization for Edge Application
von: Miriyala, Srinivas, et al.
Veröffentlicht: (2026)