Tensor Train Low-rank Approximation (TT-LoRA): Democratizing AI with Accelerated LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Anjum, Afia, Eren, Maksim E., Boureima, Ismael, Alexandrov, Boian, Bhattarai, Manish |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs
von: Wanna, Selma, et al.
Veröffentlicht: (2024)
von: Wanna, Selma, et al.
Veröffentlicht: (2024)
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
LoRID: Low-Rank Iterative Diffusion for Adversarial Purification
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2024)
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2024)
Tensor-Train WENO Scheme for Compressible Flows
von: Danis, Mustafa Engin, et al.
Veröffentlicht: (2024)
von: Danis, Mustafa Engin, et al.
Veröffentlicht: (2024)
Cyber-Security Knowledge Graph Generation by Hierarchical Nonnegative Matrix Factorization
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
TT-LoRA MoE: Unifying Parameter-Efficient Fine-Tuning and Sparse Mixture-of-Experts
von: Kunwar, Pradip, et al.
Veröffentlicht: (2025)
von: Kunwar, Pradip, et al.
Veröffentlicht: (2025)
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization
von: Barron, Ryan C., et al.
Veröffentlicht: (2025)
von: Barron, Ryan C., et al.
Veröffentlicht: (2025)
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization
von: Barron, Ryan C., et al.
Veröffentlicht: (2024)
von: Barron, Ryan C., et al.
Veröffentlicht: (2024)
LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging
von: Lee, Seungeon, et al.
Veröffentlicht: (2025)
von: Lee, Seungeon, et al.
Veröffentlicht: (2025)
Topic Modeling and Link-Prediction for Material Property Discovery
von: Barron, Ryan C., et al.
Veröffentlicht: (2025)
von: Barron, Ryan C., et al.
Veröffentlicht: (2025)
ST-LoRA: Low-rank Adaptation for Spatio-Temporal Forecasting
von: Ruan, Weilin, et al.
Veröffentlicht: (2024)
von: Ruan, Weilin, et al.
Veröffentlicht: (2024)
Topological Signatures of Adversaries in Multimodal Alignments
von: Vu, Minh, et al.
Veröffentlicht: (2025)
von: Vu, Minh, et al.
Veröffentlicht: (2025)
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
Catch'em all: Classification of Rare, Prominent, and Novel Malware Families
von: Eren, Maksim E., et al.
Veröffentlicht: (2024)
von: Eren, Maksim E., et al.
Veröffentlicht: (2024)
ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads
von: Zuo, Jingwei, et al.
Veröffentlicht: (2026)
von: Zuo, Jingwei, et al.
Veröffentlicht: (2026)
S-LoRA: Serving Thousands of Concurrent LoRA Adapters
von: Sheng, Ying, et al.
Veröffentlicht: (2023)
von: Sheng, Ying, et al.
Veröffentlicht: (2023)
ARCS: Agentic Retrieval-Augmented Code Synthesis with Iterative Refinement
von: Bhattarai, Manish, et al.
Veröffentlicht: (2025)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2025)
LoRA as Oracle
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
Activated LoRA: Fine-tuned LLMs for Intrinsics
von: Greenewald, Kristjan, et al.
Veröffentlicht: (2025)
von: Greenewald, Kristjan, et al.
Veröffentlicht: (2025)
FALQON: Accelerating LoRA Fine-tuning with Low-Bit Floating-Point Arithmetic
von: Choi, Kanghyun, et al.
Veröffentlicht: (2025)
von: Choi, Kanghyun, et al.
Veröffentlicht: (2025)
LoRAFusion: Efficient LoRA Fine-Tuning for LLMs
von: Zhu, Zhanda, et al.
Veröffentlicht: (2025)
von: Zhu, Zhanda, et al.
Veröffentlicht: (2025)
Localized LoRA: A Structured Low-Rank Approximation for Efficient Fine-Tuning
von: Barazandeh, Babak, et al.
Veröffentlicht: (2025)
von: Barazandeh, Babak, et al.
Veröffentlicht: (2025)
HiLoRA: Adaptive Hierarchical LoRA Routing for Training-Free Domain Generalization
von: Han, Ziyi, et al.
Veröffentlicht: (2025)
von: Han, Ziyi, et al.
Veröffentlicht: (2025)
FinLoRA: Benchmarking LoRA Methods for Fine-Tuning LLMs on Financial Datasets
von: Wang, Dannong, et al.
Veröffentlicht: (2025)
von: Wang, Dannong, et al.
Veröffentlicht: (2025)
Matrix Factorization for Inferring Associations and Missing Links
von: Barron, Ryan, et al.
Veröffentlicht: (2025)
von: Barron, Ryan, et al.
Veröffentlicht: (2025)
LoRA Done RITE: Robust Invariant Transformation Equilibration for LoRA Optimization
von: Yen, Jui-Nan, et al.
Veröffentlicht: (2024)
von: Yen, Jui-Nan, et al.
Veröffentlicht: (2024)
Kron-LoRA: Hybrid Kronecker-LoRA Adapters for Scalable, Sustainable Fine-tuning
von: Shen, Yixin
Veröffentlicht: (2025)
von: Shen, Yixin
Veröffentlicht: (2025)
R-LoRA: Randomized Multi-Head LoRA for Efficient Multi-Task Learning
von: Liu, Jinda, et al.
Veröffentlicht: (2025)
von: Liu, Jinda, et al.
Veröffentlicht: (2025)
LoRA-Gen: Specializing Large Language Model via Online LoRA Generation
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
DR-LoRA: Dynamic Rank LoRA for Fine-Tuning Mixture-of-Experts Models
von: Deng, Guanzhi, et al.
Veröffentlicht: (2026)
von: Deng, Guanzhi, et al.
Veröffentlicht: (2026)
LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing
von: Li, Wenbing, et al.
Veröffentlicht: (2025)
von: Li, Wenbing, et al.
Veröffentlicht: (2025)
LoRA-Squeeze: Simple and Effective Post-Tuning and In-Tuning Compression of LoRA Modules
von: Vulić, Ivan, et al.
Veröffentlicht: (2026)
von: Vulić, Ivan, et al.
Veröffentlicht: (2026)
LaFA: Latent Feature Attacks on Non-negative Matrix Factorization
von: Vu, Minh, et al.
Veröffentlicht: (2024)
von: Vu, Minh, et al.
Veröffentlicht: (2024)
Tensor-Train Operator Inference
von: Danis, Engin, et al.
Veröffentlicht: (2025)
von: Danis, Engin, et al.
Veröffentlicht: (2025)
AC-LoRA: (Almost) Training-Free Access Control-Aware Multi-Modal LLMs
von: Lazier, Lara Magdalena, et al.
Veröffentlicht: (2025)
von: Lazier, Lara Magdalena, et al.
Veröffentlicht: (2025)
LoRA is All You Need for Safety Alignment of Reasoning LLMs
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
A Note on LoRA
von: Fomenko, Vlad, et al.
Veröffentlicht: (2024)
von: Fomenko, Vlad, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
von: Barron, Ryan, et al.
Veröffentlicht: (2024) -
TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs
von: Wanna, Selma, et al.
Veröffentlicht: (2024) -
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024) -
LoRID: Low-Rank Iterative Diffusion for Adversarial Purification
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2024) -
Tensor-Train WENO Scheme for Compressible Flows
von: Danis, Mustafa Engin, et al.
Veröffentlicht: (2024)