TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Mingxue, Xu, Yao Lei, Mandic, Danilo P. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TensorSLM: Energy-efficient Embedding Compression of Sub-billion Parameter Language Models on Low-end Devices
by: Xu, Mingxue, et al.
Published: (2025)
by: Xu, Mingxue, et al.
Published: (2025)
Geometry is All You Need: A Unified Taxonomy of Matrix and Tensor Factorization for Compression of Generative Language Models
by: Xu, Mingxue, et al.
Published: (2024)
by: Xu, Mingxue, et al.
Published: (2024)
Efficient Design of Compliant Mechanisms Using Multi-Objective Optimization
by: Humer, Alexander, et al.
Published: (2025)
by: Humer, Alexander, et al.
Published: (2025)
A Wachspress-based transfinite formulation for exactly enforcing Dirichlet boundary conditions on convex polygonal domains in physics-informed neural networks
by: Sukumar, N., et al.
Published: (2026)
by: Sukumar, N., et al.
Published: (2026)
TT-SNN: Tensor Train Decomposition for Efficient Spiking Neural Network Training
by: Lee, Donghyun, et al.
Published: (2024)
by: Lee, Donghyun, et al.
Published: (2024)
FastVPINNs: Tensor-Driven Acceleration of VPINNs for Complex Geometries
by: Anandh, Thivin, et al.
Published: (2024)
by: Anandh, Thivin, et al.
Published: (2024)
SO-PIFRNN: Self-optimization physics-informed Fourier-features randomized neural network for solving partial differential equations
by: Linghu, Jiale, et al.
Published: (2025)
by: Linghu, Jiale, et al.
Published: (2025)
From LIF to QIF: Toward Differentiable Spiking Neurons for Scientific Machine Learning
by: Wan, Ruyin, et al.
Published: (2025)
by: Wan, Ruyin, et al.
Published: (2025)
Expressivity and Approximation Properties of Deep Neural Networks with ReLU$^k$ Activation
by: He, Juncai, et al.
Published: (2023)
by: He, Juncai, et al.
Published: (2023)
HiPreNets: High-Precision Neural Networks through Progressive Training
by: Mulle, Ethan, et al.
Published: (2025)
by: Mulle, Ethan, et al.
Published: (2025)
Physics-Informed Neural Network based inverse framework for time-fractional differential equations for rheology
by: Thakur, Sukirt, et al.
Published: (2024)
by: Thakur, Sukirt, et al.
Published: (2024)
QuAKE: Speeding up Model Inference Using Quick and Approximate Kernels for Exponential Non-Linearities
by: Narayanaswami, Sai Kiran, et al.
Published: (2024)
by: Narayanaswami, Sai Kiran, et al.
Published: (2024)
Auto-weighted Bayesian Physics-Informed Neural Networks and robust estimations for multitask inverse problems in pore-scale imaging of dissolution
by: Perez, Sarah, et al.
Published: (2023)
by: Perez, Sarah, et al.
Published: (2023)
rKAN: Rational Kolmogorov-Arnold Networks
by: Aghaei, Alireza Afzal
Published: (2024)
by: Aghaei, Alireza Afzal
Published: (2024)
Semialgebraic Neural Networks: From roots to representations
by: Mis, S. David, et al.
Published: (2025)
by: Mis, S. David, et al.
Published: (2025)
Goat Optimization Algorithm: A Novel Bio-Inspired Metaheuristic for Global Optimization
by: Nozari, Hamed, et al.
Published: (2025)
by: Nozari, Hamed, et al.
Published: (2025)
A third-order finite difference weighted essentially non-oscillatory scheme with shallow neural network
by: Park, Kwanghyuk, et al.
Published: (2024)
by: Park, Kwanghyuk, et al.
Published: (2024)
A fast and accurate physics-informed neural network reduced order model with shallow masked autoencoder
by: Kim, Youngkyu, et al.
Published: (2020)
by: Kim, Youngkyu, et al.
Published: (2020)
Deep greedy unfolding: Sorting out argsorting in greedy sparse recovery algorithms
by: Mohammad-Taheri, Sina, et al.
Published: (2025)
by: Mohammad-Taheri, Sina, et al.
Published: (2025)
Robust Parameter and State Estimation in Multiscale Neuronal Systems Using Physics-Informed Neural Networks
by: Wei, Changliang, et al.
Published: (2026)
by: Wei, Changliang, et al.
Published: (2026)
Heuristic Solution to Joint Deployment and Beamforming Design for STAR-RIS Aided Networks
by: Yan, Bai, et al.
Published: (2024)
by: Yan, Bai, et al.
Published: (2024)
Structured and Balanced Multi-Component and Multi-Layer Neural Networks
by: Zhang, Shijun, et al.
Published: (2024)
by: Zhang, Shijun, et al.
Published: (2024)
Approximating Matrix Functions with Deep Neural Networks and Transformers
by: Padmanabhan, Rahul, et al.
Published: (2026)
by: Padmanabhan, Rahul, et al.
Published: (2026)
Universality and approximation bounds for echo state networks with random weights
by: Li, Zhen, et al.
Published: (2022)
by: Li, Zhen, et al.
Published: (2022)
Randomized Forward Mode Gradient for Spiking Neural Networks in Scientific Machine Learning
by: Wan, Ruyin, et al.
Published: (2024)
by: Wan, Ruyin, et al.
Published: (2024)
Linking Microscopic and Macroscopic Models for Evolution: Markov Chain Network Training and Conservation Law Approximations
by: Melnik, Roderick V. N.
Published: (2007)
by: Melnik, Roderick V. N.
Published: (2007)
Approximation Rates in Besov Norms and Sample-Complexity of Kolmogorov-Arnold Networks with Residual Connections
by: Kratsios, Anastasis, et al.
Published: (2025)
by: Kratsios, Anastasis, et al.
Published: (2025)
Neuro-Vesicles: Neuromodulation Should Be a Dynamical System, Not a Tensor Decoration
by: Li, Zilin, et al.
Published: (2025)
by: Li, Zilin, et al.
Published: (2025)
Splitting physics-informed neural networks for inferring the dynamics of integer- and fractional-order neuron models
by: Shekarpaz, Simin, et al.
Published: (2023)
by: Shekarpaz, Simin, et al.
Published: (2023)
Approximation Rates and VC-Dimension Bounds for (P)ReLU MLP Mixture of Experts
by: Kratsios, Anastasis, et al.
Published: (2024)
by: Kratsios, Anastasis, et al.
Published: (2024)
Is In-Context Universality Enough? MLPs are Also Universal In-Context
by: Kratsios, Anastasis, et al.
Published: (2025)
by: Kratsios, Anastasis, et al.
Published: (2025)
Stability and Discretization Error of State Space Model Neural Operators
by: Bendahi, Abderrahim, et al.
Published: (2026)
by: Bendahi, Abderrahim, et al.
Published: (2026)
A Practitioner's Guide to Kolmogorov-Arnold Networks
by: Noorizadegan, Amir, et al.
Published: (2025)
by: Noorizadegan, Amir, et al.
Published: (2025)
Tensorized Ant Colony Optimization for GPU Acceleration
by: Yang, Luming, et al.
Published: (2024)
by: Yang, Luming, et al.
Published: (2024)
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
by: Zhu, Rui-Jie, et al.
Published: (2023)
by: Zhu, Rui-Jie, et al.
Published: (2023)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
by: Briesch, Martin, et al.
Published: (2023)
by: Briesch, Martin, et al.
Published: (2023)
Hierarchical Frequency Tagging Probe (HFTP): A Unified Approach to Investigate Syntactic Structure Representations in Large Language Models and the Human Brain
by: An, Jingmin, et al.
Published: (2025)
by: An, Jingmin, et al.
Published: (2025)
Tensorized NeuroEvolution of Augmenting Topologies for GPU Acceleration
by: Wang, Lishuang, et al.
Published: (2024)
by: Wang, Lishuang, et al.
Published: (2024)
KAN versus MLP on Irregular or Noisy Functions
by: Zeng, Chen, et al.
Published: (2024)
by: Zeng, Chen, et al.
Published: (2024)
Activations Through Extensions: A Framework To Boost Performance Of Neural Networks
by: Kamanchi, Chandramouli, et al.
Published: (2024)
by: Kamanchi, Chandramouli, et al.
Published: (2024)
Similar Items
-
TensorSLM: Energy-efficient Embedding Compression of Sub-billion Parameter Language Models on Low-end Devices
by: Xu, Mingxue, et al.
Published: (2025) -
Geometry is All You Need: A Unified Taxonomy of Matrix and Tensor Factorization for Compression of Generative Language Models
by: Xu, Mingxue, et al.
Published: (2024) -
Efficient Design of Compliant Mechanisms Using Multi-Objective Optimization
by: Humer, Alexander, et al.
Published: (2025) -
A Wachspress-based transfinite formulation for exactly enforcing Dirichlet boundary conditions on convex polygonal domains in physics-informed neural networks
by: Sukumar, N., et al.
Published: (2026) -
TT-SNN: Tensor Train Decomposition for Efficient Spiking Neural Network Training
by: Lee, Donghyun, et al.
Published: (2024)