Enhancing Computation Efficiency in Large Language Models through Weight and Activation Quantization

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lee, Janghwan, Kim, Minsoo, Baek, Seungcheol, Hwang, Seok Joong, Sung, Wonyong, Choi, Jungwook
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!

Similar Items