HAT: Hardware-Aware Transformers for Efficient Natural Language Processing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hanrui, Wu, Zhanghao, Liu, Zhijian, Cai, Han, Zhu, Ligeng, Gan, Chuang, Han, Song |
|---|---|
| Format: | Preprint |
| Published: |
2020
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024)
by: Tang, Kaiwen, et al.
Published: (2024)
Hierarchical Frequency Tagging Probe (HFTP): A Unified Approach to Investigate Syntactic Structure Representations in Large Language Models and the Human Brain
by: An, Jingmin, et al.
Published: (2025)
by: An, Jingmin, et al.
Published: (2025)
Projection-Free Evolution Strategies for Continuous Prompt Search
by: Cai, Yu, et al.
Published: (2026)
by: Cai, Yu, et al.
Published: (2026)
BrainTransformers: SNN-LLM
by: Tang, Zhengzheng, et al.
Published: (2024)
by: Tang, Zhengzheng, et al.
Published: (2024)
Simulated Language Acquisition in a Biologically Realistic Model of the Brain
by: Mitropolsky, Daniel, et al.
Published: (2025)
by: Mitropolsky, Daniel, et al.
Published: (2025)
Personalized Language Models via Privacy-Preserving Evolutionary Model Merging
by: Kim, Kyuyoung, et al.
Published: (2025)
by: Kim, Kyuyoung, et al.
Published: (2025)
PaceLLM: Brain-Inspired Large Language Models for Long-Context Understanding
by: Li, Kangcong, et al.
Published: (2025)
by: Li, Kangcong, et al.
Published: (2025)
Evolutionary Computation in the Era of Large Language Model: Survey and Roadmap
by: Wu, Xingyu, et al.
Published: (2024)
by: Wu, Xingyu, et al.
Published: (2024)
Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs
by: Li, Xiaoxia, et al.
Published: (2024)
by: Li, Xiaoxia, et al.
Published: (2024)
A Landscape-Aware Differential Evolution for Multimodal Optimization Problems
by: Lin, Guo-Yun, et al.
Published: (2024)
by: Lin, Guo-Yun, et al.
Published: (2024)
RNC: Efficient RRAM-aware NAS and Compilation for DNNs on Resource-Constrained Edge Devices
by: Loong, Kam Chi, et al.
Published: (2024)
by: Loong, Kam Chi, et al.
Published: (2024)
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
by: Zhu, Rui-Jie, et al.
Published: (2023)
by: Zhu, Rui-Jie, et al.
Published: (2023)
On the Power of Convolution Augmented Transformer
by: Li, Mingchen, et al.
Published: (2024)
by: Li, Mingchen, et al.
Published: (2024)
Interpreting and learning voice commands with a Large Language Model for a robot system
by: Stankevich, Stanislau, et al.
Published: (2024)
by: Stankevich, Stanislau, et al.
Published: (2024)
Toward Relative Positional Encoding in Spiking Transformers
by: Lv, Changze, et al.
Published: (2025)
by: Lv, Changze, et al.
Published: (2025)
Neural Information Organizing and Processing -- Neural Machines
by: Petrila, Iosif Iulian
Published: (2024)
by: Petrila, Iosif Iulian
Published: (2024)
What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search
by: Zhang, Xinhao, et al.
Published: (2026)
by: Zhang, Xinhao, et al.
Published: (2026)
A Comprehensive Empirical Evaluation of Existing Word Embedding Approaches
by: Zaland, Obaidullah, et al.
Published: (2023)
by: Zaland, Obaidullah, et al.
Published: (2023)
Generalization and Completeness of Stochastic Local Search Algorithms
by: Loscos, Daniel, et al.
Published: (2026)
by: Loscos, Daniel, et al.
Published: (2026)
Spiking Convolutional Neural Networks for Text Classification
by: Lv, Changze, et al.
Published: (2024)
by: Lv, Changze, et al.
Published: (2024)
A survey on learning models of spiking neural membrane systems and spiking neural networks
by: Paul, Prithwineel, et al.
Published: (2024)
by: Paul, Prithwineel, et al.
Published: (2024)
Children's Acquisition of Tail-recursion Sequences: A Review of Locative Recursion and Possessive Recursion as Examples
by: Wang, Xiaoyi, et al.
Published: (2024)
by: Wang, Xiaoyi, et al.
Published: (2024)
Towards Faster k-Nearest-Neighbor Machine Translation
by: Shi, Xiangyu, et al.
Published: (2023)
by: Shi, Xiangyu, et al.
Published: (2023)
Bio-Inspired Mamba: Temporal Locality and Bioplausible Learning in Selective State Space Models
by: Qin, Jiahao
Published: (2024)
by: Qin, Jiahao
Published: (2024)
VietNormalizer: An Open-Source, Dependency-Free Python Library for Vietnamese Text Normalization in TTS and NLP Applications
by: Nguyen, Hung Vu, et al.
Published: (2026)
by: Nguyen, Hung Vu, et al.
Published: (2026)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
by: Csordás, Róbert, et al.
Published: (2023)
by: Csordás, Róbert, et al.
Published: (2023)
TEFormer: Structured Bidirectional Temporal Enhancement Modeling in Spiking Transformers
by: Shen, Sicheng, et al.
Published: (2026)
by: Shen, Sicheng, et al.
Published: (2026)
Scalable Network Emulation on Analog Neuromorphic Hardware
by: Arnold, Elias, et al.
Published: (2024)
by: Arnold, Elias, et al.
Published: (2024)
MAR: Efficient Large Language Models via Module-aware Architecture Refinement
by: Cai, Junhong, et al.
Published: (2026)
by: Cai, Junhong, et al.
Published: (2026)
BEExformer: A Fast Inferencing Binarized Transformer with Early Exits
by: Ansar, Wazib, et al.
Published: (2024)
by: Ansar, Wazib, et al.
Published: (2024)
Efficient Heuristics Generation for Solving Combinatorial Optimization Problems Using Large Language Models
by: Wu, Xuan, et al.
Published: (2025)
by: Wu, Xuan, et al.
Published: (2025)
Efficient and Effective Time-Series Forecasting with Spiking Neural Networks
by: Lv, Changze, et al.
Published: (2024)
by: Lv, Changze, et al.
Published: (2024)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
A Hormone-inspired Emotion Layer for Transformer language models (HELT)
by: Reda, Eslam, et al.
Published: (2026)
by: Reda, Eslam, et al.
Published: (2026)
Bridging Quantized Artificial Neural Networks and Neuromorphic Hardware
by: Chen, Zhenhui, et al.
Published: (2025)
by: Chen, Zhenhui, et al.
Published: (2025)
Demonstrating the Advantages of Analog Wafer-Scale Neuromorphic Hardware
by: Schmidt, Hartmut, et al.
Published: (2024)
by: Schmidt, Hartmut, et al.
Published: (2024)
Amortized Inference of Neuron Parameters on Analog Neuromorphic Hardware
by: Kaiser, Jakob, et al.
Published: (2026)
by: Kaiser, Jakob, et al.
Published: (2026)
Neuronal Group Communication for Efficient Neural representation
by: Pei, Zhengqi, et al.
Published: (2025)
by: Pei, Zhengqi, et al.
Published: (2025)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
Design and Implementation of Hardware Accelerators for Neural Processing Applications
by: Mayannavar, Shilpa, et al.
Published: (2024)
by: Mayannavar, Shilpa, et al.
Published: (2024)
Similar Items
-
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024) -
Hierarchical Frequency Tagging Probe (HFTP): A Unified Approach to Investigate Syntactic Structure Representations in Large Language Models and the Human Brain
by: An, Jingmin, et al.
Published: (2025) -
Projection-Free Evolution Strategies for Continuous Prompt Search
by: Cai, Yu, et al.
Published: (2026) -
BrainTransformers: SNN-LLM
by: Tang, Zhengzheng, et al.
Published: (2024) -
Simulated Language Acquisition in a Biologically Realistic Model of the Brain
by: Mitropolsky, Daniel, et al.
Published: (2025)