A Hormone-inspired Emotion Layer for Transformer language models (HELT)
Fuente:
arXiv
Saved in:
| Main Authors: | Reda, Eslam, El-Metwally, Sara |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Univariate Radial Basis Function Layers: Brain-inspired Deep Neural Layers for Low-Dimensional Inputs
by: Jost, Daniel, et al.
Published: (2023)
by: Jost, Daniel, et al.
Published: (2023)
Intelligent Neural Networks: From Layered Architectures to Graph-Organized Intelligence
by: Salomon, Antoine
Published: (2025)
by: Salomon, Antoine
Published: (2025)
On the Power of Convolution Augmented Transformer
by: Li, Mingchen, et al.
Published: (2024)
by: Li, Mingchen, et al.
Published: (2024)
BrainTransformers: SNN-LLM
by: Tang, Zhengzheng, et al.
Published: (2024)
by: Tang, Zhengzheng, et al.
Published: (2024)
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024)
by: Tang, Kaiwen, et al.
Published: (2024)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
by: Csordás, Róbert, et al.
Published: (2023)
by: Csordás, Róbert, et al.
Published: (2023)
HAT: Hardware-Aware Transformers for Efficient Natural Language Processing
by: Wang, Hanrui, et al.
Published: (2020)
by: Wang, Hanrui, et al.
Published: (2020)
Enriching language models with graph-based context information to better understand textual data
by: Roethel, Albert, et al.
Published: (2023)
by: Roethel, Albert, et al.
Published: (2023)
Comply: Learning Sentences with Complex Weights inspired by Fruit Fly Olfaction
by: Figueroa, Alexei, et al.
Published: (2025)
by: Figueroa, Alexei, et al.
Published: (2025)
Collapse-Free Prototype Readout Layer for Transformer Encoders
by: Cirrincione, Giansalvo, et al.
Published: (2026)
by: Cirrincione, Giansalvo, et al.
Published: (2026)
Understanding Textual Emotion Through Emoji Prediction
by: Gordon, Ethan, et al.
Published: (2025)
by: Gordon, Ethan, et al.
Published: (2025)
Why Prompt Optimization Works, and Why It Sometimes Doesn't: A Causal-Inspired Edit-Level Analysis
by: Gong, Shuzhi, et al.
Published: (2026)
by: Gong, Shuzhi, et al.
Published: (2026)
A Transformer-based Neural Architecture Search Method
by: Wang, Shang, et al.
Published: (2025)
by: Wang, Shang, et al.
Published: (2025)
Representation Learning in a Decomposed Encoder Design for Bio-inspired Hebbian Learning
by: Jaziri, Achref, et al.
Published: (2023)
by: Jaziri, Achref, et al.
Published: (2023)
S-TLLR: STDP-inspired Temporal Local Learning Rule for Spiking Neural Networks
by: Apolinario, Marco Paul E., et al.
Published: (2023)
by: Apolinario, Marco Paul E., et al.
Published: (2023)
GLU Attention Improve Transformer
by: Wang, Zehao
Published: (2025)
by: Wang, Zehao
Published: (2025)
Decomposing Evolutionary Mixture-of-LoRA Architectures: The Routing Lever, the Lifecycle Penalty, and a Substrate-Conditional Boundary
by: Kumaresan, Ramchand
Published: (2026)
by: Kumaresan, Ramchand
Published: (2026)
EvoX: Meta-Evolution for Automated Discovery
by: Liu, Shu, et al.
Published: (2026)
by: Liu, Shu, et al.
Published: (2026)
EvolKV: Evolutionary KV Cache Compression for LLM Inference
by: Yu, Bohan, et al.
Published: (2025)
by: Yu, Bohan, et al.
Published: (2025)
An In-depth Walkthrough on Evolution of Neural Machine Translation
by: Jagtap, Rohan, et al.
Published: (2020)
by: Jagtap, Rohan, et al.
Published: (2020)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
SpikeLM: Towards General Spike-Driven Language Modeling via Elastic Bi-Spiking Mechanisms
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
by: Briesch, Martin, et al.
Published: (2023)
by: Briesch, Martin, et al.
Published: (2023)
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
by: Majumdar, Somshubra, et al.
Published: (2024)
by: Majumdar, Somshubra, et al.
Published: (2024)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)
by: Kadlčík, Marek, et al.
Published: (2025)
Large Language Models for Tuning Evolution Strategies
by: Kramer, Oliver
Published: (2024)
by: Kramer, Oliver
Published: (2024)
AP-BMM: Approximating Capability-Cost Pareto Sets of LLMs via Asynchronous Prior-Guided Bayesian Model Merging
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models
by: Shen, Shuaijie, et al.
Published: (2024)
by: Shen, Shuaijie, et al.
Published: (2024)
An enhanced Teaching-Learning-Based Optimization (TLBO) with Grey Wolf Optimizer (GWO) for text feature selection and clustering
by: Azarshab, Mahsa, et al.
Published: (2024)
by: Azarshab, Mahsa, et al.
Published: (2024)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
ComplicaCode: Enhancing Disease Complication Detection in Electronic Health Records through ICD Path Generation
by: Zhou, Xiaofan
Published: (2023)
by: Zhou, Xiaofan
Published: (2023)
Improving Sequence-to-Sequence Models for Abstractive Text Summarization Using Meta Heuristic Approaches
by: Saxena, Aditya, et al.
Published: (2024)
by: Saxena, Aditya, et al.
Published: (2024)
Improving Language Plasticity via Pretraining with Active Forgetting
by: Chen, Yihong, et al.
Published: (2023)
by: Chen, Yihong, et al.
Published: (2023)
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
by: Zhu, Rui-Jie, et al.
Published: (2023)
by: Zhu, Rui-Jie, et al.
Published: (2023)
B'MOJO: Hybrid State Space Realizations of Foundation Models with Eidetic and Fading Memory
by: Zancato, Luca, et al.
Published: (2024)
by: Zancato, Luca, et al.
Published: (2024)
Neural Information Organizing and Processing -- Neural Machines
by: Petrila, Iosif Iulian
Published: (2024)
by: Petrila, Iosif Iulian
Published: (2024)
NOBLE: Accelerating Transformers with Nonlinear Low-Rank Branches
by: Smith, Ethan
Published: (2026)
by: Smith, Ethan
Published: (2026)
Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics
by: Ge, Zhangyu, et al.
Published: (2025)
by: Ge, Zhangyu, et al.
Published: (2025)
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
Similar Items
-
Univariate Radial Basis Function Layers: Brain-inspired Deep Neural Layers for Low-Dimensional Inputs
by: Jost, Daniel, et al.
Published: (2023) -
Intelligent Neural Networks: From Layered Architectures to Graph-Organized Intelligence
by: Salomon, Antoine
Published: (2025) -
On the Power of Convolution Augmented Transformer
by: Li, Mingchen, et al.
Published: (2024) -
BrainTransformers: SNN-LLM
by: Tang, Zhengzheng, et al.
Published: (2024) -
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024)