Neural Attention: A Novel Mechanism for Enhanced Expressive Power in Transformer Models
Fuente:
arXiv
Saved in:
| Main Authors: | DiGiugno, Andrew, Mahmood, Ausif |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spiking Neural Network Architecture Search: A Survey
by: Svoboda, Kama, et al.
Published: (2025)
by: Svoboda, Kama, et al.
Published: (2025)
Interactive LLM-assisted Curriculum Learning for Multi-Task Evolutionary Policy Search
by: Sakallioglu, Berfin, et al.
Published: (2026)
by: Sakallioglu, Berfin, et al.
Published: (2026)
Enhanced Protein Intrinsic Disorder Prediction Through Dual-View Multiscale Features and Multi-objective Evolutionary Algorithm
by: Wang, Shaokuan, et al.
Published: (2026)
by: Wang, Shaokuan, et al.
Published: (2026)
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
by: Bömer, Thomas, et al.
Published: (2026)
by: Bömer, Thomas, et al.
Published: (2026)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
by: Ji, Binbin, et al.
Published: (2025)
by: Ji, Binbin, et al.
Published: (2025)
Theoretical Analysis of the Advantage of Deepening Neural Networks
by: Esaki, Yasushi, et al.
Published: (2020)
by: Esaki, Yasushi, et al.
Published: (2020)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
by: Sun, Yuhui, et al.
Published: (2025)
by: Sun, Yuhui, et al.
Published: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
by: Reddy, Sandeep, et al.
Published: (2025)
by: Reddy, Sandeep, et al.
Published: (2025)
Vector Symbolic Architectures answer Jackendoff's challenges for cognitive neuroscience
by: Gayler, Ross W.
Published: (2004)
by: Gayler, Ross W.
Published: (2004)
METER: Multi-modal Evidence-based Thinking and Explainable Reasoning -- Algorithm and Benchmark
by: Yang, Xu, et al.
Published: (2025)
by: Yang, Xu, et al.
Published: (2025)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
by: Zolduoarrati, Elijah, et al.
Published: (2025)
by: Zolduoarrati, Elijah, et al.
Published: (2025)
WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents
by: Liu, Bingnan, et al.
Published: (2026)
by: Liu, Bingnan, et al.
Published: (2026)
Evolving Programmatic Skill Networks
by: Shi, Haochen, et al.
Published: (2026)
by: Shi, Haochen, et al.
Published: (2026)
Survey Transfer Learning: Recycling Data with Silicon Responses
by: Amini, Ali
Published: (2025)
by: Amini, Ali
Published: (2025)
Reward-Modulated Local Learning in Spiking Encoders: Controlled Benchmarks with STDP and Hybrid Rate Readouts
by: Chakraborty, Debjyoti
Published: (2026)
by: Chakraborty, Debjyoti
Published: (2026)
A Survey on Large Language Models with some Insights on their Capabilities and Limitations
by: Matarazzo, Andrea, et al.
Published: (2025)
by: Matarazzo, Andrea, et al.
Published: (2025)
Expanding continual few-shot learning benchmarks to include recognition of specific instances
by: Kowadlo, Gideon, et al.
Published: (2022)
by: Kowadlo, Gideon, et al.
Published: (2022)
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
by: Atad, Ido Andrew, et al.
Published: (2026)
by: Atad, Ido Andrew, et al.
Published: (2026)
Collapse or Preserve: Data-Dependent Temporal Aggregation for Spiking Neural Network Acceleration
by: Qin, Jiahao
Published: (2026)
by: Qin, Jiahao
Published: (2026)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
by: Mandal, Paul K.
Published: (2025)
by: Mandal, Paul K.
Published: (2025)
Less is More: Learning Graph Tasks with Just LLMs
by: Shirai, Sola, et al.
Published: (2025)
by: Shirai, Sola, et al.
Published: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
by: Schneider, Felix, et al.
Published: (2026)
by: Schneider, Felix, et al.
Published: (2026)
Deep Convolutional Autoencoder for Assessment of Drive-Cycle Anomalies in Connected Vehicle Sensor Data
by: Geglio, Anthony, et al.
Published: (2022)
by: Geglio, Anthony, et al.
Published: (2022)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
by: Wu, Jason, et al.
Published: (2026)
by: Wu, Jason, et al.
Published: (2026)
Using Deep Learning to Generate Semantically Correct Hindi Captions
by: Khan, Wasim Akram, et al.
Published: (2026)
by: Khan, Wasim Akram, et al.
Published: (2026)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
Phase-Coded Memory and Morphological Resonance: A Next-Generation Retrieval-Augmented Generator Architecture
by: Saklakov, Denis V.
Published: (2025)
by: Saklakov, Denis V.
Published: (2025)
VLM-VPI: A Vision-Language Reasoning Framework for Improving Automated Vehicle-Pedestrian Interactions
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Neuromorphic Parameter Estimation for Power Converter Health Monitoring Using Spiking Neural Networks
by: Baik, Hyeongmeen, et al.
Published: (2026)
by: Baik, Hyeongmeen, et al.
Published: (2026)
REMoH: A Reflective Evolution of Multi-objective Heuristics approach via Large Language Models
by: Forniés-Tabuenca, Diego, et al.
Published: (2025)
by: Forniés-Tabuenca, Diego, et al.
Published: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
by: Fadli, Samih
Published: (2025)
by: Fadli, Samih
Published: (2025)
SCULPT: Constraint-Guided Pruned MCTS that Carves Efficient Paths for Mathematical Reasoning
by: Fang, Qitong, et al.
Published: (2026)
by: Fang, Qitong, et al.
Published: (2026)
On the Limits of Learned Importance Scoring for KV Cache Compression
by: Steele, Brady
Published: (2026)
by: Steele, Brady
Published: (2026)
Scaling Trends for Multi-Hop Contextual Reasoning in Mid-Scale Language Models
by: Steele, Brady, et al.
Published: (2026)
by: Steele, Brady, et al.
Published: (2026)
The Reasoning-Creativity Trade-off: Toward Creativity-Driven Problem Solving
by: Luyten, Max Ruiz, et al.
Published: (2026)
by: Luyten, Max Ruiz, et al.
Published: (2026)
Similar Items
-
Spiking Neural Network Architecture Search: A Survey
by: Svoboda, Kama, et al.
Published: (2025) -
Interactive LLM-assisted Curriculum Learning for Multi-Task Evolutionary Policy Search
by: Sakallioglu, Berfin, et al.
Published: (2026) -
Enhanced Protein Intrinsic Disorder Prediction Through Dual-View Multiscale Features and Multi-objective Evolutionary Algorithm
by: Wang, Shaokuan, et al.
Published: (2026) -
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
by: Bömer, Thomas, et al.
Published: (2026) -
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
by: Ji, Binbin, et al.
Published: (2025)