Efficient transformer with reinforced position embedding for language models
Fuente:
arXiv
Saved in:
| Main Authors: | Hsiao, Yen-Che, Dutta, Abhishek |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Autonomous Agents: Adaptive-planning, Reasoning, and Acting in Language Models
by: Dutta, Abhishek, et al.
Published: (2024)
by: Dutta, Abhishek, et al.
Published: (2024)
Unveiling Reasoning Thresholds in Language Models: Scaling, Fine-Tuning, and Interpretability through Attention Maps
by: Hsiao, Yen-Che, et al.
Published: (2025)
by: Hsiao, Yen-Che, et al.
Published: (2025)
Hybrid Coordinate Descent for Efficient Neural Network Learning Using Line Search and Gradient Descent
by: Hsiao, Yen-Che, et al.
Published: (2024)
by: Hsiao, Yen-Che, et al.
Published: (2024)
Adaptive Reasoning and Acting in Medical Language Agents
by: Dutta, Abhishek, et al.
Published: (2024)
by: Dutta, Abhishek, et al.
Published: (2024)
Enhanced Arabic-language cyberbullying detection: deep embedding and transformer (BERT) approaches
by: Aljohani, Ebtesam Jaber, et al.
Published: (2025)
by: Aljohani, Ebtesam Jaber, et al.
Published: (2025)
Why transformers are obviously good models of language
by: Hill, Felix
Published: (2024)
by: Hill, Felix
Published: (2024)
Interpreting the linear structure of vision-language model embedding spaces
by: Papadimitriou, Isabel, et al.
Published: (2025)
by: Papadimitriou, Isabel, et al.
Published: (2025)
Prompt reinforcing for long-term planning of large language models
by: Lin, Hsien-Chin, et al.
Published: (2025)
by: Lin, Hsien-Chin, et al.
Published: (2025)
Vocabulary embeddings organize linguistic structure early in language model training
by: Papadimitriou, Isabel, et al.
Published: (2025)
by: Papadimitriou, Isabel, et al.
Published: (2025)
A semantic embedding space based on large language models for modelling human beliefs
by: Lee, Byunghwee, et al.
Published: (2024)
by: Lee, Byunghwee, et al.
Published: (2024)
NLD-LLM: A systematic framework for evaluating small language transformer models on natural language description
by: Jelodar, Hamed, et al.
Published: (2025)
by: Jelodar, Hamed, et al.
Published: (2025)
Derivation of Back-propagation for Graph Convolutional Networks using Matrix Calculus and its Application to Explainable Artificial Intelligence
by: Hsiao, Yen-Che, et al.
Published: (2024)
by: Hsiao, Yen-Che, et al.
Published: (2024)
Lost without translation -- Can transformer (language models) understand mood states?
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
Mixture-of-Depths: Dynamically allocating compute in transformer-based language models
by: Raposo, David, et al.
Published: (2024)
by: Raposo, David, et al.
Published: (2024)
Survey on reinforcement learning for language processing
by: Uc-Cetina, Victor, et al.
Published: (2021)
by: Uc-Cetina, Victor, et al.
Published: (2021)
Humans and transformer LMs: Abstraction drives language learning
by: Jian, Jasper, et al.
Published: (2026)
by: Jian, Jasper, et al.
Published: (2026)
Physical models realizing the transformer architecture of large language models
by: Chen, Zeqian
Published: (2025)
by: Chen, Zeqian
Published: (2025)
Comparison of different Unique hard attention transformer models by the formal languages they can recognize
by: Ryvkin, Leonid
Published: (2025)
by: Ryvkin, Leonid
Published: (2025)
A comparative study of transformer-based embeddings for topic coherence
by: Ding, Alex, et al.
Published: (2026)
by: Ding, Alex, et al.
Published: (2026)
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
by: Thamma, Abishek, et al.
Published: (2025)
by: Thamma, Abishek, et al.
Published: (2025)
Spatio-temporal transformer to support automatic sign language translation
by: Ruiz, Christian, et al.
Published: (2025)
by: Ruiz, Christian, et al.
Published: (2025)
Multilingual acoustic word embeddings for zero-resource languages
by: Jacobs, Christiaan
Published: (2024)
by: Jacobs, Christiaan
Published: (2024)
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
by: Kumar, Aditya, et al.
Published: (2025)
by: Kumar, Aditya, et al.
Published: (2025)
A path to natural language through tokenisation and transformers
by: Berman, David S., et al.
Published: (2026)
by: Berman, David S., et al.
Published: (2026)
Efficient argument classification with compact language models and ChatGPT-4 refinements
by: Pietron, Marcin, et al.
Published: (2024)
by: Pietron, Marcin, et al.
Published: (2024)
'Neural howlround' in large language models: a self-reinforcing bias phenomenon, and a dynamic attenuation solution
by: Drake, Seth
Published: (2025)
by: Drake, Seth
Published: (2025)
Attention-based transformer models for image captioning across languages: An in-depth survey and evaluation
by: Albadarneh, Israa A., et al.
Published: (2025)
by: Albadarneh, Israa A., et al.
Published: (2025)
Can large language models understand uncommon meanings of common words?
by: Wu, Jinyang, et al.
Published: (2024)
by: Wu, Jinyang, et al.
Published: (2024)
A multitask transformer to sign language translation using motion gesture primitives
by: López, Fredy Alejandro Mendoza, et al.
Published: (2025)
by: López, Fredy Alejandro Mendoza, et al.
Published: (2025)
Large language models have learned to use language
by: Lupyan, Gary
Published: (2025)
by: Lupyan, Gary
Published: (2025)
Tiny language models
by: Gross, Ronit D., et al.
Published: (2025)
by: Gross, Ronit D., et al.
Published: (2025)
Detecting out-of-distribution text using topological features of transformer-based language models
by: Pollano, Andres, et al.
Published: (2023)
by: Pollano, Andres, et al.
Published: (2023)
Can we teach language models to gloss endangered languages?
by: Ginn, Michael, et al.
Published: (2024)
by: Ginn, Michael, et al.
Published: (2024)
Retrieval augmentation of large language models for lay language generation
by: Guo, Yue, et al.
Published: (2022)
by: Guo, Yue, et al.
Published: (2022)
Do large language models resemble humans in language use?
by: Cai, Zhenguang G., et al.
Published: (2023)
by: Cai, Zhenguang G., et al.
Published: (2023)
Studies with impossible languages falsify LMs as models of human language
by: Bowers, Jeffrey S., et al.
Published: (2025)
by: Bowers, Jeffrey S., et al.
Published: (2025)
Generalist embedding models are better at short-context clinical semantic search than specialized embedding models
by: Excoffier, Jean-Baptiste, et al.
Published: (2024)
by: Excoffier, Jean-Baptiste, et al.
Published: (2024)
Why do language models perform worse for morphologically complex languages?
by: Arnett, Catherine, et al.
Published: (2024)
by: Arnett, Catherine, et al.
Published: (2024)
Prompting language influences diagnostic reasoning and accuracy of large language models
by: Bazoge, Adrien, et al.
Published: (2026)
by: Bazoge, Adrien, et al.
Published: (2026)
A systematic review of geospatial location embedding approaches in large language models: A path to spatial AI systems
by: Tucker, Sean
Published: (2024)
by: Tucker, Sean
Published: (2024)
Similar Items
-
Towards Autonomous Agents: Adaptive-planning, Reasoning, and Acting in Language Models
by: Dutta, Abhishek, et al.
Published: (2024) -
Unveiling Reasoning Thresholds in Language Models: Scaling, Fine-Tuning, and Interpretability through Attention Maps
by: Hsiao, Yen-Che, et al.
Published: (2025) -
Hybrid Coordinate Descent for Efficient Neural Network Learning Using Line Search and Gradient Descent
by: Hsiao, Yen-Che, et al.
Published: (2024) -
Adaptive Reasoning and Acting in Medical Language Agents
by: Dutta, Abhishek, et al.
Published: (2024) -
Enhanced Arabic-language cyberbullying detection: deep embedding and transformer (BERT) approaches
by: Aljohani, Ebtesam Jaber, et al.
Published: (2025)