Lightweight Transformer Architectures for Edge Devices in Real-Time Applications
Fuente:
arXiv
Guardado en:
| Autor principal: | Samson, Hema Hariharan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Aligning LLMs for Multilingual Consistency in Enterprise Applications
por: Agarwal, Amit, et al.
Publicado: (2025)
por: Agarwal, Amit, et al.
Publicado: (2025)
A Comparative Study of Feature Selection in Tsetlin Machines
por: Halenka, Vojtech, et al.
Publicado: (2025)
por: Halenka, Vojtech, et al.
Publicado: (2025)
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
por: Aman, La Ode, et al.
Publicado: (2024)
por: Aman, La Ode, et al.
Publicado: (2024)
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities
por: Giacomelli, Stefano, et al.
Publicado: (2024)
por: Giacomelli, Stefano, et al.
Publicado: (2024)
Graded Transformers
por: Shaska Sr, Tony
Publicado: (2025)
por: Shaska Sr, Tony
Publicado: (2025)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
por: Mitchell, Rupert, et al.
Publicado: (2025)
por: Mitchell, Rupert, et al.
Publicado: (2025)
Benchmarking Catastrophic Forgetting Mitigation Methods in Federated Time Series Forecasting
por: Hallak, Khaled, et al.
Publicado: (2025)
por: Hallak, Khaled, et al.
Publicado: (2025)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
por: Radosky, Lukas, et al.
Publicado: (2026)
por: Radosky, Lukas, et al.
Publicado: (2026)
Enhancing Diversity in Multi-objective Feature Selection
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
por: Danieli, Federico, et al.
Publicado: (2025)
por: Danieli, Federico, et al.
Publicado: (2025)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
por: Pather, Kaviraj, et al.
Publicado: (2025)
por: Pather, Kaviraj, et al.
Publicado: (2025)
Complex-Valued Phase-Coherent Transformer
por: Hioki, Leona
Publicado: (2026)
por: Hioki, Leona
Publicado: (2026)
MVTamperBench: Evaluating Robustness of Vision-Language Models
por: Agarwal, Amit, et al.
Publicado: (2024)
por: Agarwal, Amit, et al.
Publicado: (2024)
ReliabilityBench: Evaluating LLM Agent Reliability Under Production-Like Stress Conditions
por: Gupta, Aayush
Publicado: (2026)
por: Gupta, Aayush
Publicado: (2026)
VideoMind: An Omni-Modal Video Dataset with Intent Grounding for Deep-Cognitive Video Understanding
por: Yang, Baoyao, et al.
Publicado: (2025)
por: Yang, Baoyao, et al.
Publicado: (2025)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
por: Du, Wenzhang
Publicado: (2025)
por: Du, Wenzhang
Publicado: (2025)
Measuring Intent Comprehension in LLMs
por: Kunievsky, Nadav, et al.
Publicado: (2025)
por: Kunievsky, Nadav, et al.
Publicado: (2025)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
por: Sáez, Arnau Igualde, et al.
Publicado: (2025)
por: Sáez, Arnau Igualde, et al.
Publicado: (2025)
XAutoLM: Efficient Fine-Tuning of Language Models via Meta-Learning and AutoML
por: Estevanell-Valladares, Ernesto L., et al.
Publicado: (2025)
por: Estevanell-Valladares, Ernesto L., et al.
Publicado: (2025)
Optimized Gradient Clipping for Noisy Label Learning
por: Ye, Xichen, et al.
Publicado: (2024)
por: Ye, Xichen, et al.
Publicado: (2024)
Boosting Test Performance with Importance Sampling--a Subpopulation Perspective
por: Shen, Hongyu, et al.
Publicado: (2024)
por: Shen, Hongyu, et al.
Publicado: (2024)
A Systematic Review of Digital Twin-Driven Predictive Maintenance in Industrial Engineering: Taxonomy, Architectural Elements, and Future Research Directions
por: Ismail, Leila, et al.
Publicado: (2025)
por: Ismail, Leila, et al.
Publicado: (2025)
A Computational Approach to Modeling Conversational Systems: Analyzing Large-Scale Quasi-Patterned Dialogue Flows
por: Ammar, Mohamed Achref Ben, et al.
Publicado: (2025)
por: Ammar, Mohamed Achref Ben, et al.
Publicado: (2025)
Transparent but Powerful: Explainability, Accuracy, and Generalizability in ADHD Detection from Social Media Data
por: Wiechmann, D., et al.
Publicado: (2024)
por: Wiechmann, D., et al.
Publicado: (2024)
FedDAF: Federated Domain Adaptation Using Model Functional Distance
por: Sen, Mrinmay, et al.
Publicado: (2025)
por: Sen, Mrinmay, et al.
Publicado: (2025)
SLEAN: Simple Lightweight Ensemble Analysis Network for Multi-Provider LLM Coordination: Design, Implementation, and Vibe Coding Bug Investigation Case Study
por: Vargas, Matheus J. T.
Publicado: (2025)
por: Vargas, Matheus J. T.
Publicado: (2025)
EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
por: Nauen, Tobias Christian, et al.
Publicado: (2024)
por: Nauen, Tobias Christian, et al.
Publicado: (2024)
Can Agentic AI Match the Performance of Human Data Scientists?
por: Luo, An, et al.
Publicado: (2025)
por: Luo, An, et al.
Publicado: (2025)
AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science
por: Luo, An, et al.
Publicado: (2026)
por: Luo, An, et al.
Publicado: (2026)
Generative AI and the Transformation of Software Development Practices
por: Acharya, Vivek
Publicado: (2025)
por: Acharya, Vivek
Publicado: (2025)
Graph Connectionist Temporal Classification for Phoneme Recognition
por: Grafé, Henry, et al.
Publicado: (2025)
por: Grafé, Henry, et al.
Publicado: (2025)
Symphonym: Universal Phonetic Embeddings for Cross-Script Name Matching
por: Gadd, Stephen
Publicado: (2026)
por: Gadd, Stephen
Publicado: (2026)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
por: Bhadra, Dipayan, et al.
Publicado: (2025)
por: Bhadra, Dipayan, et al.
Publicado: (2025)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
por: Qesaraku, Bjorna, et al.
Publicado: (2025)
por: Qesaraku, Bjorna, et al.
Publicado: (2025)
Semantic Modeling for World-Centered Architectures
por: Mantsivoda, Andrei, et al.
Publicado: (2026)
por: Mantsivoda, Andrei, et al.
Publicado: (2026)
Introducing COGENT3: An AI Architecture for Emergent Cognition
por: Salazar, Eduardo
Publicado: (2025)
por: Salazar, Eduardo
Publicado: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
por: Fang, Biyi, et al.
Publicado: (2025)
por: Fang, Biyi, et al.
Publicado: (2025)
SmartEdge: Smart Healthcare End-to-End Integrated Edge and Cloud Computing System for Diabetes Prediction Enabled by Ensemble Machine Learning
por: Hennebelle, Alain, et al.
Publicado: (2025)
por: Hennebelle, Alain, et al.
Publicado: (2025)
A Mathematical Theory of Discursive Networks
por: Gutiérrez, Juan B.
Publicado: (2025)
por: Gutiérrez, Juan B.
Publicado: (2025)
Ejemplares similares
-
Aligning LLMs for Multilingual Consistency in Enterprise Applications
por: Agarwal, Amit, et al.
Publicado: (2025) -
A Comparative Study of Feature Selection in Tsetlin Machines
por: Halenka, Vojtech, et al.
Publicado: (2025) -
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
por: Aman, La Ode, et al.
Publicado: (2024) -
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities
por: Giacomelli, Stefano, et al.
Publicado: (2024) -
Graded Transformers
por: Shaska Sr, Tony
Publicado: (2025)