Balancing Computation Load and Representation Expressivity in Parallel Hybrid Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Moradi, Mohammad Mahdi, Ahmed, Walid, Wen, Shuangyue, Mudur, Sudhir, Zhang, Weiwei, Liu, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
by: Moradi, Mohammad Mahdi, et al.
Published: (2025)
by: Moradi, Mohammad Mahdi, et al.
Published: (2025)
DiSCTT: Consensus-Guided Self-Curriculum for Efficient Test-Time Adaptation in Reasoning
by: Moradi, Mohammad Mahdi, et al.
Published: (2026)
by: Moradi, Mohammad Mahdi, et al.
Published: (2026)
GC-KBVQA: A New Four-Stage Framework for Enhancing Knowledge Based Visual Question Answering Performance
by: Moradi, Mohammad Mahdi, et al.
Published: (2025)
by: Moradi, Mohammad Mahdi, et al.
Published: (2025)
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
by: Amer, Hossam, et al.
Published: (2026)
by: Amer, Hossam, et al.
Published: (2026)
Training Acceleration of Low-Rank Decomposed Networks using Sequential Freezing and Rank Quantization
by: Hajimolahoseini, Habib, et al.
Published: (2023)
by: Hajimolahoseini, Habib, et al.
Published: (2023)
Accelerating the Low-Rank Decomposed Models
by: Hajimolahoseini, Habib, et al.
Published: (2024)
by: Hajimolahoseini, Habib, et al.
Published: (2024)
Number Representations in LLMs: A Computational Parallel to Human Perception
by: AlquBoj, H. V., et al.
Published: (2025)
by: AlquBoj, H. V., et al.
Published: (2025)
Bayesian Mixture of Experts For Large Language Models
by: Dialameh, Maryam, et al.
Published: (2025)
by: Dialameh, Maryam, et al.
Published: (2025)
Gender Encoding Patterns in Pretrained Language Model Representations
by: Zakizadeh, Mahdi, et al.
Published: (2025)
by: Zakizadeh, Mahdi, et al.
Published: (2025)
ETT: Expanding the Long Context Understanding Capability of LLMs at Test-Time
by: Zahirnia, Kiarash, et al.
Published: (2025)
by: Zahirnia, Kiarash, et al.
Published: (2025)
Expert Threshold Routing for Autoregressive Language Modeling with Dynamic Computation Allocation and Load Balancing
by: Sun, Hanchi, et al.
Published: (2026)
by: Sun, Hanchi, et al.
Published: (2026)
TIC: Translate-Infer-Compile for accurate "text to plan" using LLMs and Logical Representations
by: Agarwal, Sudhir, et al.
Published: (2024)
by: Agarwal, Sudhir, et al.
Published: (2024)
Transformers are Stateless Differentiable Neural Computers
by: Tang, Bo, et al.
Published: (2026)
by: Tang, Bo, et al.
Published: (2026)
Distributed Hybrid Parallelism for Large Language Models: Comparative Study and System Design Guide
by: Amer, Hossam, et al.
Published: (2026)
by: Amer, Hossam, et al.
Published: (2026)
Lower Bounds on the Expressivity of Recurrent Neural Language Models
by: Svete, Anej, et al.
Published: (2024)
by: Svete, Anej, et al.
Published: (2024)
Envisage: Towards Expressive Visual Graph Querying
by: Wen, Xiaolin, et al.
Published: (2025)
by: Wen, Xiaolin, et al.
Published: (2025)
An In-depth Walkthrough on Evolution of Neural Machine Translation
by: Jagtap, Rohan, et al.
Published: (2020)
by: Jagtap, Rohan, et al.
Published: (2020)
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute
by: Wen, Hao, et al.
Published: (2025)
by: Wen, Hao, et al.
Published: (2025)
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-Experts
by: Yang, Jiajie
Published: (2025)
by: Yang, Jiajie
Published: (2025)
Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models
by: Qiu, Zihan, et al.
Published: (2025)
by: Qiu, Zihan, et al.
Published: (2025)
Neural Digital Twins: Toward Next-Generation Brain-Computer Interfaces
by: Bina, Mohammad Mahdi Habibi, et al.
Published: (2026)
by: Bina, Mohammad Mahdi Habibi, et al.
Published: (2026)
DrVoice: Parallel Speech-Text Voice Conversation Model via Dual-Resolution Speech Representations
by: Tan, Chao-Hong, et al.
Published: (2025)
by: Tan, Chao-Hong, et al.
Published: (2025)
Balancing Stylization and Truth via Disentangled Representation Steering
by: Shen, Chenglei, et al.
Published: (2025)
by: Shen, Chenglei, et al.
Published: (2025)
Computational Narrative Understanding for Expressive Text-to-Speech
by: Michel, Gaspard, et al.
Published: (2025)
by: Michel, Gaspard, et al.
Published: (2025)
ColBERT: Using BERT Sentence Embedding in Parallel Neural Networks for Computational Humor
by: Annamoradnejad, Issa, et al.
Published: (2020)
by: Annamoradnejad, Issa, et al.
Published: (2020)
Beyond Hard and Soft: Hybrid Context Compression for Balancing Local and Global Information Retention
by: Liao, Huanxuan, et al.
Published: (2025)
by: Liao, Huanxuan, et al.
Published: (2025)
Unlocking the Potential of ChatGPT: A Comprehensive Exploration of its Applications, Advantages, Limitations, and Future Directions in Natural Language Processing
by: Hariri, Walid
Published: (2023)
by: Hariri, Walid
Published: (2023)
Analyzing the Performance of ChatGPT in Cardiology and Vascular Pathologies
by: Hariri, Walid
Published: (2023)
by: Hariri, Walid
Published: (2023)
On The Expressivity of Recurrent Neural Cascades
by: Knorozova, Nadezda Alexandrovna, et al.
Published: (2023)
by: Knorozova, Nadezda Alexandrovna, et al.
Published: (2023)
Expressivity and Speech Synthesis
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Parallel Loop Transformer for Efficient Test-Time Computation Scaling
by: Wu, Bohong, et al.
Published: (2025)
by: Wu, Bohong, et al.
Published: (2025)
Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
by: Wang, Lean, et al.
Published: (2024)
by: Wang, Lean, et al.
Published: (2024)
FinGPT-HPC: Efficient Pretraining and Finetuning Large Language Models for Financial Applications with High-Performance Computing
by: Liu, Xiao-Yang, et al.
Published: (2024)
by: Liu, Xiao-Yang, et al.
Published: (2024)
GNN-CNN: An Efficient Hybrid Model of Convolutional and Graph Neural Networks for Text Representation
by: Rastakhiz, Fardin
Published: (2025)
by: Rastakhiz, Fardin
Published: (2025)
Partitioning Unstructured Sparse Tensor Algebra for Load-Balanced Parallel Execution
by: Chougule, Atharva, et al.
Published: (2026)
by: Chougule, Atharva, et al.
Published: (2026)
Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning
by: Wu, Tong, et al.
Published: (2025)
by: Wu, Tong, et al.
Published: (2025)
$D^2LoRA$: Data-Driven LoRA Initialization for Low Resource Tasks
by: SeraJ, Javad, et al.
Published: (2025)
by: SeraJ, Javad, et al.
Published: (2025)
Evaluating Computational Representations of Character: An Austen Character Similarity Benchmark
by: Yang, Funing, et al.
Published: (2024)
by: Yang, Funing, et al.
Published: (2024)
Kimi Linear: An Expressive, Efficient Attention Architecture
by: Kimi Team, et al.
Published: (2025)
by: Kimi Team, et al.
Published: (2025)
Balancing the Reasoning Load: Difficulty-Differentiated Policy Optimization with Length Redistribution for Efficient and Robust Reinforcement Learning
by: Xia, Yinan, et al.
Published: (2026)
by: Xia, Yinan, et al.
Published: (2026)
Similar Items
-
Continuous Self-Improvement of Large Language Models by Test-time Training with Verifier-Driven Sample Selection
by: Moradi, Mohammad Mahdi, et al.
Published: (2025) -
DiSCTT: Consensus-Guided Self-Curriculum for Efficient Test-Time Adaptation in Reasoning
by: Moradi, Mohammad Mahdi, et al.
Published: (2026) -
GC-KBVQA: A New Four-Stage Framework for Enhancing Knowledge Based Visual Question Answering Performance
by: Moradi, Mohammad Mahdi, et al.
Published: (2025) -
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
by: Amer, Hossam, et al.
Published: (2026) -
Training Acceleration of Low-Rank Decomposed Networks using Sequential Freezing and Rank Quantization
by: Hajimolahoseini, Habib, et al.
Published: (2023)