Saved in:
| Main Author: | Mohammed, Ahmed Abdelmuniem Abdalla |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.05222 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing
by: Filippova, Anastasiia, et al.
Published: (2026)
by: Filippova, Anastasiia, et al.
Published: (2026)
Learning to Route LLMs with Confidence Tokens
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
How Transformers Learn to Plan via Multi-Token Prediction
by: Huang, Jianhao, et al.
Published: (2026)
by: Huang, Jianhao, et al.
Published: (2026)
Learning to Route: Per-Sample Adaptive Routing for Multimodal Multitask Prediction
by: Ajirak, Marzieh, et al.
Published: (2025)
by: Ajirak, Marzieh, et al.
Published: (2025)
TARo: Token-level Adaptive Routing for LLM Test-time Alignment
by: Rai, Arushi, et al.
Published: (2026)
by: Rai, Arushi, et al.
Published: (2026)
Token-Level LLM Collaboration via FusionRoute
by: Xiong, Nuoya, et al.
Published: (2026)
by: Xiong, Nuoya, et al.
Published: (2026)
TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment
by: Wang, Jiaxuan, et al.
Published: (2026)
by: Wang, Jiaxuan, et al.
Published: (2026)
Global-Lens Transformers: Adaptive Token Mixing for Dynamic Link Prediction
by: Zou, Tao, et al.
Published: (2025)
by: Zou, Tao, et al.
Published: (2025)
BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute
by: Ding, Dujian, et al.
Published: (2025)
by: Ding, Dujian, et al.
Published: (2025)
Route Experts by Sequence, not by Token
by: Wen, Tiansheng, et al.
Published: (2025)
by: Wen, Tiansheng, et al.
Published: (2025)
CADENCE: Context-Adaptive Depth Estimation for Navigation and Computational Efficiency
by: Johnsen, Timothy K, et al.
Published: (2026)
by: Johnsen, Timothy K, et al.
Published: (2026)
MedFormer-UR: Uncertainty-Routed Transformer for Medical Image Classification
by: Sibhai, Mohammed Maaz, et al.
Published: (2026)
by: Sibhai, Mohammed Maaz, et al.
Published: (2026)
An Adaptive Simulated Annealing-Based Machine Learning Approach for Developing an E-Triage Tool for Hospital Emergency Operations
by: Ahmed, Abdulaziz, et al.
Published: (2022)
by: Ahmed, Abdulaziz, et al.
Published: (2022)
Directional Routing in Transformers
by: Taylor, Kevin
Published: (2026)
by: Taylor, Kevin
Published: (2026)
Subjective Depth and Timescale Transformers: Learning Where and When to Compute
by: Wieser, Frederico, et al.
Published: (2025)
by: Wieser, Frederico, et al.
Published: (2025)
Probing the Embedding Space of Transformers via Minimal Token Perturbations
by: Conti, Eddie, et al.
Published: (2025)
by: Conti, Eddie, et al.
Published: (2025)
VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic Routing
by: Wang, Yixiao, et al.
Published: (2025)
by: Wang, Yixiao, et al.
Published: (2025)
Cross-Domain Few-Shot Learning via Adaptive Transformer Networks
by: Paeedeh, Naeem, et al.
Published: (2024)
by: Paeedeh, Naeem, et al.
Published: (2024)
Steering Frozen LLMs: Adaptive Social Alignment via Online Prompt Routing
by: Zhang, Zeyu, et al.
Published: (2026)
by: Zhang, Zeyu, et al.
Published: (2026)
SliceMoE: Routing Embedding Slices Instead of Tokens for Fine-Grained and Balanced Transformer Scaling
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
Adaptive Computation Pruning for the Forgetting Transformer
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Geometric Mixture-of-Experts with Curvature-Guided Adaptive Routing for Graph Representation Learning
by: Cao, Haifang, et al.
Published: (2026)
by: Cao, Haifang, et al.
Published: (2026)
Token-Level Prompt Mixture with Parameter-Free Routing for Federated Domain Generalization
by: Gong, Shuai, et al.
Published: (2025)
by: Gong, Shuai, et al.
Published: (2025)
Depth-Adaptive Graph Neural Networks via Learnable Bakry-'Emery Curvature
by: Hevapathige, Asela, et al.
Published: (2025)
by: Hevapathige, Asela, et al.
Published: (2025)
SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning
by: Xu, Zhongling, et al.
Published: (2026)
by: Xu, Zhongling, et al.
Published: (2026)
Graph Tokenization for Bridging Graphs and Transformers
by: Guo, Zeyuan, et al.
Published: (2026)
by: Guo, Zeyuan, et al.
Published: (2026)
Efficient Real-Time Aircraft ETA Prediction via Feature Tokenization Transformer
by: Huang, Liping, et al.
Published: (2025)
by: Huang, Liping, et al.
Published: (2025)
RouteFormer: A Transformer-Based Routing Framework for Autonomous Vehicles
by: Youssef, Yazan, et al.
Published: (2025)
by: Youssef, Yazan, et al.
Published: (2025)
Adaptive Token-Weighted Differential Privacy for LLMs: Not All Tokens Require Equal Protection
by: Yu, Manjiang, et al.
Published: (2025)
by: Yu, Manjiang, et al.
Published: (2025)
Adaptive Information Routing for Multimodal Time Series Forecasting
by: Seo, Jun, et al.
Published: (2025)
by: Seo, Jun, et al.
Published: (2025)
SkillOrchestra: Learning to Route Agents via Skill Transfer
by: Wang, Jiayu, et al.
Published: (2026)
by: Wang, Jiayu, et al.
Published: (2026)
Enhancing Cross-Problem Vehicle Routing via Federated Learning
by: Meng, Xiangchi, et al.
Published: (2026)
by: Meng, Xiangchi, et al.
Published: (2026)
Enhanced Graph Transformer with Serialized Graph Tokens
by: Wang, Ruixiang, et al.
Published: (2026)
by: Wang, Ruixiang, et al.
Published: (2026)
Rethinking Tokenized Graph Transformers for Node Classification
by: Chen, Jinsong, et al.
Published: (2025)
by: Chen, Jinsong, et al.
Published: (2025)
PR-CapsNet: Pseudo-Riemannian Capsule Network with Adaptive Curvature Routing for Graph Learning
by: Qin, Ye, et al.
Published: (2025)
by: Qin, Ye, et al.
Published: (2025)
ARROW: An Adaptive Rollout and Routing Method for Global Weather Forecasting
by: Tian, Jindong, et al.
Published: (2025)
by: Tian, Jindong, et al.
Published: (2025)
SaDiT: Efficient Protein Backbone Design via Latent Structural Tokenization and Diffusion Transformers
by: Mo, Shentong, et al.
Published: (2026)
by: Mo, Shentong, et al.
Published: (2026)
Adaptive Online Learning with LSTM Networks for Energy Price Prediction
by: Salihoglu, Salih, et al.
Published: (2025)
by: Salihoglu, Salih, et al.
Published: (2025)
The Depth Delusion: Why Transformers Should Be Wider, Not Deeper
by: Fahim, Md Muhtasim Munif, et al.
Published: (2026)
by: Fahim, Md Muhtasim Munif, et al.
Published: (2026)
Gradient Routing: Masking Gradients to Localize Computation in Neural Networks
by: Cloud, Alex, et al.
Published: (2024)
by: Cloud, Alex, et al.
Published: (2024)
Similar Items
-
Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing
by: Filippova, Anastasiia, et al.
Published: (2026) -
Learning to Route LLMs with Confidence Tokens
by: Chuang, Yu-Neng, et al.
Published: (2024) -
How Transformers Learn to Plan via Multi-Token Prediction
by: Huang, Jianhao, et al.
Published: (2026) -
Learning to Route: Per-Sample Adaptive Routing for Multimodal Multitask Prediction
by: Ajirak, Marzieh, et al.
Published: (2025) -
TARo: Token-level Adaptive Routing for LLM Test-time Alignment
by: Rai, Arushi, et al.
Published: (2026)