Duo-LLM: A Framework for Studying Adaptive Computation in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Alizadeh, Keivan, Mirzadeh, Iman, Shahrokhi, Hooman, Belenko, Dmitry, Sun, Frank, Cho, Minsik, Sekhavat, Mohammad Hossein, Nabi, Moin, Farajtabar, Mehrdad |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Computational Bottlenecks of Training Small-scale Large Language Models
by: Ashkboos, Saleh, et al.
Published: (2024)
by: Ashkboos, Saleh, et al.
Published: (2024)
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
by: Alizadeh, Keivan, et al.
Published: (2023)
by: Alizadeh, Keivan, et al.
Published: (2023)
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
by: Mirzadeh, Iman, et al.
Published: (2024)
by: Mirzadeh, Iman, et al.
Published: (2024)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
by: Samragh, Mohammad, et al.
Published: (2024)
by: Samragh, Mohammad, et al.
Published: (2024)
Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context
by: Alizadeh, Keivan, et al.
Published: (2026)
by: Alizadeh, Keivan, et al.
Published: (2026)
SALSA: Soup-based Alignment Learning for Stronger Adaptation in RLHF
by: Chegini, Atoosa, et al.
Published: (2024)
by: Chegini, Atoosa, et al.
Published: (2024)
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
by: Shojaee, Parshin, et al.
Published: (2025)
by: Shojaee, Parshin, et al.
Published: (2025)
OpenELM: An Efficient Language Model Family with Open Training and Inference Framework
by: Mehta, Sachin, et al.
Published: (2024)
by: Mehta, Sachin, et al.
Published: (2024)
Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity
by: Joudaki, Amir, et al.
Published: (2025)
by: Joudaki, Amir, et al.
Published: (2025)
From Dense to Dynamic: Token-Difficulty Driven MoEfication of Pre-Trained LLMs
by: Nishu, Kumari, et al.
Published: (2025)
by: Nishu, Kumari, et al.
Published: (2025)
Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential
by: Samragh, Mohammad, et al.
Published: (2025)
by: Samragh, Mohammad, et al.
Published: (2025)
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
by: Mehta, Sachin, et al.
Published: (2024)
by: Mehta, Sachin, et al.
Published: (2024)
MemoryLLM: Plug-n-Play Interpretable Feed-Forward Memory for Transformers
by: Jaiswal, Ajay, et al.
Published: (2026)
by: Jaiswal, Ajay, et al.
Published: (2026)
Contrastive Perplexity for Controlled Generation: An Application in Detoxifying Large Language Models
by: Klein, Tassilo, et al.
Published: (2024)
by: Klein, Tassilo, et al.
Published: (2024)
TIDE: Every Layer Knows the Token Beneath the Context
by: Jaiswal, Ajay, et al.
Published: (2026)
by: Jaiswal, Ajay, et al.
Published: (2026)
MoE-PHDS: One MoE checkpoint for flexible runtime sparsity
by: Hannah, Lauren. A, et al.
Published: (2025)
by: Hannah, Lauren. A, et al.
Published: (2025)
TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining
by: Li, Jeffrey, et al.
Published: (2025)
by: Li, Jeffrey, et al.
Published: (2025)
KV-Runahead: Scalable Causal LLM Inference by Parallel Key-Value Cache Generation
by: Cho, Minsik, et al.
Published: (2024)
by: Cho, Minsik, et al.
Published: (2024)
RL for Reasoning by Adaptively Revealing Rationales
by: Amani, Mohammad Hossein, et al.
Published: (2025)
by: Amani, Mohammad Hossein, et al.
Published: (2025)
Learning Private Representations through Entropy-based Adversarial Training
by: Klein, Tassilo, et al.
Published: (2025)
by: Klein, Tassilo, et al.
Published: (2025)
Network-Based Video Recommendation Using Viewing Patterns and Modularity Analysis: An Integrated Framework
by: Maghsoudi, Mehrdad, et al.
Published: (2023)
by: Maghsoudi, Mehrdad, et al.
Published: (2023)
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
by: Armandpour, Mohammadreza, et al.
Published: (2026)
by: Armandpour, Mohammadreza, et al.
Published: (2026)
Gray-Box Computed Torque Control for Differential-Drive Mobile Robot Tracking
by: Pishkhani, Arman Javan Sekhavat
Published: (2025)
by: Pishkhani, Arman Javan Sekhavat
Published: (2025)
SPD: Sync-Point Drop for Efficient Tensor Parallelism of Large Language Models
by: Kim, Han-Byul, et al.
Published: (2025)
by: Kim, Han-Byul, et al.
Published: (2025)
DeFi Liquidation Risk Modeling Using Geometric Brownian Motion
by: Belenko, Timofei, et al.
Published: (2025)
by: Belenko, Timofei, et al.
Published: (2025)
The path towards contact-based physical human-robot interaction
by: Farajtabar, Mohammad, et al.
Published: (2024)
by: Farajtabar, Mohammad, et al.
Published: (2024)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
by: Fu, Qichen, et al.
Published: (2024)
by: Fu, Qichen, et al.
Published: (2024)
Adaptive Observer‐Based Event‐Triggered Fault‐Tolerant Control of Multiagent Systems Subject to Output Quantization, Input Nonlinearities, and Unknown Actuator and Sensor Faults
by: Maryam Beigizadeh, et al.
Published: (2025)
by: Maryam Beigizadeh, et al.
Published: (2025)
AeroDuo: Aerial Duo for UAV-based Vision and Language Navigation
by: Wu, Ruipu, et al.
Published: (2025)
by: Wu, Ruipu, et al.
Published: (2025)
Towards Low-bit Communication for Tensor Parallel LLM Inference
by: Dong, Harry, et al.
Published: (2024)
by: Dong, Harry, et al.
Published: (2024)
SpecMD: A Comprehensive Study On Speculative Expert Prefetching
by: Hoang, Duc, et al.
Published: (2026)
by: Hoang, Duc, et al.
Published: (2026)
LLM4C2Rust: Large Language Models for Automated Memory-Safe Code Transpilation
by: Bedell, Sarah, et al.
Published: (2026)
by: Bedell, Sarah, et al.
Published: (2026)
SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents
by: Saberi, Mehrdad, et al.
Published: (2026)
by: Saberi, Mehrdad, et al.
Published: (2026)
Generalized convex functions and their applications in optimality conditions
by: Alizadeh, Mohammad Hossein, et al.
Published: (2024)
by: Alizadeh, Mohammad Hossein, et al.
Published: (2024)
Conformal Prediction Sets for Deep Generative Models via Reduction to Conformal Regression
by: Shahrokhi, Hooman, et al.
Published: (2025)
by: Shahrokhi, Hooman, et al.
Published: (2025)
The hybrid dilemma -- do hybrid technologies play a transitionary or stationary role in transitions processes?
by: Phirouzabadi, Amir Mirzadeh
Published: (2025)
by: Phirouzabadi, Amir Mirzadeh
Published: (2025)
Push‐Out Bond Strength of Fiber‐Reinforced Post Using Various Post Space Irrigation Treatments
by: Amir Mohammad Shahrokhi, et al.
Published: (2024)
by: Amir Mohammad Shahrokhi, et al.
Published: (2024)
DuoLungo: Usability Study of Duo 2FA
by: Prapty, Renascence Tarafder, et al.
Published: (2026)
by: Prapty, Renascence Tarafder, et al.
Published: (2026)
DuoCast: Duo-Probabilistic Diffusion for Precipitation Nowcasting
by: Wen, Penghui, et al.
Published: (2024)
by: Wen, Penghui, et al.
Published: (2024)
Impact of Geometric Uncertainty on the Computation of Abdominal Aortic Aneurysm Wall Strain
by: Sekhavat, Saeideh, et al.
Published: (2025)
by: Sekhavat, Saeideh, et al.
Published: (2025)
Similar Items
-
Computational Bottlenecks of Training Small-scale Large Language Models
by: Ashkboos, Saleh, et al.
Published: (2024) -
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
by: Alizadeh, Keivan, et al.
Published: (2023) -
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
by: Mirzadeh, Iman, et al.
Published: (2024) -
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
by: Samragh, Mohammad, et al.
Published: (2024) -
Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context
by: Alizadeh, Keivan, et al.
Published: (2026)