Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Vendrell, Victor Conchello, Masdemont, Arnau Padres, Grillo, Niccolò, Ros-Giralt, Jordi, Behboodi, Arash, Massoli, Fabio Valerio |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
M3Kang: Evaluating Multilingual Multimodal Mathematical Reasoning in Vision-Language Models
by: Torres-Camps, Aleix, et al.
Published: (2026)
by: Torres-Camps, Aleix, et al.
Published: (2026)
Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck
by: Massoli, Fabio Valerio, et al.
Published: (2026)
by: Massoli, Fabio Valerio, et al.
Published: (2026)
Variational Learning ISTA
by: Massoli, Fabio Valerio, et al.
Published: (2024)
by: Massoli, Fabio Valerio, et al.
Published: (2024)
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
by: Cencerrado, Iván Vicente Moreno, et al.
Published: (2025)
by: Cencerrado, Iván Vicente Moreno, et al.
Published: (2025)
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
by: Rajaee, Sara, et al.
Published: (2025)
by: Rajaee, Sara, et al.
Published: (2025)
Reinforcement Learning of Adaptive Acquisition Policies for Inverse Problems
by: Silvestri, Gianluigi, et al.
Published: (2024)
by: Silvestri, Gianluigi, et al.
Published: (2024)
Fundamental bounds on efficiency-confidence trade-off for transductive conformal prediction
by: Behboodi, Arash, et al.
Published: (2025)
by: Behboodi, Arash, et al.
Published: (2025)
An Information Theoretic Perspective on Conformal Prediction
by: Correia, Alvaro H. C., et al.
Published: (2024)
by: Correia, Alvaro H. C., et al.
Published: (2024)
Simulating, Fast and Slow: Learning Policies for Black-Box Optimization
by: Massoli, Fabio Valerio, et al.
Published: (2024)
by: Massoli, Fabio Valerio, et al.
Published: (2024)
Adaptive Loops and Memory in Transformers: Think Harder or Know More?
by: Frey, Markus, et al.
Published: (2026)
by: Frey, Markus, et al.
Published: (2026)
Vision-Assisted Digital Twin Creation for mmWave Beam Management
by: Arnold, Maximilian, et al.
Published: (2024)
by: Arnold, Maximilian, et al.
Published: (2024)
GameTalk: Training LLMs for Strategic Conversation
by: Vendrell, Victor Conchello, et al.
Published: (2026)
by: Vendrell, Victor Conchello, et al.
Published: (2026)
LUMINA: Long-horizon Understanding for Multi-turn Interactive Agents
by: Rakhsha, Amin, et al.
Published: (2026)
by: Rakhsha, Amin, et al.
Published: (2026)
Parallel Loop Transformer for Efficient Test-Time Computation Scaling
by: Wu, Bohong, et al.
Published: (2025)
by: Wu, Bohong, et al.
Published: (2025)
Graph Memory Transformer (GMT)
by: Zanarini, Nicola, et al.
Published: (2026)
by: Zanarini, Nicola, et al.
Published: (2026)
LoopRPT: Reinforcement Pre-Training for Looped Language Models
by: Tang, Guo, et al.
Published: (2026)
by: Tang, Guo, et al.
Published: (2026)
Towards General Loop Invariant Generation: A Benchmark of Programs with Memory Manipulation
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
Efficient Reasoning on the Edge
by: Bondarenko, Yelysei, et al.
Published: (2026)
by: Bondarenko, Yelysei, et al.
Published: (2026)
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
by: Chen, Guanxu, et al.
Published: (2026)
by: Chen, Guanxu, et al.
Published: (2026)
LoopFormer: Elastic-Depth Looped Transformers for Latent Reasoning via Shortcut Modulation
by: Jeddi, Ahmadreza, et al.
Published: (2026)
by: Jeddi, Ahmadreza, et al.
Published: (2026)
Learning What to Remember: Adaptive Probabilistic Memory Retention for Memory-Efficient Language Models
by: Rafiuddin, S M, et al.
Published: (2025)
by: Rafiuddin, S M, et al.
Published: (2025)
Neural Topic Modeling with Large Language Models in the Loop
by: Yang, Xiaohao, et al.
Published: (2024)
by: Yang, Xiaohao, et al.
Published: (2024)
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
by: Cesa, Gabriele, et al.
Published: (2026)
by: Cesa, Gabriele, et al.
Published: (2026)
Language Model Memory and Memory Models for Language
by: Badger, Benjamin L.
Published: (2026)
by: Badger, Benjamin L.
Published: (2026)
Scaling Latent Reasoning via Looped Language Models
by: Zhu, Rui-Jie, et al.
Published: (2025)
by: Zhu, Rui-Jie, et al.
Published: (2025)
HMT: Hierarchical Memory Transformer for Efficient Long Context Language Processing
by: He, Zifan, et al.
Published: (2024)
by: He, Zifan, et al.
Published: (2024)
From Tool Calling to Symbolic Thinking: LLMs in a Persistent Lisp Metaprogramming Loop
by: de la Torre, Jordi
Published: (2025)
by: de la Torre, Jordi
Published: (2025)
Simultaneous Computation and Memory Efficient Zeroth-Order Optimizer for Fine-Tuning Large Language Models
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Minerva: A Programmable Memory Test Benchmark for Language Models
by: Xia, Menglin, et al.
Published: (2025)
by: Xia, Menglin, et al.
Published: (2025)
Understanding and Enhancing Mamba-Transformer Hybrids for Memory Recall and Language Modeling
by: Lee, Hyunji, et al.
Published: (2025)
by: Lee, Hyunji, et al.
Published: (2025)
Structured Token Retention and Computational Memory Paths in Large Language Models
by: Delena, Jonathan, et al.
Published: (2025)
by: Delena, Jonathan, et al.
Published: (2025)
Sparse Layers are Critical to Scaling Looped Language Models
by: Lee, Ryan, et al.
Published: (2026)
by: Lee, Ryan, et al.
Published: (2026)
LOOPRAG: Enhancing Loop Transformation Optimization with Retrieval-Augmented Large Language Models
by: Zhi, Yijie, et al.
Published: (2025)
by: Zhi, Yijie, et al.
Published: (2025)
MLP Memory: A Retriever-Pretrained Memory for Large Language Models
by: Wei, Rubin, et al.
Published: (2025)
by: Wei, Rubin, et al.
Published: (2025)
MINI-LLM: Memory-Efficient Structured Pruning for Large Language Models
by: Cheng, Hongrong, et al.
Published: (2024)
by: Cheng, Hongrong, et al.
Published: (2024)
KVPruner: Structural Pruning for Faster and Memory-Efficient Large Language Models
by: Lv, Bo, et al.
Published: (2024)
by: Lv, Bo, et al.
Published: (2024)
MemoryFormer: Minimize Transformer Computation by Removing Fully-Connected Layers
by: Ding, Ning, et al.
Published: (2024)
by: Ding, Ning, et al.
Published: (2024)
Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation
by: Hu, Zhanghao, et al.
Published: (2026)
by: Hu, Zhanghao, et al.
Published: (2026)
BALI: Branch-Aware Loop Invariant Inference with Large Language Models
by: Wang, Mingxiu, et al.
Published: (2025)
by: Wang, Mingxiu, et al.
Published: (2025)
Reasoning with Latent Thoughts: On the Power of Looped Transformers
by: Saunshi, Nikunj, et al.
Published: (2025)
by: Saunshi, Nikunj, et al.
Published: (2025)
Similar Items
-
M3Kang: Evaluating Multilingual Multimodal Mathematical Reasoning in Vision-Language Models
by: Torres-Camps, Aleix, et al.
Published: (2026) -
Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck
by: Massoli, Fabio Valerio, et al.
Published: (2026) -
Variational Learning ISTA
by: Massoli, Fabio Valerio, et al.
Published: (2024) -
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
by: Cencerrado, Iván Vicente Moreno, et al.
Published: (2025) -
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
by: Rajaee, Sara, et al.
Published: (2025)