SMR: State Memory Replay for Long Sequence Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Qi, Biqing, Gao, Junqi, Zhang, Kaiyan, Li, Dong, Liu, Jianxing, Wu, Ligang, Zhou, Bowen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Adversarial Transferability via Information Bottleneck Constraints
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Interactive Continual Learning: Fast and Slow Thinking
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning
by: Li, Dong, et al.
Published: (2024)
by: Li, Dong, et al.
Published: (2024)
Contrastive Augmented Graph2Graph Memory Interaction for Few Shot Continual Learning
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Less is More: Efficient Model Merging with Binary Task Switch
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Exploring Adversarial Robustness of Deep State Space Models
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Investigating Deep Watermark Security: An Adversarial Transferability Perspective
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
WIST: Web-Grounded Iterative Self-Play Tree for Domain-Targeted Reasoning Improvement
by: Li, Fangyuan, et al.
Published: (2026)
by: Li, Fangyuan, et al.
Published: (2026)
Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression
by: Gao, Junqi, et al.
Published: (2026)
by: Gao, Junqi, et al.
Published: (2026)
Fast and Slow Gradient Approximation for Binary Neural Network Optimization
by: Chen, Xinquan, et al.
Published: (2024)
by: Chen, Xinquan, et al.
Published: (2024)
PDAC: Efficient Coreset Selection for Continual Learning via Probability Density Awareness
by: Gao, Junqi, et al.
Published: (2025)
by: Gao, Junqi, et al.
Published: (2025)
Bohdi: Heterogeneous LLM Fusion with Automatic Data Exploration
by: Gao, Junqi, et al.
Published: (2025)
by: Gao, Junqi, et al.
Published: (2025)
Perturbation Towards Easy Samples Improves Targeted Adversarial Transferability
by: Gao, Junqi, et al.
Published: (2024)
by: Gao, Junqi, et al.
Published: (2024)
Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism
by: Jiang, Yuhua, et al.
Published: (2025)
by: Jiang, Yuhua, et al.
Published: (2025)
LongVQ: Long Sequence Modeling with Vector Quantization on Structured Memory
by: Liu, Zicheng, et al.
Published: (2024)
by: Liu, Zicheng, et al.
Published: (2024)
Towards Building Specialized Generalist AI with System 1 and System 2 Fusion
by: Zhang, Kaiyan, et al.
Published: (2024)
by: Zhang, Kaiyan, et al.
Published: (2024)
Evolution of Thought: Diverse and High-Quality Reasoning via Multi-Objective Optimization
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
SR-CIS: Self-Reflective Incremental System with Decoupled Memory and Reasoning
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Hybrid Memory Replay: Blending Real and Distilled Data for Class Incremental Learning
by: Kong, Jiangtao, et al.
Published: (2024)
by: Kong, Jiangtao, et al.
Published: (2024)
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning
by: Zhao, Kaiyan, et al.
Published: (2024)
by: Zhao, Kaiyan, et al.
Published: (2024)
SDAR: A Synergistic Diffusion-AutoRegression Paradigm for Scalable Sequence Generation
by: Cheng, Shuang, et al.
Published: (2025)
by: Cheng, Shuang, et al.
Published: (2025)
MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation
by: Wang, Shijie, et al.
Published: (2026)
by: Wang, Shijie, et al.
Published: (2026)
Enhancing LLM Agents for Code Generation with Possibility and Pass-rate Prioritized Experience Replay
by: Chen, Yuyang, et al.
Published: (2024)
by: Chen, Yuyang, et al.
Published: (2024)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
by: Liu, Jinyi, et al.
Published: (2023)
by: Liu, Jinyi, et al.
Published: (2023)
Non-Uniform Memory Sampling in Experience Replay
by: Krutsylo, Andrii
Published: (2025)
by: Krutsylo, Andrii
Published: (2025)
Back-stepping Experience Replay with Application to Model-free Reinforcement Learning for a Soft Snake Robot
by: Qi, Xinda, et al.
Published: (2024)
by: Qi, Xinda, et al.
Published: (2024)
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2026)
by: Feng, Yujie, et al.
Published: (2026)
Mini-Sequence Transformer: Optimizing Intermediate Memory for Long Sequences Training
by: Luo, Cheng, et al.
Published: (2024)
by: Luo, Cheng, et al.
Published: (2024)
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling
by: Liu, Runze, et al.
Published: (2025)
by: Liu, Runze, et al.
Published: (2025)
CoMemNet: Contrastive Sampling with Memory Replay Network for Continual Traffic Prediction
by: Wu, Mei, et al.
Published: (2026)
by: Wu, Mei, et al.
Published: (2026)
Graph Continual Learning with Debiased Lossless Memory Replay
by: Niu, Chaoxi, et al.
Published: (2024)
by: Niu, Chaoxi, et al.
Published: (2024)
On Large Language Models' Hallucination with Regard to Known Facts
by: Jiang, Che, et al.
Published: (2024)
by: Jiang, Che, et al.
Published: (2024)
KOSS: Kalman-Optimal Selective State Spaces for Long-Term Sequence Modeling
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
by: Jiang, Yuhua, et al.
Published: (2025)
by: Jiang, Yuhua, et al.
Published: (2025)
Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention
by: Lv, Xingtai, et al.
Published: (2024)
by: Lv, Xingtai, et al.
Published: (2024)
Information-Theoretic Generalization Bounds of Replay-based Continual Learning
by: Wen, Wen, et al.
Published: (2025)
by: Wen, Wen, et al.
Published: (2025)
Technologies on Effectiveness and Efficiency: A Survey of State Spaces Models
by: Lv, Xingtai, et al.
Published: (2025)
by: Lv, Xingtai, et al.
Published: (2025)
Compute-in-Memory Implementation of State Space Models for Event Sequence Processing
by: Zhang, Xiaoyu, et al.
Published: (2025)
by: Zhang, Xiaoyu, et al.
Published: (2025)
Better Generative Replay for Continual Federated Learning
by: Qi, Daiqing, et al.
Published: (2023)
by: Qi, Daiqing, et al.
Published: (2023)
Similar Items
-
Enhancing Adversarial Transferability via Information Bottleneck Constraints
by: Qi, Biqing, et al.
Published: (2024) -
Interactive Continual Learning: Fast and Slow Thinking
by: Qi, Biqing, et al.
Published: (2024) -
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
by: Qi, Biqing, et al.
Published: (2024) -
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning
by: Li, Dong, et al.
Published: (2024) -
Contrastive Augmented Graph2Graph Memory Interaction for Few Shot Continual Learning
by: Qi, Biqing, et al.
Published: (2024)