Memory Efficient Neural Processes via Constant Memory Attention Block
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Leo, Tung, Frederick, Hajimirsadeghi, Hossein, Bengio, Yoshua, Ahmed, Mohamed Osama |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tree Cross Attention
by: Feng, Leo, et al.
Published: (2023)
by: Feng, Leo, et al.
Published: (2023)
Attention as an RNN
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets
by: Liu, Zhen, et al.
Published: (2024)
by: Liu, Zhen, et al.
Published: (2024)
Learning What Matters: Steering Diffusion via Spectrally Anisotropic Forward Noise
by: Scimeca, Luca, et al.
Published: (2025)
by: Scimeca, Luca, et al.
Published: (2025)
Memory-Efficient 3D Denoising Diffusion Models for Medical Image Processing
by: Bieder, Florentin, et al.
Published: (2023)
by: Bieder, Florentin, et al.
Published: (2023)
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
VCR: A Task for Pixel-Level Complex Reasoning in Vision Language Models via Restoring Occluded Text
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
Mitigating Shortcut Learning with Diffusion Counterfactuals and Diverse Ensembles
by: Scimeca, Luca, et al.
Published: (2023)
by: Scimeca, Luca, et al.
Published: (2023)
Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
Time-, Memory- and Parameter-Efficient Visual Adaptation
by: Mercea, Otniel-Bogdan, et al.
Published: (2024)
by: Mercea, Otniel-Bogdan, et al.
Published: (2024)
MAIS: Memory-Attention for Interactive Segmentation
by: Orbes-Arteaga, Mauricio, et al.
Published: (2025)
by: Orbes-Arteaga, Mauricio, et al.
Published: (2025)
Assessing SAM for Tree Crown Instance Segmentation from Drone Imagery
by: Teng, Mélisande, et al.
Published: (2025)
by: Teng, Mélisande, et al.
Published: (2025)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
by: Choi, Seokeon, et al.
Published: (2025)
by: Choi, Seokeon, et al.
Published: (2025)
ELSA: Exact Linear-Scan Attention for Fast and Memory-Light Vision Transformers
by: Hsu, Chih-Chung, et al.
Published: (2026)
by: Hsu, Chih-Chung, et al.
Published: (2026)
QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
Attention-aware Inference Optimizations for Large Vision-Language Models with Memory-efficient Decoding
by: Ilhan, Fatih, et al.
Published: (2026)
by: Ilhan, Fatih, et al.
Published: (2026)
Memory Backdoor Attacks on Neural Networks
by: Luzon, Eden, et al.
Published: (2024)
by: Luzon, Eden, et al.
Published: (2024)
Attend Locally, Remember Linearly: Linear Attention as Cross-Frame Memory for Autoregressive Video Diffusion
by: Li, Kunyang, et al.
Published: (2026)
by: Li, Kunyang, et al.
Published: (2026)
LightCache: Memory-Efficient, Training-Free Acceleration for Video Generation
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
Topology-Guided Knowledge Distillation for Efficient Point Cloud Processing
by: Hai, Luu Tung, et al.
Published: (2025)
by: Hai, Luu Tung, et al.
Published: (2025)
Lipschitz Constant Meets Condition Number: Learning Robust and Compact Deep Neural Networks
by: Feng, Yangqi, et al.
Published: (2025)
by: Feng, Yangqi, et al.
Published: (2025)
FETCH: A Memory-Efficient Replay Approach for Continual Learning in Image Classification
by: Weißflog, Markus, et al.
Published: (2024)
by: Weißflog, Markus, et al.
Published: (2024)
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024)
by: Venkatraman, Siddarth, et al.
Published: (2024)
Adversarially Diversified Rehearsal Memory (ADRM): Mitigating Memory Overfitting Challenge in Continual Learning
by: Khan, Hikmat, et al.
Published: (2024)
by: Khan, Hikmat, et al.
Published: (2024)
MABViT -- Modified Attention Block Enhances Vision Transformers
by: Ramesh, Mahesh, et al.
Published: (2023)
by: Ramesh, Mahesh, et al.
Published: (2023)
Associative Memories in the Feature Space
by: Salvatori, Tommaso, et al.
Published: (2024)
by: Salvatori, Tommaso, et al.
Published: (2024)
Deep Clustering with Associative Memories
by: Saha, Bishwajit, et al.
Published: (2026)
by: Saha, Bishwajit, et al.
Published: (2026)
Learning from Memory: Non-Parametric Memory Augmented Self-Supervised Learning of Visual Features
by: Silva, Thalles, et al.
Published: (2024)
by: Silva, Thalles, et al.
Published: (2024)
Grow, Assess, Compress: Adaptive Backbone Scaling for Memory-Efficient Class Incremental Learning
by: Garcia-Castañeda, Adrian, et al.
Published: (2026)
by: Garcia-Castañeda, Adrian, et al.
Published: (2026)
Memory-efficient Continual Learning with Neural Collapse Contrastive
by: Dang, Trung-Anh, et al.
Published: (2024)
by: Dang, Trung-Anh, et al.
Published: (2024)
Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights
by: Wen, Qishuai, et al.
Published: (2026)
by: Wen, Qishuai, et al.
Published: (2026)
ChronoSelect: Robust Learning with Noisy Labels via Dynamics Temporal Memory
by: Wang, Jianchao, et al.
Published: (2025)
by: Wang, Jianchao, et al.
Published: (2025)
What is Memory? A Homological Perspective
by: Li, Xin
Published: (2023)
by: Li, Xin
Published: (2023)
Synthesizer Based Efficient Self-Attention for Vision Tasks
by: Zhu, Guangyang, et al.
Published: (2022)
by: Zhu, Guangyang, et al.
Published: (2022)
Memory-Efficient 4-bit Preconditioned Stochastic Optimization
by: Li, Jingyang, et al.
Published: (2024)
by: Li, Jingyang, et al.
Published: (2024)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
by: Gu, Youping, et al.
Published: (2025)
by: Gu, Youping, et al.
Published: (2025)
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
by: Wang, Hongyu, et al.
Published: (2025)
by: Wang, Hongyu, et al.
Published: (2025)
SURGEON: Memory-Adaptive Fully Test-Time Adaptation via Dynamic Activation Sparsity
by: Ma, Ke, et al.
Published: (2025)
by: Ma, Ke, et al.
Published: (2025)
TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision
by: Gillani, Syeda Anshrah, et al.
Published: (2025)
by: Gillani, Syeda Anshrah, et al.
Published: (2025)
Similar Items
-
Tree Cross Attention
by: Feng, Leo, et al.
Published: (2023) -
Attention as an RNN
by: Feng, Leo, et al.
Published: (2024) -
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024) -
Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets
by: Liu, Zhen, et al.
Published: (2024) -
Learning What Matters: Steering Diffusion via Spectrally Anisotropic Forward Noise
by: Scimeca, Luca, et al.
Published: (2025)