Understanding Input Selectivity in Mamba: Impact on Approximation Power, Memorization, and Associative Recall Capacity
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Ningyuan, Sarabia, Miguel, Moudgil, Abhinav, Rodriguez, Pau, Zappella, Luca, Danieli, Federico |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attention to Mamba: A Recipe for Cross-Architecture Distillation
by: Moudgil, Abhinav, et al.
Published: (2026)
by: Moudgil, Abhinav, et al.
Published: (2026)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
by: Danieli, Federico, et al.
Published: (2025)
by: Danieli, Federico, et al.
Published: (2025)
Machine Learning for Physical Simulation Challenge Results and Retrospective Analysis: Power Grid Use Case
by: Leyli-Abadi, Milad, et al.
Published: (2025)
by: Leyli-Abadi, Milad, et al.
Published: (2025)
ECG-FM: An Open Electrocardiogram Foundation Model
by: McKeen, Kaden, et al.
Published: (2024)
by: McKeen, Kaden, et al.
Published: (2024)
Feature Relevancy, Necessity and Usefulness: Complexity and Algorithms
by: Capdevielle, Tomás, et al.
Published: (2025)
by: Capdevielle, Tomás, et al.
Published: (2025)
A Theoretical Framework for Adaptive Utility-Weighted Benchmarking
by: Waggoner, Philip
Published: (2026)
by: Waggoner, Philip
Published: (2026)
From Language Models to Practical Self-Improving Computer Agents
by: Sheng, Alex
Published: (2024)
by: Sheng, Alex
Published: (2024)
A Taxonomy of Omnicidal Futures Involving Artificial Intelligence
by: Critch, Andrew, et al.
Published: (2025)
by: Critch, Andrew, et al.
Published: (2025)
Enhancing PyKEEN with Multiple Negative Sampling Solutions for Knowledge Graph Embedding Models
by: d'Amato, Claudia, et al.
Published: (2025)
by: d'Amato, Claudia, et al.
Published: (2025)
Driving down Poisson error can offset classification error in clinical tasks
by: Delahunt, Charles B., et al.
Published: (2024)
by: Delahunt, Charles B., et al.
Published: (2024)
Attack Selection Reduces Safety in Concentrated AI Control Settings against Trusted Monitoring
by: Schaeffer, Joachim, et al.
Published: (2026)
by: Schaeffer, Joachim, et al.
Published: (2026)
A ZeNN architecture to avoid the Gaussian trap
by: Carvalho, Luís, et al.
Published: (2025)
by: Carvalho, Luís, et al.
Published: (2025)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
by: Bian, Yijun, et al.
Published: (2024)
by: Bian, Yijun, et al.
Published: (2024)
Approximating Discrimination Within Models When Faced With Several Non-Binary Sensitive Attributes
by: Bian, Yijun, et al.
Published: (2024)
by: Bian, Yijun, et al.
Published: (2024)
Quantifying Behavioral Dissimilarity Between Mathematical Expressions
by: Mežnar, Sebastian, et al.
Published: (2024)
by: Mežnar, Sebastian, et al.
Published: (2024)
Intelligence as Computation
by: Brock, Oliver
Published: (2024)
by: Brock, Oliver
Published: (2024)
Report of the 2025 Workshop on Next-Generation Ecosystems for Scientific Computing: Harnessing Community, Software, and AI for Cross-Disciplinary Team Science
by: McInnes, Lois Curfman, et al.
Published: (2025)
by: McInnes, Lois Curfman, et al.
Published: (2025)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
Instilling Organisational Values in Firefighters through Simulation-Based Training
by: Osman, Nardine, et al.
Published: (2025)
by: Osman, Nardine, et al.
Published: (2025)
Value-Aware Multiagent Systems
by: Osman, Nardine
Published: (2025)
by: Osman, Nardine
Published: (2025)
Dynamic Observation Policies in Observation Cost-Sensitive Reinforcement Learning
by: Bellinger, Colin, et al.
Published: (2023)
by: Bellinger, Colin, et al.
Published: (2023)
On the Invariants of Softmax Attention
by: Lee, Wonsuk
Published: (2026)
by: Lee, Wonsuk
Published: (2026)
Deploying Large Language Models With Retrieval Augmented Generation
by: Prabhune, Sonal, et al.
Published: (2024)
by: Prabhune, Sonal, et al.
Published: (2024)
Heckerthoughts
by: Heckerman, David
Published: (2023)
by: Heckerman, David
Published: (2023)
Charting the Future of Scholarly Knowledge with AI: A Community Perspective
by: Jiomekong, Azanzi, et al.
Published: (2025)
by: Jiomekong, Azanzi, et al.
Published: (2025)
Achieving Distributive Justice in Federated Learning via Uncertainty Quantification
by: Carey, Alycia, et al.
Published: (2025)
by: Carey, Alycia, et al.
Published: (2025)
Return of the Schema: Building Complete Datasets for Machine Learning and Reasoning on Knowledge Graphs
by: Diliso, Ivan, et al.
Published: (2026)
by: Diliso, Ivan, et al.
Published: (2026)
ATEX-CF: Attack-Informed Counterfactual Explanations for Graph Neural Networks
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
Compressible Softmax-Attended Language under Incompressible Attention
by: Lee, Wonsuk
Published: (2026)
by: Lee, Wonsuk
Published: (2026)
Proceedings of the 20th International Conference on Knowledge, Information and Creativity Support Systems (KICSS 2025)
by: Hayama, Edited by Tessai, et al.
Published: (2025)
by: Hayama, Edited by Tessai, et al.
Published: (2025)
Vibe-Creation: The Epistemology of Human-AI Emergent Cognition
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Dancing in the Shadows: Harnessing Ambiguity for Fairer Classifiers
by: Barrainkua, Ainhize, et al.
Published: (2024)
by: Barrainkua, Ainhize, et al.
Published: (2024)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
by: Lomasto, Luigi, et al.
Published: (2026)
by: Lomasto, Luigi, et al.
Published: (2026)
Combination of Weak Learners eXplanations to Improve Random Forest eXplicability Robustness
by: Pala, Riccardo, et al.
Published: (2024)
by: Pala, Riccardo, et al.
Published: (2024)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
by: Abhishek, Alok, et al.
Published: (2026)
by: Abhishek, Alok, et al.
Published: (2026)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
by: de Haan, Ian B., et al.
Published: (2026)
by: de Haan, Ian B., et al.
Published: (2026)
Ideological Isolation in Online Social Networks: A Survey of Computational Definitions, Metrics, and Mitigation Strategies
by: Wang, Xiaodan, et al.
Published: (2026)
by: Wang, Xiaodan, et al.
Published: (2026)
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
by: Zhao, Zelin, et al.
Published: (2023)
by: Zhao, Zelin, et al.
Published: (2023)
Similar Items
-
Attention to Mamba: A Recipe for Cross-Architecture Distillation
by: Moudgil, Abhinav, et al.
Published: (2026) -
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
by: Danieli, Federico, et al.
Published: (2025) -
Machine Learning for Physical Simulation Challenge Results and Retrospective Analysis: Power Grid Use Case
by: Leyli-Abadi, Milad, et al.
Published: (2025) -
ECG-FM: An Open Electrocardiogram Foundation Model
by: McKeen, Kaden, et al.
Published: (2024) -
Feature Relevancy, Necessity and Usefulness: Complexity and Algorithms
by: Capdevielle, Tomás, et al.
Published: (2025)