The Implicit Bias of Structured State Space Models Can Be Poisoned With Clean Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Slutzky, Yonatan, Alexander, Yotam, Razin, Noam, Cohen, Nadav |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States
by: Razin, Noam, et al.
Published: (2024)
by: Razin, Noam, et al.
Published: (2024)
Do Neural Networks Need Gradient Descent to Generalize? A Theoretical Study
by: Alexander, Yotam, et al.
Published: (2025)
by: Alexander, Yotam, et al.
Published: (2025)
Why Does Agentic Safety Fail to Generalize Across Tasks?
by: Slutzky, Yonatan, et al.
Published: (2026)
by: Slutzky, Yonatan, et al.
Published: (2026)
Lecture Notes on Linear Neural Networks: A Tale of Optimization and Generalization in Deep Learning
by: Cohen, Nadav, et al.
Published: (2024)
by: Cohen, Nadav, et al.
Published: (2024)
What Makes Data Suitable for a Locally Connected Neural Network? A Necessary and Sufficient Condition Based on Quantum Entanglement
by: Alexander, Yotam, et al.
Published: (2023)
by: Alexander, Yotam, et al.
Published: (2023)
Understanding Deep Learning via Notions of Rank
by: Razin, Noam
Published: (2024)
by: Razin, Noam
Published: (2024)
Why is Your Language Model a Poor Implicit Reward Model?
by: Razin, Noam, et al.
Published: (2025)
by: Razin, Noam, et al.
Published: (2025)
Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data
by: Ran-Milo, Yuval, et al.
Published: (2026)
by: Ran-Milo, Yuval, et al.
Published: (2026)
The Implicit Bias of Logit Regularization
by: Beck, Alon, et al.
Published: (2026)
by: Beck, Alon, et al.
Published: (2026)
Provable Benefits of Complex Parameterizations for Structured State Space Models
by: Ran-Milo, Yuval, et al.
Published: (2024)
by: Ran-Milo, Yuval, et al.
Published: (2024)
When Errors Can Be Beneficial: A Categorization of Imperfect Rewards for Policy Gradient
by: Shang, Shuning, et al.
Published: (2026)
by: Shang, Shuning, et al.
Published: (2026)
BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF
by: Duan, Kaiwen, et al.
Published: (2025)
by: Duan, Kaiwen, et al.
Published: (2025)
On the Expressive Power of Sparse Geometric MPNNs
by: Sverdlov, Yonatan, et al.
Published: (2024)
by: Sverdlov, Yonatan, et al.
Published: (2024)
Certified Robustness to Clean-Label Poisoning Using Diffusion Denoising
by: Hong, Sanghyun, et al.
Published: (2024)
by: Hong, Sanghyun, et al.
Published: (2024)
The Ultimate Cookbook for Invisible Poison: Crafting Subtle Clean-Label Text Backdoors with Style Attributes
by: You, Wencong, et al.
Published: (2025)
by: You, Wencong, et al.
Published: (2025)
Poisoning the Inner Prediction Logic of Graph Neural Networks for Clean-Label Backdoor Attacks
by: Zhang, Yuxiang, et al.
Published: (2026)
by: Zhang, Yuxiang, et al.
Published: (2026)
Can Implicit Bias Imply Adversarial Robustness?
by: Min, Hancheng, et al.
Published: (2024)
by: Min, Hancheng, et al.
Published: (2024)
Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting
by: Chen, Howard, et al.
Published: (2025)
by: Chen, Howard, et al.
Published: (2025)
Wicked Oddities: Selectively Poisoning for Effective Clean-Label Backdoor Attacks
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations
by: Sverdlov, Yonatan, et al.
Published: (2024)
by: Sverdlov, Yonatan, et al.
Published: (2024)
Monotone and Separable Set Functions: Characterizations and Neural Models
by: Sarangi, Soutrik, et al.
Published: (2025)
by: Sarangi, Soutrik, et al.
Published: (2025)
Mamba Knockout for Unraveling Factual Information Flow
by: Endy, Nir, et al.
Published: (2025)
by: Endy, Nir, et al.
Published: (2025)
Contextual Plackett-Luce: An Efficient Neural Model for Probabilistic Sequence Selection under Ambiguity
by: Mizrachi, Noam, et al.
Published: (2026)
by: Mizrachi, Noam, et al.
Published: (2026)
Tuning Frequency Bias of State Space Models
by: Yu, Annan, et al.
Published: (2024)
by: Yu, Annan, et al.
Published: (2024)
Implicit Regularization Towards Rank Minimization in ReLU Networks
by: Timor, Nadav, et al.
Published: (2022)
by: Timor, Nadav, et al.
Published: (2022)
Understanding Nonlinear Implicit Bias via Region Counts in Input Space
by: Li, Jingwei, et al.
Published: (2025)
by: Li, Jingwei, et al.
Published: (2025)
Decoupled Weight Decay for Any $p$ Norm
by: Outmezguine, Nadav Joseph, et al.
Published: (2024)
by: Outmezguine, Nadav Joseph, et al.
Published: (2024)
Adversarial Bias: Data Poisoning Attacks on Fairness
by: Chan, Eunice, et al.
Published: (2025)
by: Chan, Eunice, et al.
Published: (2025)
BXRL: Behavior-Explainable Reinforcement Learning
by: Rachum, Ram, et al.
Published: (2026)
by: Rachum, Ram, et al.
Published: (2026)
FSW-GNN: A Bi-Lipschitz WL-Equivalent Graph Neural Network
by: Sverdlov, Yonatan, et al.
Published: (2024)
by: Sverdlov, Yonatan, et al.
Published: (2024)
When and How to Canonize: A Generalization Perspective
by: Sverdlov, Yonatan, et al.
Published: (2026)
by: Sverdlov, Yonatan, et al.
Published: (2026)
Short-Range Oversquashing
by: Mishayev, Yaaqov, et al.
Published: (2025)
by: Mishayev, Yaaqov, et al.
Published: (2025)
Partial-Label Learning with Conformal Candidate Cleaning
by: Fuchs, Tobias, et al.
Published: (2025)
by: Fuchs, Tobias, et al.
Published: (2025)
What Makes a Reward Model a Good Teacher? An Optimization Perspective
by: Razin, Noam, et al.
Published: (2025)
by: Razin, Noam, et al.
Published: (2025)
Some Robustness Properties of Label Cleaning
by: Cheng, Chen, et al.
Published: (2025)
by: Cheng, Chen, et al.
Published: (2025)
SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot
by: Tuo, Kaiwen, et al.
Published: (2025)
by: Tuo, Kaiwen, et al.
Published: (2025)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
by: Itzhak, Itay, et al.
Published: (2023)
by: Itzhak, Itay, et al.
Published: (2023)
Learning Implicit Bias in Generative Spaces for Accelerating Protein Dynamics Emulation
by: Cheng, Kaihui, et al.
Published: (2026)
by: Cheng, Kaihui, et al.
Published: (2026)
Inertial Navigation Meets Deep Learning: A Survey of Current Trends and Future Directions
by: Cohen, Nadav, et al.
Published: (2023)
by: Cohen, Nadav, et al.
Published: (2023)
Blameless Users in a Clean Room: Defining Copyright Protection for Generative Models
by: Cohen, Aloni
Published: (2025)
by: Cohen, Aloni
Published: (2025)
Similar Items
-
Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States
by: Razin, Noam, et al.
Published: (2024) -
Do Neural Networks Need Gradient Descent to Generalize? A Theoretical Study
by: Alexander, Yotam, et al.
Published: (2025) -
Why Does Agentic Safety Fail to Generalize Across Tasks?
by: Slutzky, Yonatan, et al.
Published: (2026) -
Lecture Notes on Linear Neural Networks: A Tale of Optimization and Generalization in Deep Learning
by: Cohen, Nadav, et al.
Published: (2024) -
What Makes Data Suitable for a Locally Connected Neural Network? A Necessary and Sufficient Condition Based on Quantum Entanglement
by: Alexander, Yotam, et al.
Published: (2023)