Dictionary Learning: The Complexity of Learning Sparse Superposed Features with Feedback
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Kumar, Akash |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Superpose Task-specific Features for Model Merging
von: Qiu, Haiquan, et al.
Veröffentlicht: (2025)
von: Qiu, Haiquan, et al.
Veröffentlicht: (2025)
Identifying Functionally Important Features with End-to-End Sparse Dictionary Learning
von: Braun, Dan, et al.
Veröffentlicht: (2024)
von: Braun, Dan, et al.
Veröffentlicht: (2024)
Improving Dictionary Learning with Gated Sparse Autoencoders
von: Rajamanoharan, Senthooran, et al.
Veröffentlicht: (2024)
von: Rajamanoharan, Senthooran, et al.
Veröffentlicht: (2024)
A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
Impact of Label Noise on Learning Complex Features
von: Vashisht, Rahul, et al.
Veröffentlicht: (2024)
von: Vashisht, Rahul, et al.
Veröffentlicht: (2024)
Features that Make a Difference: Leveraging Gradients for Improved Dictionary Learning
von: Olmo, Jeffrey, et al.
Veröffentlicht: (2024)
von: Olmo, Jeffrey, et al.
Veröffentlicht: (2024)
Learning Multi-Level Features with Matryoshka Sparse Autoencoders
von: Bussmann, Bart, et al.
Veröffentlicht: (2025)
von: Bussmann, Bart, et al.
Veröffentlicht: (2025)
Learning Smooth Distance Functions via Queries
von: Kumar, Akash, et al.
Veröffentlicht: (2024)
von: Kumar, Akash, et al.
Veröffentlicht: (2024)
A Gap Between Decision Trees and Neural Networks
von: Kumar, Akash
Veröffentlicht: (2026)
von: Kumar, Akash
Veröffentlicht: (2026)
Discrete Dictionary-based Decomposition Layer for Structured Representation Learning
von: Park, Taewon, et al.
Veröffentlicht: (2024)
von: Park, Taewon, et al.
Veröffentlicht: (2024)
TraCeS: Trajectory Based Credit Assignment From Sparse Safety Feedback
von: Low, Siow Meng, et al.
Veröffentlicht: (2025)
von: Low, Siow Meng, et al.
Veröffentlicht: (2025)
Learning to Detect Critical Nodes in Sparse Graphs via Feature Importance Awareness
von: Tan, Xuwei, et al.
Veröffentlicht: (2021)
von: Tan, Xuwei, et al.
Veröffentlicht: (2021)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
von: Arnob, Samin Yeasar, et al.
Veröffentlicht: (2025)
von: Arnob, Samin Yeasar, et al.
Veröffentlicht: (2025)
Personalized Federated Dictionary Learning for Modeling Heterogeneity in Multi-site fMRI Data
von: Zhang, Yipu, et al.
Veröffentlicht: (2025)
von: Zhang, Yipu, et al.
Veröffentlicht: (2025)
PathletRL++: Optimizing Trajectory Pathlet Extraction and Dictionary Formation via Reinforcement Learning
von: Alix, Gian, et al.
Veröffentlicht: (2024)
von: Alix, Gian, et al.
Veröffentlicht: (2024)
Discourse-Aware In-Context Learning for Temporal Expression Normalization
von: Gautam, Akash Kumar, et al.
Veröffentlicht: (2024)
von: Gautam, Akash Kumar, et al.
Veröffentlicht: (2024)
Orion-MSP: Multi-Scale Sparse Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
Adaptive Multi-Fidelity Reinforcement Learning for Variance Reduction in Engineering Design Optimization
von: Agrawal, Akash, et al.
Veröffentlicht: (2025)
von: Agrawal, Akash, et al.
Veröffentlicht: (2025)
Dy-mer: An Explainable DNA Sequence Representation Scheme using Dictionary Learning
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2024)
Adaptive Learning of Design Strategies over Non-Hierarchical Multi-Fidelity Models via Policy Alignment
von: Agrawal, Akash, et al.
Veröffentlicht: (2024)
von: Agrawal, Akash, et al.
Veröffentlicht: (2024)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
von: Ayonrinde, Kola
Veröffentlicht: (2024)
von: Ayonrinde, Kola
Veröffentlicht: (2024)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
von: Hartman, Max, et al.
Veröffentlicht: (2025)
von: Hartman, Max, et al.
Veröffentlicht: (2025)
Contrastive Preference Learning: Learning from Human Feedback without RL
von: Hejna, Joey, et al.
Veröffentlicht: (2023)
von: Hejna, Joey, et al.
Veröffentlicht: (2023)
RLAF: Reinforcement Learning from Automaton Feedback
von: Alinejad, Mahyar, et al.
Veröffentlicht: (2025)
von: Alinejad, Mahyar, et al.
Veröffentlicht: (2025)
Reward Learning from Multiple Feedback Types
von: Metz, Yannick, et al.
Veröffentlicht: (2025)
von: Metz, Yannick, et al.
Veröffentlicht: (2025)
Understanding the Learning Dynamics of Alignment with Human Feedback
von: Im, Shawn, et al.
Veröffentlicht: (2024)
von: Im, Shawn, et al.
Veröffentlicht: (2024)
Swap-guided Preference Learning for Personalized Reinforcement Learning from Human Feedback
von: Kim, Gihoon, et al.
Veröffentlicht: (2026)
von: Kim, Gihoon, et al.
Veröffentlicht: (2026)
Data Whitening Improves Sparse Autoencoder Learning
von: Saraswatula, Ashwin, et al.
Veröffentlicht: (2025)
von: Saraswatula, Ashwin, et al.
Veröffentlicht: (2025)
LInK: Learning Joint Representations of Design and Performance Spaces through Contrastive Learning for Mechanism Synthesis
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2024)
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2024)
Sparse Adapter Fusion for Continual Learning in NLP
von: Zeng, Min, et al.
Veröffentlicht: (2026)
von: Zeng, Min, et al.
Veröffentlicht: (2026)
Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
von: Karvonen, Adam, et al.
Veröffentlicht: (2024)
von: Karvonen, Adam, et al.
Veröffentlicht: (2024)
The Sample Complexity of Multiclass and Sparse Contextual Bandits
von: Erez, Liad, et al.
Veröffentlicht: (2026)
von: Erez, Liad, et al.
Veröffentlicht: (2026)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
A Gap Between the Gaussian RKHS and Neural Networks: An Infinite-Center Asymptotic Analysis
von: Kumar, Akash, et al.
Veröffentlicht: (2025)
von: Kumar, Akash, et al.
Veröffentlicht: (2025)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
Negative Feedback System as Optimizer for Machine Learning Systems
von: Hasan, Md Munir, et al.
Veröffentlicht: (2021)
von: Hasan, Md Munir, et al.
Veröffentlicht: (2021)
Nonlinearity, Feedback and Uniform Consistency in Causal Structural Learning
von: Wang, Shuyan
Veröffentlicht: (2023)
von: Wang, Shuyan
Veröffentlicht: (2023)
Corruption Robust Offline Reinforcement Learning with Human Feedback
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
Privacy Preserving Reinforcement Learning with One-Sided Feedback
von: Cong, Lin William, et al.
Veröffentlicht: (2026)
von: Cong, Lin William, et al.
Veröffentlicht: (2026)
CANDERE-COACH: Reinforcement Learning from Noisy Feedback
von: Li, Yuxuan, et al.
Veröffentlicht: (2024)
von: Li, Yuxuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Superpose Task-specific Features for Model Merging
von: Qiu, Haiquan, et al.
Veröffentlicht: (2025) -
Identifying Functionally Important Features with End-to-End Sparse Dictionary Learning
von: Braun, Dan, et al.
Veröffentlicht: (2024) -
Improving Dictionary Learning with Gated Sparse Autoencoders
von: Rajamanoharan, Senthooran, et al.
Veröffentlicht: (2024) -
A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima
von: Tang, Yiming, et al.
Veröffentlicht: (2025) -
Impact of Label Noise on Learning Complex Features
von: Vashisht, Rahul, et al.
Veröffentlicht: (2024)