Stochastic Parameter Decomposition
Fuente:
arXiv
Saved in:
| Main Authors: | Bushnaq, Lucius, Braun, Dan, Sharkey, Lee |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Identifying Sparsely Active Circuits Through Local Loss Landscape Decomposition
by: Chrisman, Brianna, et al.
Published: (2025)
by: Chrisman, Brianna, et al.
Published: (2025)
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition
by: Braun, Dan, et al.
Published: (2025)
by: Braun, Dan, et al.
Published: (2025)
Identifying Functionally Important Features with End-to-End Sparse Dictionary Learning
by: Braun, Dan, et al.
Published: (2024)
by: Braun, Dan, et al.
Published: (2024)
Interpretability as Compression: Reconsidering SAE Explanations of Neural Activations with MDL-SAEs
by: Ayonrinde, Kola, et al.
Published: (2024)
by: Ayonrinde, Kola, et al.
Published: (2024)
Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs
by: Ye, Charles, et al.
Published: (2026)
by: Ye, Charles, et al.
Published: (2026)
Deterministic Decomposition of Stochastic Generative Dynamics
by: Song, Xingyu, et al.
Published: (2026)
by: Song, Xingyu, et al.
Published: (2026)
Parameter Expanded Stochastic Gradient Markov Chain Monte Carlo
by: Kim, Hyunsu, et al.
Published: (2025)
by: Kim, Hyunsu, et al.
Published: (2025)
Clarifying Shampoo: Adapting Spectral Descent to Stochasticity and the Parameter Trajectory
by: Eschenhagen, Runa, et al.
Published: (2026)
by: Eschenhagen, Runa, et al.
Published: (2026)
From Memorization to Reasoning in the Spectrum of Loss Curvature
by: Merullo, Jack, et al.
Published: (2025)
by: Merullo, Jack, et al.
Published: (2025)
Transformation Categorization Based on Group Decomposition Theory Using Parameter Division
by: Komatsu, Takayuki, et al.
Published: (2026)
by: Komatsu, Takayuki, et al.
Published: (2026)
Using Degeneracy in the Loss Landscape for Mechanistic Interpretability
by: Bushnaq, Lucius, et al.
Published: (2024)
by: Bushnaq, Lucius, et al.
Published: (2024)
Language Models as Causal Effect Generators
by: Bynum, Lucius E. J., et al.
Published: (2024)
by: Bynum, Lucius E. J., et al.
Published: (2024)
A New Paradigm for Counterfactual Reasoning in Fairness and Recourse
by: Bynum, Lucius E. J., et al.
Published: (2024)
by: Bynum, Lucius E. J., et al.
Published: (2024)
Efficient CNN-LSTM based Parameter Estimation of Levy Driven Stochastic Differential Equations
by: Li, Shuaiyu, et al.
Published: (2024)
by: Li, Shuaiyu, et al.
Published: (2024)
Attention-based Iterative Decomposition for Tensor Product Representation
by: Park, Taewon, et al.
Published: (2024)
by: Park, Taewon, et al.
Published: (2024)
Sparse Autoencoders Do Not Find Canonical Units of Analysis
by: Leask, Patrick, et al.
Published: (2025)
by: Leask, Patrick, et al.
Published: (2025)
CorDA: Context-Oriented Decomposition Adaptation of Large Language Models for Task-Aware Parameter-Efficient Fine-tuning
by: Yang, Yibo, et al.
Published: (2024)
by: Yang, Yibo, et al.
Published: (2024)
Discrete Dictionary-based Decomposition Layer for Structured Representation Learning
by: Park, Taewon, et al.
Published: (2024)
by: Park, Taewon, et al.
Published: (2024)
Controllable Probabilistic Forecasting with Stochastic Decomposition Layers
by: Schreck, John S., et al.
Published: (2025)
by: Schreck, John S., et al.
Published: (2025)
Q-function Decomposition with Intervention Semantics with Factored Action Spaces
by: Lee, Junkyu, et al.
Published: (2025)
by: Lee, Junkyu, et al.
Published: (2025)
Catalyst: a Novel Regularizer for Structured Pruning with Auxiliary Extension of Parameter Space
by: Jung, Jaeheun, et al.
Published: (2025)
by: Jung, Jaeheun, et al.
Published: (2025)
Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines
by: Nilayam, Prashanti, et al.
Published: (2026)
by: Nilayam, Prashanti, et al.
Published: (2026)
MUXQ: Mixed-to-Uniform Precision MatriX Quantization via Low-Rank Outlier Decomposition
by: Lee, Seoungsub, et al.
Published: (2026)
by: Lee, Seoungsub, et al.
Published: (2026)
On Discovery of Local Independence over Continuous Variables via Neural Contextual Decomposition
by: Hwang, Inwoo, et al.
Published: (2024)
by: Hwang, Inwoo, et al.
Published: (2024)
Parameters vs FLOPs: Scaling Laws for Optimal Sparsity for Mixture-of-Experts Language Models
by: Abnar, Samira, et al.
Published: (2025)
by: Abnar, Samira, et al.
Published: (2025)
The Local Interaction Basis: Identifying Computationally-Relevant and Sparsely Interacting Features in Neural Networks
by: Bushnaq, Lucius, et al.
Published: (2024)
by: Bushnaq, Lucius, et al.
Published: (2024)
Unveiling Options with Neural Decomposition
by: Alikhasi, Mahdi, et al.
Published: (2024)
by: Alikhasi, Mahdi, et al.
Published: (2024)
The Unseen Frontier: Pushing the Limits of LLM Sparsity with Surrogate-Free ADMM
by: Lee, Kwanhee, et al.
Published: (2025)
by: Lee, Kwanhee, et al.
Published: (2025)
CleaR: Towards Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Label Learning
by: Kim, Yeachan, et al.
Published: (2024)
by: Kim, Yeachan, et al.
Published: (2024)
Sparse Decomposition of Graph Neural Networks
by: Hu, Yaochen, et al.
Published: (2024)
by: Hu, Yaochen, et al.
Published: (2024)
Coefficient Decomposition for Spectral Graph Convolution
by: Huang, Feng, et al.
Published: (2024)
by: Huang, Feng, et al.
Published: (2024)
A Unified Frequency Domain Decomposition Framework for Interpretable and Robust Time Series Forecasting
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
Explaining Datasets in Words: Statistical Models with Natural Language Parameters
by: Zhong, Ruiqi, et al.
Published: (2024)
by: Zhong, Ruiqi, et al.
Published: (2024)
Hierarchical Sparse Circuit Extraction from Billion-Parameter Language Models through Scalable Attribution Graph Decomposition
by: Uddin, Mohammed Mudassir, et al.
Published: (2026)
by: Uddin, Mohammed Mudassir, et al.
Published: (2026)
Stochastic activations
by: Lomeli, Maria, et al.
Published: (2025)
by: Lomeli, Maria, et al.
Published: (2025)
See Further for Parameter Efficient Fine-tuning by Standing on the Shoulders of Decomposition
by: Si, Chongjie, et al.
Published: (2024)
by: Si, Chongjie, et al.
Published: (2024)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
by: Nguyen-Hien, T. Duy, et al.
Published: (2025)
by: Nguyen-Hien, T. Duy, et al.
Published: (2025)
Parameter-Efficient Fine-Tuning for Foundation Models
by: Zhang, Dan, et al.
Published: (2025)
by: Zhang, Dan, et al.
Published: (2025)
Bilinear Convolution Decomposition for Causal RL Interpretability
by: Oozeer, Narmeen, et al.
Published: (2024)
by: Oozeer, Narmeen, et al.
Published: (2024)
Deep Classifier Mimicry without Data Access
by: Braun, Steven, et al.
Published: (2023)
by: Braun, Steven, et al.
Published: (2023)
Similar Items
-
Identifying Sparsely Active Circuits Through Local Loss Landscape Decomposition
by: Chrisman, Brianna, et al.
Published: (2025) -
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition
by: Braun, Dan, et al.
Published: (2025) -
Identifying Functionally Important Features with End-to-End Sparse Dictionary Learning
by: Braun, Dan, et al.
Published: (2024) -
Interpretability as Compression: Reconsidering SAE Explanations of Neural Activations with MDL-SAEs
by: Ayonrinde, Kola, et al.
Published: (2024) -
Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs
by: Ye, Charles, et al.
Published: (2026)