Tackling Polysemanticity with Neuron Embeddings
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Foote, Alex |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling Polysemantic Neurons with a Null-Calibrated Polysemanticity Index and Causal Patch Interventions
von: Gupta, Manan, et al.
Veröffentlicht: (2025)
von: Gupta, Manan, et al.
Veröffentlicht: (2025)
NeuronScope: A Multi-Agent Framework for Explaining Polysemantic Neurons in Language Models
von: Liu, Weiqi, et al.
Veröffentlicht: (2026)
von: Liu, Weiqi, et al.
Veröffentlicht: (2026)
PURE: Turning Polysemantic Neurons Into Pure Features by Identifying Relevant Circuits
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2024)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2024)
The Rate-Distortion-Polysemanticity Tradeoff in SAEs
von: Mencattini, Tommaso, et al.
Veröffentlicht: (2026)
von: Mencattini, Tommaso, et al.
Veröffentlicht: (2026)
Polysemanticity and Capacity in Neural Networks
von: Scherlis, Adam, et al.
Veröffentlicht: (2022)
von: Scherlis, Adam, et al.
Veröffentlicht: (2022)
Disentangling Polysemantic Channels in Convolutional Neural Networks
von: Hesse, Robin, et al.
Veröffentlicht: (2025)
von: Hesse, Robin, et al.
Veröffentlicht: (2025)
Interpretability Without Tradeoffs: Disentangling Polysemanticity At Equal Predictive Performance
von: Bağcı, Doğukan, et al.
Veröffentlicht: (2026)
von: Bağcı, Doğukan, et al.
Veröffentlicht: (2026)
Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs
von: Ye, Charles, et al.
Veröffentlicht: (2026)
von: Ye, Charles, et al.
Veröffentlicht: (2026)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
von: Kopf, Laura, et al.
Veröffentlicht: (2025)
von: Kopf, Laura, et al.
Veröffentlicht: (2025)
Taming Polysemanticity in LLMs: Provable Feature Recovery via Sparse Autoencoders
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
TRIP: A Nonparametric Test to Diagnose Biased Feature Importance Scores
von: Foote, Aaron, et al.
Veröffentlicht: (2025)
von: Foote, Aaron, et al.
Veröffentlicht: (2025)
What Causes Polysemanticity? An Alternative Origin Story of Mixed Selectivity from Incidental Causes
von: Lecomte, Victor, et al.
Veröffentlicht: (2023)
von: Lecomte, Victor, et al.
Veröffentlicht: (2023)
Tackling air quality with SAPIENS
von: Bona, Marcella, et al.
Veröffentlicht: (2026)
von: Bona, Marcella, et al.
Veröffentlicht: (2026)
Decomposing Attention To Find Context-Sensitive Neurons
von: Gibson, Alex
Veröffentlicht: (2025)
von: Gibson, Alex
Veröffentlicht: (2025)
Tackling GNARLy Problems: Graph Neural Algorithmic Reasoning Reimagined through Reinforcement Learning
von: Schutz, Alex, et al.
Veröffentlicht: (2025)
von: Schutz, Alex, et al.
Veröffentlicht: (2025)
Tackling Noisy Labels with Network Parameter Additive Decomposition
von: Wang, Jingyi, et al.
Veröffentlicht: (2024)
von: Wang, Jingyi, et al.
Veröffentlicht: (2024)
Tackling the Zero-Shot Reinforcement Learning Loss Directly
von: Ollivier, Yann
Veröffentlicht: (2025)
von: Ollivier, Yann
Veröffentlicht: (2025)
Neuronal Fluctuations: Learning Rates vs Participating Neurons
von: Pareek, Darsh, et al.
Veröffentlicht: (2025)
von: Pareek, Darsh, et al.
Veröffentlicht: (2025)
Tackling Data Heterogeneity in Federated Learning via Loss Decomposition
von: Zeng, Shuang, et al.
Veröffentlicht: (2024)
von: Zeng, Shuang, et al.
Veröffentlicht: (2024)
Tackling Dimensional Collapse toward Comprehensive Universal Domain Adaptation
von: Fang, Hung-Chieh, et al.
Veröffentlicht: (2024)
von: Fang, Hung-Chieh, et al.
Veröffentlicht: (2024)
STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning
von: Chen, Zekai, et al.
Veröffentlicht: (2026)
von: Chen, Zekai, et al.
Veröffentlicht: (2026)
Tackling Fake Forgetting through Uncertainty Quantification
von: Shi, Yingdan, et al.
Veröffentlicht: (2025)
von: Shi, Yingdan, et al.
Veröffentlicht: (2025)
SPARC: Spectral Architectures Tackling the Cold-Start Problem in Graph Learning
von: Jacobs, Yahel, et al.
Veröffentlicht: (2024)
von: Jacobs, Yahel, et al.
Veröffentlicht: (2024)
Invertible Fourier Neural Operators for Tackling Both Forward and Inverse Problems
von: Long, Da, et al.
Veröffentlicht: (2024)
von: Long, Da, et al.
Veröffentlicht: (2024)
Tackling Time-Series Forecasting Generalization via Mitigating Concept Drift
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2025)
Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
SafeNeuron: Neuron-Level Safety Alignment for Large Language Models
von: Wang, Zhaoxin, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoxin, et al.
Veröffentlicht: (2026)
Variational Distributional Neuron
von: Ruffenach, Yves
Veröffentlicht: (2026)
von: Ruffenach, Yves
Veröffentlicht: (2026)
Expand Neurons, Not Parameters
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
Threshold Neuron: A Brain-inspired Artificial Neuron for Efficient On-device Inference
von: Zheng, Zihao, et al.
Veröffentlicht: (2024)
von: Zheng, Zihao, et al.
Veröffentlicht: (2024)
Improving Model Fusion by Training-time Neuron Alignment with Fixed Neuron Anchors
von: Li, Zexi, et al.
Veröffentlicht: (2024)
von: Li, Zexi, et al.
Veröffentlicht: (2024)
Tackling Data Heterogeneity in Federated Learning through Knowledge Distillation with Inequitable Aggregation
von: Ma, Xing
Veröffentlicht: (2025)
von: Ma, Xing
Veröffentlicht: (2025)
Tackling prediction tasks in relational databases with LLMs
von: Wydmuch, Marek, et al.
Veröffentlicht: (2024)
von: Wydmuch, Marek, et al.
Veröffentlicht: (2024)
Simple Linear Neuron Boosting
von: Munoz, Daniel
Veröffentlicht: (2025)
von: Munoz, Daniel
Veröffentlicht: (2025)
Tackling Feature-Classifier Mismatch in Federated Learning via Prompt-Driven Feature Transformation
von: Wu, Xinghao, et al.
Veröffentlicht: (2024)
von: Wu, Xinghao, et al.
Veröffentlicht: (2024)
FedKL: Tackling Data Heterogeneity in Federated Reinforcement Learning by Penalizing KL Divergence
von: Xie, Zhijie, et al.
Veröffentlicht: (2022)
von: Xie, Zhijie, et al.
Veröffentlicht: (2022)
Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space
von: He, Xin, et al.
Veröffentlicht: (2025)
von: He, Xin, et al.
Veröffentlicht: (2025)
Model-Robust and Adaptive-Optimal Transfer Learning for Tackling Concept Shifts in Nonparametric Regression
von: Lin, Haotian, et al.
Veröffentlicht: (2025)
von: Lin, Haotian, et al.
Veröffentlicht: (2025)
FedRC: Tackling Diverse Distribution Shifts Challenge in Federated Learning by Robust Clustering
von: Guo, Yongxin, et al.
Veröffentlicht: (2023)
von: Guo, Yongxin, et al.
Veröffentlicht: (2023)
Non-exchangeable Conformal Prediction with Optimal Transport: Tackling Distribution Shifts with Unlabeled Data
von: Correia, Alvaro H. C., et al.
Veröffentlicht: (2025)
von: Correia, Alvaro H. C., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Disentangling Polysemantic Neurons with a Null-Calibrated Polysemanticity Index and Causal Patch Interventions
von: Gupta, Manan, et al.
Veröffentlicht: (2025) -
NeuronScope: A Multi-Agent Framework for Explaining Polysemantic Neurons in Language Models
von: Liu, Weiqi, et al.
Veröffentlicht: (2026) -
PURE: Turning Polysemantic Neurons Into Pure Features by Identifying Relevant Circuits
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2024) -
The Rate-Distortion-Polysemanticity Tradeoff in SAEs
von: Mencattini, Tommaso, et al.
Veröffentlicht: (2026) -
Polysemanticity and Capacity in Neural Networks
von: Scherlis, Adam, et al.
Veröffentlicht: (2022)