Bridging Explainability and Embeddings: BEE Aware of Spuriousness
Fuente:
arXiv
Saved in:
| Main Authors: | Păduraru, Cristian Daniel, Bărbălau, Antonio, Filipescu, Radu, Nicolicioiu, Andrei Liviu, Burceanu, Elena |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Sparse Autoencoders: Select-and-Project for Fairness and Control from Encoder Features Alone
by: Bărbălau, Antonio, et al.
Published: (2025)
by: Bărbălau, Antonio, et al.
Published: (2025)
Not All Splits Are Equal: Rethinking Attribute Generalization Across Unrelated Categories
by: Fircă, Liviu Nicolae, et al.
Published: (2025)
by: Fircă, Liviu Nicolae, et al.
Published: (2025)
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
by: Dragomir, Alexandra, et al.
Published: (2026)
by: Dragomir, Alexandra, et al.
Published: (2026)
Robust Novelty Detection through Style-Conscious Feature Ranking
by: Smeu, Stefan, et al.
Published: (2023)
by: Smeu, Stefan, et al.
Published: (2023)
Curriculum-enhanced GroupDRO: Challenging the Norm of Avoiding Curriculum Learning in Subpopulation Shift Setups
by: Barbalau, Antonio
Published: (2024)
by: Barbalau, Antonio
Published: (2024)
ChronoGraph: A Real-World Graph-Based Multivariate Time Series Dataset
by: Lutu, Adrian Catalin, et al.
Published: (2025)
by: Lutu, Adrian Catalin, et al.
Published: (2025)
Spurious Correlation-Aware Embedding Regularization for Worst-Group Robustness
by: Park, Subeen, et al.
Published: (2025)
by: Park, Subeen, et al.
Published: (2025)
idSCD: Identifying Training Datasets through Semantic Correlation Descriptors
by: Gobeaja, Andrada, et al.
Published: (2026)
by: Gobeaja, Andrada, et al.
Published: (2026)
Exploring the potential of prototype-based soft-labels data distillation for imbalanced data classification
by: Rosu, Radu-Andrei, et al.
Published: (2024)
by: Rosu, Radu-Andrei, et al.
Published: (2024)
Out of Spuriousity: Improving Robustness to Spurious Correlations without Group Annotations
by: Le, Phuong Quynh, et al.
Published: (2024)
by: Le, Phuong Quynh, et al.
Published: (2024)
Neural Redshift: Random Networks are not Random Functions
by: Teney, Damien, et al.
Published: (2024)
by: Teney, Damien, et al.
Published: (2024)
LAVA: Explainability for Unsupervised Latent Embeddings
by: Stresec, Ivan, et al.
Published: (2025)
by: Stresec, Ivan, et al.
Published: (2025)
Revisiting Spurious Correlation in Domain Generalization
by: Qin, Bin, et al.
Published: (2024)
by: Qin, Bin, et al.
Published: (2024)
Severing Spurious Correlations with Data Pruning
by: Mulchandani, Varun, et al.
Published: (2025)
by: Mulchandani, Varun, et al.
Published: (2025)
Mitigating Spurious Correlations via Disagreement Probability
by: Han, Hyeonggeun, et al.
Published: (2024)
by: Han, Hyeonggeun, et al.
Published: (2024)
Shortcut to Nowhere: Demystifying Deep Spurious Regression
by: Xu, Guanrong, et al.
Published: (2026)
by: Xu, Guanrong, et al.
Published: (2026)
Spurious Rewards: Rethinking Training Signals in RLVR
by: Shao, Rulin, et al.
Published: (2025)
by: Shao, Rulin, et al.
Published: (2025)
BEE: Metric-Adapted Explanations via Baseline Exploration-Exploitation
by: Barkan, Oren, et al.
Published: (2024)
by: Barkan, Oren, et al.
Published: (2024)
Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation
by: Airinei, Daniel, et al.
Published: (2025)
by: Airinei, Daniel, et al.
Published: (2025)
System-Embedded Diffusion Bridge Models
by: Sobieski, Bartlomiej, et al.
Published: (2025)
by: Sobieski, Bartlomiej, et al.
Published: (2025)
ScoresActivation: A New Activation Function for Model Agnostic Global Explainability by Design
by: Covaci, Emanuel, et al.
Published: (2025)
by: Covaci, Emanuel, et al.
Published: (2025)
A Trace-Based Assurance Framework for Agentic AI Orchestration: Contracts, Testing, and Governance
by: Paduraru, Ciprian, et al.
Published: (2026)
by: Paduraru, Ciprian, et al.
Published: (2026)
SES: Bridging the Gap Between Explainability and Prediction of Graph Neural Networks
by: Huang, Zhenhua, et al.
Published: (2024)
by: Huang, Zhenhua, et al.
Published: (2024)
Combating Spurious Correlations in Graph Interpretability via Self-Reflection
by: Cai, Kecheng, et al.
Published: (2026)
by: Cai, Kecheng, et al.
Published: (2026)
Explainable Graph Spectral Clustering For GloVe-like Text Embeddings
by: Kłopotek, Mieczysław A., et al.
Published: (2025)
by: Kłopotek, Mieczysław A., et al.
Published: (2025)
Embedding-Aware Feature Discovery: Bridging Latent Representations and Interpretable Features in Event Sequences
by: Sakhno, Artem, et al.
Published: (2026)
by: Sakhno, Artem, et al.
Published: (2026)
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
by: Mikriukov, Georgii, et al.
Published: (2026)
by: Mikriukov, Georgii, et al.
Published: (2026)
Correlation-Aware Feature Attribution Based Explainable AI
by: Sengupta, Poushali, et al.
Published: (2025)
by: Sengupta, Poushali, et al.
Published: (2025)
The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory
by: Tang, Luoxi, et al.
Published: (2026)
by: Tang, Luoxi, et al.
Published: (2026)
Evaluating Explainability in Machine Learning Predictions through Explainer-Agnostic Metrics
by: Munoz, Cristian, et al.
Published: (2023)
by: Munoz, Cristian, et al.
Published: (2023)
The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter
by: Bell, Samuel J., et al.
Published: (2024)
by: Bell, Samuel J., et al.
Published: (2024)
Navigating Shortcuts, Spurious Correlations, and Confounders: From Origins via Detection to Mitigation
by: Steinmann, David, et al.
Published: (2024)
by: Steinmann, David, et al.
Published: (2024)
Information as Structural Alignment: A Dynamical Theory of Continual Learning
by: Negulescu, Radu
Published: (2026)
by: Negulescu, Radu
Published: (2026)
Robust Uncertainty Quantification Using Conformalised Monte Carlo Prediction
by: Bethell, Daniel, et al.
Published: (2023)
by: Bethell, Daniel, et al.
Published: (2023)
Not Only the Last-Layer Features for Spurious Correlations: All Layer Deep Feature Reweighting
by: Hameed, Humza Wajid, et al.
Published: (2024)
by: Hameed, Humza Wajid, et al.
Published: (2024)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026)
by: Moya, Christian, et al.
Published: (2026)
SCL-GNN: Towards Generalizable Graph Neural Networks via Spurious Correlation Learning
by: Zhang, Yuxiang, et al.
Published: (2026)
by: Zhang, Yuxiang, et al.
Published: (2026)
Reveal-to-Revise: Explainable Bias-Aware Generative Modeling with Multimodal Attention
by: Mohammad, Noor Islam S., et al.
Published: (2025)
by: Mohammad, Noor Islam S., et al.
Published: (2025)
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant
by: Nicolicioiu, Armand, et al.
Published: (2024)
by: Nicolicioiu, Armand, et al.
Published: (2024)
Task-Informed Anti-Curriculum by Masking Improves Downstream Performance on Text
by: Jarca, Andrei, et al.
Published: (2025)
by: Jarca, Andrei, et al.
Published: (2025)
Similar Items
-
Rethinking Sparse Autoencoders: Select-and-Project for Fairness and Control from Encoder Features Alone
by: Bărbălau, Antonio, et al.
Published: (2025) -
Not All Splits Are Equal: Rethinking Attribute Generalization Across Unrelated Categories
by: Fircă, Liviu Nicolae, et al.
Published: (2025) -
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
by: Dragomir, Alexandra, et al.
Published: (2026) -
Robust Novelty Detection through Style-Conscious Feature Ranking
by: Smeu, Stefan, et al.
Published: (2023) -
Curriculum-enhanced GroupDRO: Challenging the Norm of Avoiding Curriculum Learning in Subpopulation Shift Setups
by: Barbalau, Antonio
Published: (2024)