Learning to Receive Help: Intervention-Aware Concept Embedding Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zarlenga, Mateo Espinosa, Collins, Katherine M., Dvijotham, Krishnamurthy, Weller, Adrian, Shams, Zohreh, Jamnik, Mateja |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Avoiding Leakage Poisoning: Concept Interventions Under Distribution Shifts
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
Hierarchical Concept-based Interpretable Models
by: Hill, Oscar, et al.
Published: (2026)
by: Hill, Oscar, et al.
Published: (2026)
Digging Deeper: Learning Multi-Level Concept Hierarchies
by: Hill, Oscar, et al.
Published: (2026)
by: Hill, Oscar, et al.
Published: (2026)
Do Concept Bottleneck Models Respect Localities?
by: Raman, Naveen, et al.
Published: (2024)
by: Raman, Naveen, et al.
Published: (2024)
Efficient Bias Mitigation Without Privileged Information
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
Understanding Inter-Concept Relationships in Concept-Based Models
by: Raman, Naveen, et al.
Published: (2024)
by: Raman, Naveen, et al.
Published: (2024)
Foundations of Interpretable Models
by: Barbiero, Pietro, et al.
Published: (2025)
by: Barbiero, Pietro, et al.
Published: (2025)
Actionable Interpretability Must Be Defined in Terms of Symmetries
by: Barbiero, Pietro, et al.
Published: (2026)
by: Barbiero, Pietro, et al.
Published: (2026)
Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization
by: Penaloza, Emiliano, et al.
Published: (2025)
by: Penaloza, Emiliano, et al.
Published: (2025)
Interpretable Neural-Symbolic Concept Reasoning
by: Barbiero, Pietro, et al.
Published: (2023)
by: Barbiero, Pietro, et al.
Published: (2023)
Causal Concept Graph Models: Beyond Causal Opacity in Deep Learning
by: Dominici, Gabriele, et al.
Published: (2024)
by: Dominici, Gabriele, et al.
Published: (2024)
Multimodal Lego: Model Merging and Fine-Tuning Across Topologies and Modalities in Biomedicine
by: Hemker, Konstantin, et al.
Published: (2024)
by: Hemker, Konstantin, et al.
Published: (2024)
Estimation of Concept Explanations Should be Uncertainty Aware
by: Piratla, Vihari, et al.
Published: (2023)
by: Piratla, Vihari, et al.
Published: (2023)
LLM Embeddings for Deep Learning on Tabular Data
by: Koloski, Boshko, et al.
Published: (2025)
by: Koloski, Boshko, et al.
Published: (2025)
Don't Lose Focus: Activation Steering via Key-Orthogonal Projections
by: Luo, Haoyan, et al.
Published: (2026)
by: Luo, Haoyan, et al.
Published: (2026)
HEALNet: Multimodal Fusion for Heterogeneous Biomedical Data
by: Hemker, Konstantin, et al.
Published: (2023)
by: Hemker, Konstantin, et al.
Published: (2023)
TabMDA: Tabular Manifold Data Augmentation for Any Classifier using Transformers with In-context Subsetting
by: Margeloiu, Andrei, et al.
Published: (2024)
by: Margeloiu, Andrei, et al.
Published: (2024)
Achieving the Tightest Relaxation of Sigmoids for Formal Verification
by: Chevalier, Samuel, et al.
Published: (2024)
by: Chevalier, Samuel, et al.
Published: (2024)
Mixture of Concept Bottleneck Experts
by: De Santis, Francesco, et al.
Published: (2026)
by: De Santis, Francesco, et al.
Published: (2026)
Sphere Neural-Networks for Rational Reasoning
by: Dong, Tiansi, et al.
Published: (2024)
by: Dong, Tiansi, et al.
Published: (2024)
Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models
by: Kim, Kyuyoung, et al.
Published: (2024)
by: Kim, Kyuyoung, et al.
Published: (2024)
Provably Bounding Neural Network Preimages
by: Kotha, Suhas, et al.
Published: (2023)
by: Kotha, Suhas, et al.
Published: (2023)
Verified Neural Compressed Sensing
by: Bunel, Rudy, et al.
Published: (2024)
by: Bunel, Rudy, et al.
Published: (2024)
An AI Monkey Gets Grapes for Sure -- Sphere Neural Networks for Reliable Decision-Making
by: Dong, Tiansi, et al.
Published: (2026)
by: Dong, Tiansi, et al.
Published: (2026)
Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate Experts
by: Pugnana, Andrea, et al.
Published: (2025)
by: Pugnana, Andrea, et al.
Published: (2025)
Mitigating Shortcut Learning with InterpoLated Learning
by: Korakakis, Michalis, et al.
Published: (2025)
by: Korakakis, Michalis, et al.
Published: (2025)
Correlated Noise Provably Beats Independent Noise for Differentially Private Learning
by: Choquette-Choo, Christopher A., et al.
Published: (2023)
by: Choquette-Choo, Christopher A., et al.
Published: (2023)
ALVIN: Active Learning Via INterpolation
by: Korakakis, Michalis, et al.
Published: (2024)
by: Korakakis, Michalis, et al.
Published: (2024)
PATHS: A Hierarchical Transformer for Efficient Whole Slide Image Analysis
by: Buzzard, Zak, et al.
Published: (2024)
by: Buzzard, Zak, et al.
Published: (2024)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
Neural Interpretable Reasoning
by: Barbiero, Pietro, et al.
Published: (2025)
by: Barbiero, Pietro, et al.
Published: (2025)
Learning Personalized Decision Support Policies
by: Bhatt, Umang, et al.
Published: (2023)
by: Bhatt, Umang, et al.
Published: (2023)
Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions
by: Sastre, Ignacio, et al.
Published: (2026)
by: Sastre, Ignacio, et al.
Published: (2026)
Shaping the Future of Mathematics in the Age of AI
by: Commelin, Johan, et al.
Published: (2026)
by: Commelin, Johan, et al.
Published: (2026)
No, of Course I Can! Deeper Fine-Tuning Attacks That Bypass Token-Level Safety Mechanisms
by: Kazdan, Joshua, et al.
Published: (2025)
by: Kazdan, Joshua, et al.
Published: (2025)
Steering into New Embedding Spaces: Analyzing Cross-Lingual Alignment Induced by Model Interventions in Multilingual Language Models
by: Sundar, Anirudh, et al.
Published: (2025)
by: Sundar, Anirudh, et al.
Published: (2025)
Large Language Models Must Be Taught to Know What They Don't Know
by: Kapoor, Sanyam, et al.
Published: (2024)
by: Kapoor, Sanyam, et al.
Published: (2024)
Modulating Language Model Experiences through Frictions
by: Collins, Katherine M., et al.
Published: (2024)
by: Collins, Katherine M., et al.
Published: (2024)
Understanding Subjectivity through the Lens of Motivational Context in Model-Generated Image Satisfaction
by: Dutta, Senjuti, et al.
Published: (2024)
by: Dutta, Senjuti, et al.
Published: (2024)
CRAFT: Forgetting-Aware Intervention-Based Adaptation for Continual Learning
by: Hossen, Md Anwar, et al.
Published: (2026)
by: Hossen, Md Anwar, et al.
Published: (2026)
Similar Items
-
Avoiding Leakage Poisoning: Concept Interventions Under Distribution Shifts
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025) -
Hierarchical Concept-based Interpretable Models
by: Hill, Oscar, et al.
Published: (2026) -
Digging Deeper: Learning Multi-Level Concept Hierarchies
by: Hill, Oscar, et al.
Published: (2026) -
Do Concept Bottleneck Models Respect Localities?
by: Raman, Naveen, et al.
Published: (2024) -
Efficient Bias Mitigation Without Privileged Information
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)