Concept Gradient: Concept-based Interpretation Without Linear Assumption
Fuente:
arXiv
Saved in:
| Main Authors: | Bai, Andrew, Yeh, Chih-Kuan, Ravikumar, Pradeep, Lin, Neil Y. C., Hsieh, Cho-Jui |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Efficient Rehearsal Scheme for Catastrophic Forgetting Mitigation during Multi-stage Fine-tuning
by: Bai, Andrew, et al.
Published: (2024)
by: Bai, Andrew, et al.
Published: (2024)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
by: Bai, Andrew, et al.
Published: (2025)
by: Bai, Andrew, et al.
Published: (2025)
CLUE: Concept-Level Uncertainty Estimation for Large Language Models
by: Wang, Yu-Hsiang, et al.
Published: (2024)
by: Wang, Yu-Hsiang, et al.
Published: (2024)
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
by: Rajendran, Goutham, et al.
Published: (2024)
by: Rajendran, Goutham, et al.
Published: (2024)
A Unifying Framework for Unsupervised Concept Extraction
by: Squires, Chandler, et al.
Published: (2026)
by: Squires, Chandler, et al.
Published: (2026)
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
by: Hong, Yunqi, et al.
Published: (2025)
by: Hong, Yunqi, et al.
Published: (2025)
Data Attribution for Diffusion Models: Timestep-induced Bias in Influence Estimation
by: Xie, Tong, et al.
Published: (2024)
by: Xie, Tong, et al.
Published: (2024)
Online Continuous Hyperparameter Optimization for Generalized Linear Contextual Bandits
by: Kang, Yue, et al.
Published: (2023)
by: Kang, Yue, et al.
Published: (2023)
On the Loss of Context-awareness in General Instruction Fine-tuning
by: Wang, Yihan, et al.
Published: (2024)
by: Wang, Yihan, et al.
Published: (2024)
Embedding Space Selection for Detecting Memorization and Fingerprinting in Generative Models
by: He, Jack, et al.
Published: (2024)
by: He, Jack, et al.
Published: (2024)
Concept Bottleneck Models Without Predefined Concepts
by: Schrodi, Simon, et al.
Published: (2024)
by: Schrodi, Simon, et al.
Published: (2024)
Hierarchical Concept-based Interpretable Models
by: Hill, Oscar, et al.
Published: (2026)
by: Hill, Oscar, et al.
Published: (2026)
Tree of Concepts: Interpretable Continual Learners in Non-Stationary Clinical Domains
by: Cho, Dongkyu, et al.
Published: (2026)
by: Cho, Dongkyu, et al.
Published: (2026)
Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry
by: Hindupur, Sai Sumedh R., et al.
Published: (2025)
by: Hindupur, Sai Sumedh R., et al.
Published: (2025)
CAT: Interpretable Concept-based Taylor Additive Models
by: Duong, Viet, et al.
Published: (2024)
by: Duong, Viet, et al.
Published: (2024)
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
by: Lin, Shen, et al.
Published: (2026)
by: Lin, Shen, et al.
Published: (2026)
Efficient Frameworks for Generalized Low-Rank Matrix Bandit Problems
by: Kang, Yue, et al.
Published: (2024)
by: Kang, Yue, et al.
Published: (2024)
Low-rank Matrix Bandits with Heavy-tailed Rewards
by: Kang, Yue, et al.
Published: (2024)
by: Kang, Yue, et al.
Published: (2024)
On the Origins of Linear Representations in Large Language Models
by: Jiang, Yibo, et al.
Published: (2024)
by: Jiang, Yibo, et al.
Published: (2024)
Unlearning Concepts in Diffusion Model via Concept Domain Correction and Concept Preserving Gradient
by: Wu, Yongliang, et al.
Published: (2024)
by: Wu, Yongliang, et al.
Published: (2024)
Identifying General Mechanism Shifts in Linear Causal Representations
by: Chen, Tianyu, et al.
Published: (2024)
by: Chen, Tianyu, et al.
Published: (2024)
Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE)
by: Bhalla, Usha, et al.
Published: (2024)
by: Bhalla, Usha, et al.
Published: (2024)
Linear Adversarial Concept Erasure
by: Ravfogel, Shauli, et al.
Published: (2022)
by: Ravfogel, Shauli, et al.
Published: (2022)
Explaining Concept Shift with Interpretable Feature Attribution
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
Interpretable Reward Modeling with Active Concept Bottlenecks
by: Laguna, Sonia, et al.
Published: (2025)
by: Laguna, Sonia, et al.
Published: (2025)
ConceptCaps: a Distilled Concept Dataset for Interpretability in Music Models
by: Sienkiewicz, Bruno, et al.
Published: (2026)
by: Sienkiewicz, Bruno, et al.
Published: (2026)
Interpretable Concept-Based Memory Reasoning
by: Debot, David, et al.
Published: (2024)
by: Debot, David, et al.
Published: (2024)
Leakage and Interpretability in Concept-Based Models
by: Parisini, Enrico, et al.
Published: (2025)
by: Parisini, Enrico, et al.
Published: (2025)
Interpretable Prognostics with Concept Bottleneck Models
by: Forest, Florent, et al.
Published: (2024)
by: Forest, Florent, et al.
Published: (2024)
What Does Preference Learning Recover from Pairwise Comparison Data?
by: Pukdee, Rattana, et al.
Published: (2026)
by: Pukdee, Rattana, et al.
Published: (2026)
Provably Robust Training of Quantum Circuit Classifiers Against Parameter Noise
by: Tecot, Lucas, et al.
Published: (2025)
by: Tecot, Lucas, et al.
Published: (2025)
From Segments to Concepts: Interpretable Image Classification via Concept-Guided Segmentation
by: Eisenberg, Ran, et al.
Published: (2025)
by: Eisenberg, Ran, et al.
Published: (2025)
Understanding Augmentation-based Self-Supervised Representation Learning via RKHS Approximation and Regression
by: Zhai, Runtian, et al.
Published: (2023)
by: Zhai, Runtian, et al.
Published: (2023)
Interpretability for Multimodal Emotion Recognition using Concept Activation Vectors
by: Asokan, Ashish Ramayee, et al.
Published: (2022)
by: Asokan, Ashish Ramayee, et al.
Published: (2022)
Exploiting Interpretable Capabilities with Concept-Enhanced Diffusion and Prototype Networks
by: Carballo-Castro, Alba, et al.
Published: (2024)
by: Carballo-Castro, Alba, et al.
Published: (2024)
Federated Concept-Based Models: Interpretable models with distributed supervision
by: Fenoglio, Dario, et al.
Published: (2026)
by: Fenoglio, Dario, et al.
Published: (2026)
Interpretable Neural-Symbolic Concept Reasoning
by: Barbiero, Pietro, et al.
Published: (2023)
by: Barbiero, Pietro, et al.
Published: (2023)
Preserving Task-Relevant Information Under Linear Concept Removal
by: Holstege, Floris, et al.
Published: (2025)
by: Holstege, Floris, et al.
Published: (2025)
Interpretable Concept Bottlenecks to Align Reinforcement Learning Agents
by: Delfosse, Quentin, et al.
Published: (2024)
by: Delfosse, Quentin, et al.
Published: (2024)
Self-explaining Neural Network with Concept-based Explanations for ICU Mortality Prediction
by: Kumar, Sayantan, et al.
Published: (2021)
by: Kumar, Sayantan, et al.
Published: (2021)
Similar Items
-
An Efficient Rehearsal Scheme for Catastrophic Forgetting Mitigation during Multi-stage Fine-tuning
by: Bai, Andrew, et al.
Published: (2024) -
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
by: Bai, Andrew, et al.
Published: (2025) -
CLUE: Concept-Level Uncertainty Estimation for Large Language Models
by: Wang, Yu-Hsiang, et al.
Published: (2024) -
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
by: Rajendran, Goutham, et al.
Published: (2024) -
A Unifying Framework for Unsupervised Concept Extraction
by: Squires, Chandler, et al.
Published: (2026)