Hierarchical, Interpretable, Label-Free Concept Bottleneck Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xie, Haodong, Cai, Yujun, Maharjan, Rahul Singh, Wang, Yiwei, Tavella, Federico, Cangelosi, Angelo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
Attributes-aware Visual Emotion Representation Learning
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
Noise-Free Explanation for Driving Action Prediction
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024)
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024)
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
von: Tavella, Federico, et al.
Veröffentlicht: (2025)
von: Tavella, Federico, et al.
Veröffentlicht: (2025)
Representation Understanding via Activation Maximization
von: Zhu, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhu, Hongbo, et al.
Veröffentlicht: (2025)
Cure or Poison? Embedding Instructions Visually Alters Hallucination in Vision-Language Models
von: Wang, Zhaochen, et al.
Veröffentlicht: (2025)
von: Wang, Zhaochen, et al.
Veröffentlicht: (2025)
Concept Complement Bottleneck Model for Interpretable Medical Image Diagnosis
von: Wang, Hongmei, et al.
Veröffentlicht: (2024)
von: Wang, Hongmei, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Concept Bottleneck Models with Enhanced Interpretability
von: Zhang, Haifei, et al.
Veröffentlicht: (2025)
von: Zhang, Haifei, et al.
Veröffentlicht: (2025)
ChainMPQ: Interleaved Text-Image Reasoning Chains for Mitigating Relation Hallucinations
von: Wu, Yike, et al.
Veröffentlicht: (2025)
von: Wu, Yike, et al.
Veröffentlicht: (2025)
Signs of Language: Embodied Sign Language Fingerspelling Acquisition from Demonstrations for Human-Robot Interaction
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
PAS: A Training-Free Stabilizer for Temporal Encoding in Video LLMs
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretations
von: Xu, Xinyue, et al.
Veröffentlicht: (2024)
von: Xu, Xinyue, et al.
Veröffentlicht: (2024)
Semi-supervised Concept Bottleneck Models
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
Editable Concept Bottleneck Models
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
MRFD: Multi-Region Fusion Decoding with Self-Consistency for Mitigating Hallucinations in LVLMs
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
Text Speaks Louder than Vision: ASCII Art Reveals Textual Biases in Vision-Language Models
von: Wang, Zhaochen, et al.
Veröffentlicht: (2025)
von: Wang, Zhaochen, et al.
Veröffentlicht: (2025)
Explainable Visual Anomaly Detection via Concept Bottleneck Models
von: Stropeni, Arianna, et al.
Veröffentlicht: (2025)
von: Stropeni, Arianna, et al.
Veröffentlicht: (2025)
Zero-shot Concept Bottleneck Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
Process-Guided Concept Bottleneck Model
von: Asiyabi, Reza M., et al.
Veröffentlicht: (2026)
von: Asiyabi, Reza M., et al.
Veröffentlicht: (2026)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG
von: Fu, Honghao, et al.
Veröffentlicht: (2026)
von: Fu, Honghao, et al.
Veröffentlicht: (2026)
Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding
von: Wu, Hang, et al.
Veröffentlicht: (2026)
von: Wu, Hang, et al.
Veröffentlicht: (2026)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
CBM-RAG: Demonstrating Enhanced Interpretability in Radiology Report Generation with Multi-Agent RAG and Concept Bottleneck Models
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2025)
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2025)
EQ-CBM: A Probabilistic Concept Bottleneck with Energy-based Models and Quantized Vectors
von: Kim, Sangwon, et al.
Veröffentlicht: (2024)
von: Kim, Sangwon, et al.
Veröffentlicht: (2024)
PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
CLIP-Free, Label Free, Unsupervised Concept Bottleneck Models
von: Sammani, Fawaz, et al.
Veröffentlicht: (2025)
von: Sammani, Fawaz, et al.
Veröffentlicht: (2025)
MVP-CBM:Multi-layer Visual Preference-enhanced Concept Bottleneck Model for Explainable Medical Image Classification
von: Wang, Chunjiang, et al.
Veröffentlicht: (2025)
von: Wang, Chunjiang, et al.
Veröffentlicht: (2025)
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning
von: Wu, Hang, et al.
Veröffentlicht: (2026)
von: Wu, Hang, et al.
Veröffentlicht: (2026)
Adaptive Concept Bottleneck for Foundation Models Under Distribution Shifts
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
WLASL-LEX: a Dataset for Recognising Phonological Properties in American Sign Language
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
CoBELa: Steering Transparent Generation via Concept Bottlenecks on Energy Landscapes
von: Kim, Sangwon, et al.
Veröffentlicht: (2025)
von: Kim, Sangwon, et al.
Veröffentlicht: (2025)
Narrowing Information Bottleneck Theory for Multimodal Image-Text Representations Interpretability
von: Zhu, Zhiyu, et al.
Veröffentlicht: (2025)
von: Zhu, Zhiyu, et al.
Veröffentlicht: (2025)
HELM: Hierarchical and Explicit Label Modeling with Graph Learning for Multi-Label Image Classification
von: Stoimchev, Marjan, et al.
Veröffentlicht: (2026)
von: Stoimchev, Marjan, et al.
Veröffentlicht: (2026)
Phonology Recognition in American Sign Language
von: Tavella, Federico, et al.
Veröffentlicht: (2021)
von: Tavella, Federico, et al.
Veröffentlicht: (2021)
Revealing Temporal Label Noise in Multimodal Hateful Video Classification
von: Yang, Shuonan, et al.
Veröffentlicht: (2025)
von: Yang, Shuonan, et al.
Veröffentlicht: (2025)
Discover-then-Name: Task-Agnostic Concept Bottlenecks via Automated Concept Discovery
von: Rao, Sukrut, et al.
Veröffentlicht: (2024)
von: Rao, Sukrut, et al.
Veröffentlicht: (2024)
Mitigating Coordinate Prediction Bias from Positional Encoding Failures
von: Tao, Xingjian, et al.
Veröffentlicht: (2025)
von: Tao, Xingjian, et al.
Veröffentlicht: (2025)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
SPARC: Concept-Aligned Sparse Autoencoders for Cross-Model and Cross-Modal Interpretability
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2025)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning
von: Xie, Haodong, et al.
Veröffentlicht: (2024) -
Attributes-aware Visual Emotion Representation Learning
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025) -
Noise-Free Explanation for Driving Action Prediction
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024) -
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
von: Tavella, Federico, et al.
Veröffentlicht: (2025) -
Representation Understanding via Activation Maximization
von: Zhu, Hongbo, et al.
Veröffentlicht: (2025)