Explainable Concept Generation through Vision-Language Preference Learning for Understanding Neural Networks' Internal Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Taparia, Aditya, Sagar, Som, Senanayake, Ransalu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Failures Are Fated, But Can Be Faded: Characterizing and Mitigating Unwanted Behaviors in Large-Scale Vision and Language Models
by: Sagar, Som, et al.
Published: (2024)
by: Sagar, Som, et al.
Published: (2024)
LLM-Assisted Red Teaming of Diffusion Models through "Failures Are Fated, But Can Be Faded"
by: Sagar, Som, et al.
Published: (2024)
by: Sagar, Som, et al.
Published: (2024)
The Role of Predictive Uncertainty and Diversity in Embodied AI and Robot Learning
by: Senanayake, Ransalu
Published: (2024)
by: Senanayake, Ransalu
Published: (2024)
BaTCAVe: Trustworthy Explanations for Robot Behaviors
by: Sagar, Som, et al.
Published: (2024)
by: Sagar, Som, et al.
Published: (2024)
VLC Fusion: Vision-Language Conditioned Sensor Fusion for Robust Object Detection
by: Taparia, Aditya, et al.
Published: (2025)
by: Taparia, Aditya, et al.
Published: (2025)
If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions
by: Esfandiarpoor, Reza, et al.
Published: (2024)
by: Esfandiarpoor, Reza, et al.
Published: (2024)
Fairness in Autonomous Driving: Towards Understanding Confounding Factors in Object Detection under Challenging Weather
by: Pathiraja, Bimsara, et al.
Published: (2024)
by: Pathiraja, Bimsara, et al.
Published: (2024)
Multimodal Adaptive Retrieval Augmented Generation through Internal Representation Learning
by: Du, Ruoshuang, et al.
Published: (2026)
by: Du, Ruoshuang, et al.
Published: (2026)
Understanding Distributed Representations of Concepts in Deep Neural Networks without Supervision
by: Chang, Wonjoon, et al.
Published: (2023)
by: Chang, Wonjoon, et al.
Published: (2023)
Bayesian Multi-Scale Neural Network for Crowd Counting
by: Sagar, Abhinav
Published: (2020)
by: Sagar, Abhinav
Published: (2020)
Graph Neural Networks in Vision-Language Image Understanding: A Survey
by: Senior, Henry, et al.
Published: (2023)
by: Senior, Henry, et al.
Published: (2023)
Dynamic Scene Understanding from Vision-Language Representations
by: Pruss, Shahaf, et al.
Published: (2025)
by: Pruss, Shahaf, et al.
Published: (2025)
Transferring Textual Preferences to Vision-Language Understanding through Model Merging
by: Li, Chen-An, et al.
Published: (2025)
by: Li, Chen-An, et al.
Published: (2025)
Hierarchical Network Fusion for Multi-Modal Electron Micrograph Representation Learning with Foundational Large Language Models
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
SITUATE: Indoor Human Trajectory Prediction through Geometric Features and Self-Supervised Vision Representation
by: Capogrosso, Luigi, et al.
Published: (2024)
by: Capogrosso, Luigi, et al.
Published: (2024)
Circuit Tracing in Vision-Language Models: Understanding the Internal Mechanisms of Multimodal Thinking
by: Yang, Jingcheng, et al.
Published: (2026)
by: Yang, Jingcheng, et al.
Published: (2026)
Evaluating the Stability of Semantic Concept Representations in CNNs for Robust Explainability
by: Mikriukov, Georgii, et al.
Published: (2023)
by: Mikriukov, Georgii, et al.
Published: (2023)
Enhancing Zero-Shot Image Recognition in Vision-Language Models through Human-like Concept Guidance
by: Liu, Hui, et al.
Published: (2025)
by: Liu, Hui, et al.
Published: (2025)
MMRL: Multi-Modal Representation Learning for Vision-Language Models
by: Guo, Yuncheng, et al.
Published: (2025)
by: Guo, Yuncheng, et al.
Published: (2025)
Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning
by: Roy, Shuvendu, et al.
Published: (2024)
by: Roy, Shuvendu, et al.
Published: (2024)
Multi-Modal Instruction-Tuning Small-Scale Language-and-Vision Assistant for Semiconductor Electron Micrograph Analysis
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
An Explainable Fast Deep Neural Network for Emotion Recognition
by: Di Luzio, Francesco, et al.
Published: (2024)
by: Di Luzio, Francesco, et al.
Published: (2024)
Understanding Multimodal Deep Neural Networks: A Concept Selection View
by: Shang, Chenming, et al.
Published: (2024)
by: Shang, Chenming, et al.
Published: (2024)
Learning Decomposable and Debiased Representations via Attribute-Centric Information Bottlenecks
by: Hong, Jinyung, et al.
Published: (2024)
by: Hong, Jinyung, et al.
Published: (2024)
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers
by: Islam, Chashi Mahiul, et al.
Published: (2025)
by: Islam, Chashi Mahiul, et al.
Published: (2025)
Matryoshka Representation Learning
by: Kusupati, Aditya, et al.
Published: (2022)
by: Kusupati, Aditya, et al.
Published: (2022)
Learning Structured Representations with Hyperbolic Embeddings
by: Sinha, Aditya, et al.
Published: (2024)
by: Sinha, Aditya, et al.
Published: (2024)
Fractional Concepts in Neural Networks: Enhancing Activation Functions
by: Alijani, Zahra, et al.
Published: (2023)
by: Alijani, Zahra, et al.
Published: (2023)
Understanding Task Transfer in Vision-Language Models
by: Sachdeva, Bhuvan, et al.
Published: (2025)
by: Sachdeva, Bhuvan, et al.
Published: (2025)
Multilingual Diversity Improves Vision-Language Representations
by: Nguyen, Thao, et al.
Published: (2024)
by: Nguyen, Thao, et al.
Published: (2024)
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
by: Srivastava, Divyansh, et al.
Published: (2024)
by: Srivastava, Divyansh, et al.
Published: (2024)
Self-Evolving Visual Concept Library using Vision-Language Critics
by: Sehgal, Atharva, et al.
Published: (2025)
by: Sehgal, Atharva, et al.
Published: (2025)
Transitive Vision-Language Prompt Learning for Domain Generalization
by: Wang, Liyuan, et al.
Published: (2024)
by: Wang, Liyuan, et al.
Published: (2024)
Forget Less by Learning Together through Concept Consolidation
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
Curvature Learning for Generalization of Hyperbolic Neural Networks
by: Fan, Xiaomeng, et al.
Published: (2025)
by: Fan, Xiaomeng, et al.
Published: (2025)
Learning Topological Representations for Deep Image Understanding
by: Hu, Xiaoling
Published: (2024)
by: Hu, Xiaoling
Published: (2024)
Concept-based Analysis of Neural Networks via Vision-Language Models
by: Mangal, Ravi, et al.
Published: (2024)
by: Mangal, Ravi, et al.
Published: (2024)
Vision-Language Models Encode Clinical Guidelines for Concept-Based Medical Reasoning
by: Harmanani, Mohamed, et al.
Published: (2026)
by: Harmanani, Mohamed, et al.
Published: (2026)
Concept-skill Transferability-based Data Selection for Large Vision-Language Models
by: Lee, Jaewoo, et al.
Published: (2024)
by: Lee, Jaewoo, et al.
Published: (2024)
Concept-TRAK: Understanding how diffusion models learn concepts through concept-level attribution
by: Park, Yonghyun, et al.
Published: (2025)
by: Park, Yonghyun, et al.
Published: (2025)
Similar Items
-
Failures Are Fated, But Can Be Faded: Characterizing and Mitigating Unwanted Behaviors in Large-Scale Vision and Language Models
by: Sagar, Som, et al.
Published: (2024) -
LLM-Assisted Red Teaming of Diffusion Models through "Failures Are Fated, But Can Be Faded"
by: Sagar, Som, et al.
Published: (2024) -
The Role of Predictive Uncertainty and Diversity in Embodied AI and Robot Learning
by: Senanayake, Ransalu
Published: (2024) -
BaTCAVe: Trustworthy Explanations for Robot Behaviors
by: Sagar, Som, et al.
Published: (2024) -
VLC Fusion: Vision-Language Conditioned Sensor Fusion for Robust Object Detection
by: Taparia, Aditya, et al.
Published: (2025)