Automatic Discovery of Visual Circuits
Fuente:
arXiv
Saved in:
| Main Authors: | Rajaram, Achyuta, Chowdhury, Neil, Torralba, Antonio, Andreas, Jacob, Schwettmann, Sarah |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Line of Sight: On Linear Representations in VLLMs
by: Rajaram, Achyuta, et al.
Published: (2025)
by: Rajaram, Achyuta, et al.
Published: (2025)
A Multimodal Automated Interpretability Agent
by: Shaham, Tamar Rott, et al.
Published: (2024)
by: Shaham, Tamar Rott, et al.
Published: (2024)
Nearest Neighbor Normalization Improves Multimodal Retrieval
by: Chowdhury, Neil, et al.
Published: (2024)
by: Chowdhury, Neil, et al.
Published: (2024)
CoE: Chain-of-Explanation via Automatic Visual Concept Circuit Description and Polysemanticity Quantification
by: Yu, Wenlong, et al.
Published: (2025)
by: Yu, Wenlong, et al.
Published: (2025)
Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations
by: Kwon, Dahee, et al.
Published: (2025)
by: Kwon, Dahee, et al.
Published: (2025)
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
by: Che, Liwei, et al.
Published: (2026)
by: Che, Liwei, et al.
Published: (2026)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
by: Cazenavette, George, et al.
Published: (2025)
by: Cazenavette, George, et al.
Published: (2025)
Automatically Generating Visual Hallucination Test Cases for Multimodal Large Language Models
by: Liu, Zhongye, et al.
Published: (2024)
by: Liu, Zhongye, et al.
Published: (2024)
Visual Prompt Discovery via Semantic Exploration
by: Kim, Jaechang, et al.
Published: (2026)
by: Kim, Jaechang, et al.
Published: (2026)
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers
by: Swain, Kabir, et al.
Published: (2026)
by: Swain, Kabir, et al.
Published: (2026)
VCD: A Dataset for Visual Commonsense Discovery in Images
by: Shen, Xiangqing, et al.
Published: (2024)
by: Shen, Xiangqing, et al.
Published: (2024)
Image-Intrinsic Priors for Integrated Circuit Defect Detection and Novel Class Discovery via Self-Supervised Learning
by: Zhao, Botong., et al.
Published: (2025)
by: Zhao, Botong., et al.
Published: (2025)
Enhancing Skin Disease Diagnosis: Interpretable Visual Concept Discovery with SAM
by: Hu, Xin, et al.
Published: (2024)
by: Hu, Xin, et al.
Published: (2024)
TSOM: Small Object Motion Detection Neural Network Inspired by Avian Visual Circuit
by: Hu, Pignge, et al.
Published: (2024)
by: Hu, Pignge, et al.
Published: (2024)
TM-PATHVQA:90000+ Textless Multilingual Questions for Medical Visual Question Answering
by: Rajkhowa, Tonmoy, et al.
Published: (2024)
by: Rajkhowa, Tonmoy, et al.
Published: (2024)
TruthLens: Visual Grounding for Universal DeepFake Reasoning
by: Kundu, Rohit, et al.
Published: (2025)
by: Kundu, Rohit, et al.
Published: (2025)
Single-pass Adaptive Image Tokenization for Minimum Program Search
by: Duggal, Shivam, et al.
Published: (2025)
by: Duggal, Shivam, et al.
Published: (2025)
DiffEM: Learning from Corrupted Data with Diffusion Models via Expectation Maximization
by: Hosseintabar, Danial, et al.
Published: (2025)
by: Hosseintabar, Danial, et al.
Published: (2025)
Synthetic Data Augmentation for Table Detection: Re-evaluating TableNet's Performance with Automatically Generated Document Images
by: Sahukara, Krishna, et al.
Published: (2025)
by: Sahukara, Krishna, et al.
Published: (2025)
Adaptive Length Image Tokenization via Recurrent Allocation
by: Duggal, Shivam, et al.
Published: (2024)
by: Duggal, Shivam, et al.
Published: (2024)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
by: Tuong, Nguyen Anh, et al.
Published: (2026)
by: Tuong, Nguyen Anh, et al.
Published: (2026)
ResAgent: Entropy-based Prior Point Discovery and Visual Reasoning for Referring Expression Segmentation
by: Wang, Yihao, et al.
Published: (2026)
by: Wang, Yihao, et al.
Published: (2026)
AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs
by: Chowdhury, Sanjoy, et al.
Published: (2025)
by: Chowdhury, Sanjoy, et al.
Published: (2025)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
by: Zhu, Wanrong, et al.
Published: (2024)
by: Zhu, Wanrong, et al.
Published: (2024)
Visual Homing in Outdoor Robots Using Mushroom Body Circuits and Learning Walks
by: Gattaux, Gabriel G., et al.
Published: (2025)
by: Gattaux, Gabriel G., et al.
Published: (2025)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
by: Zheng, Sixiao, et al.
Published: (2024)
by: Zheng, Sixiao, et al.
Published: (2024)
Separating Knowledge and Perception with Procedural Data
by: Rodríguez-Muñoz, Adrián, et al.
Published: (2025)
by: Rodríguez-Muñoz, Adrián, et al.
Published: (2025)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
by: Zhang, Ziheng, et al.
Published: (2025)
by: Zhang, Ziheng, et al.
Published: (2025)
VB: Visibility Benchmark for Visibility and Perspective Reasoning in Images
by: Tripathi, Neil
Published: (2026)
by: Tripathi, Neil
Published: (2026)
MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks
by: Chowdhury, Sanjoy, et al.
Published: (2025)
by: Chowdhury, Sanjoy, et al.
Published: (2025)
Certified Circuits: Stability Guarantees for Mechanistic Circuits
by: Anani, Alaa, et al.
Published: (2026)
by: Anani, Alaa, et al.
Published: (2026)
VSA4VQA: Scaling a Vector Symbolic Architecture to Visual Question Answering on Natural Images
by: Penzkofer, Anna, et al.
Published: (2024)
by: Penzkofer, Anna, et al.
Published: (2024)
FACE: Faithful Automatic Concept Extraction
by: Bhusal, Dipkamal, et al.
Published: (2025)
by: Bhusal, Dipkamal, et al.
Published: (2025)
Automatic dental superimposition of 3D intraorals and 2D photographs for human identification
by: Villegas-Yeguas, Antonio D., et al.
Published: (2026)
by: Villegas-Yeguas, Antonio D., et al.
Published: (2026)
Classifier-to-Bias: Toward Unsupervised Automatic Bias Detection for Visual Classifiers
by: Guimard, Quentin, et al.
Published: (2025)
by: Guimard, Quentin, et al.
Published: (2025)
Automatic Medical Report Generation: Methods and Applications
by: Guo, Li, et al.
Published: (2024)
by: Guo, Li, et al.
Published: (2024)
Alignment-Aware and Reliability-Gated Multimodal Fusion for Unmanned Aerial Vehicle Detection Across Heterogeneous Thermal-Visual Sensors
by: Jahan, Ishrat, et al.
Published: (2026)
by: Jahan, Ishrat, et al.
Published: (2026)
Ontology-Guided Diffusion for Zero-Shot Visual Sim2Real Transfer
by: Youssef, Mohamed, et al.
Published: (2026)
by: Youssef, Mohamed, et al.
Published: (2026)
Generalized Class Discovery in Instance Segmentation
by: Hoang, Cuong Manh, et al.
Published: (2025)
by: Hoang, Cuong Manh, et al.
Published: (2025)
Automatized Self-Supervised Learning for Skin Lesion Screening
by: Useini, Vullnet, et al.
Published: (2023)
by: Useini, Vullnet, et al.
Published: (2023)
Similar Items
-
Line of Sight: On Linear Representations in VLLMs
by: Rajaram, Achyuta, et al.
Published: (2025) -
A Multimodal Automated Interpretability Agent
by: Shaham, Tamar Rott, et al.
Published: (2024) -
Nearest Neighbor Normalization Improves Multimodal Retrieval
by: Chowdhury, Neil, et al.
Published: (2024) -
CoE: Chain-of-Explanation via Automatic Visual Concept Circuit Description and Polysemanticity Quantification
by: Yu, Wenlong, et al.
Published: (2025) -
Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations
by: Kwon, Dahee, et al.
Published: (2025)