Automatic Discovery of Visual Circuits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rajaram, Achyuta, Chowdhury, Neil, Torralba, Antonio, Andreas, Jacob, Schwettmann, Sarah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Line of Sight: On Linear Representations in VLLMs
von: Rajaram, Achyuta, et al.
Veröffentlicht: (2025)
von: Rajaram, Achyuta, et al.
Veröffentlicht: (2025)
A Multimodal Automated Interpretability Agent
von: Shaham, Tamar Rott, et al.
Veröffentlicht: (2024)
von: Shaham, Tamar Rott, et al.
Veröffentlicht: (2024)
Nearest Neighbor Normalization Improves Multimodal Retrieval
von: Chowdhury, Neil, et al.
Veröffentlicht: (2024)
von: Chowdhury, Neil, et al.
Veröffentlicht: (2024)
CoE: Chain-of-Explanation via Automatic Visual Concept Circuit Description and Polysemanticity Quantification
von: Yu, Wenlong, et al.
Veröffentlicht: (2025)
von: Yu, Wenlong, et al.
Veröffentlicht: (2025)
Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations
von: Kwon, Dahee, et al.
Veröffentlicht: (2025)
von: Kwon, Dahee, et al.
Veröffentlicht: (2025)
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
von: Che, Liwei, et al.
Veröffentlicht: (2026)
von: Che, Liwei, et al.
Veröffentlicht: (2026)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
von: Cazenavette, George, et al.
Veröffentlicht: (2025)
von: Cazenavette, George, et al.
Veröffentlicht: (2025)
Automatically Generating Visual Hallucination Test Cases for Multimodal Large Language Models
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
Visual Prompt Discovery via Semantic Exploration
von: Kim, Jaechang, et al.
Veröffentlicht: (2026)
von: Kim, Jaechang, et al.
Veröffentlicht: (2026)
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers
von: Swain, Kabir, et al.
Veröffentlicht: (2026)
von: Swain, Kabir, et al.
Veröffentlicht: (2026)
VCD: A Dataset for Visual Commonsense Discovery in Images
von: Shen, Xiangqing, et al.
Veröffentlicht: (2024)
von: Shen, Xiangqing, et al.
Veröffentlicht: (2024)
Image-Intrinsic Priors for Integrated Circuit Defect Detection and Novel Class Discovery via Self-Supervised Learning
von: Zhao, Botong., et al.
Veröffentlicht: (2025)
von: Zhao, Botong., et al.
Veröffentlicht: (2025)
Enhancing Skin Disease Diagnosis: Interpretable Visual Concept Discovery with SAM
von: Hu, Xin, et al.
Veröffentlicht: (2024)
von: Hu, Xin, et al.
Veröffentlicht: (2024)
TSOM: Small Object Motion Detection Neural Network Inspired by Avian Visual Circuit
von: Hu, Pignge, et al.
Veröffentlicht: (2024)
von: Hu, Pignge, et al.
Veröffentlicht: (2024)
TM-PATHVQA:90000+ Textless Multilingual Questions for Medical Visual Question Answering
von: Rajkhowa, Tonmoy, et al.
Veröffentlicht: (2024)
von: Rajkhowa, Tonmoy, et al.
Veröffentlicht: (2024)
TruthLens: Visual Grounding for Universal DeepFake Reasoning
von: Kundu, Rohit, et al.
Veröffentlicht: (2025)
von: Kundu, Rohit, et al.
Veröffentlicht: (2025)
Single-pass Adaptive Image Tokenization for Minimum Program Search
von: Duggal, Shivam, et al.
Veröffentlicht: (2025)
von: Duggal, Shivam, et al.
Veröffentlicht: (2025)
DiffEM: Learning from Corrupted Data with Diffusion Models via Expectation Maximization
von: Hosseintabar, Danial, et al.
Veröffentlicht: (2025)
von: Hosseintabar, Danial, et al.
Veröffentlicht: (2025)
Synthetic Data Augmentation for Table Detection: Re-evaluating TableNet's Performance with Automatically Generated Document Images
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
Adaptive Length Image Tokenization via Recurrent Allocation
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
ResAgent: Entropy-based Prior Point Discovery and Visual Reasoning for Referring Expression Segmentation
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
Visual Homing in Outdoor Robots Using Mushroom Body Circuits and Learning Walks
von: Gattaux, Gabriel G., et al.
Veröffentlicht: (2025)
von: Gattaux, Gabriel G., et al.
Veröffentlicht: (2025)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
von: Zheng, Sixiao, et al.
Veröffentlicht: (2024)
von: Zheng, Sixiao, et al.
Veröffentlicht: (2024)
Separating Knowledge and Perception with Procedural Data
von: Rodríguez-Muñoz, Adrián, et al.
Veröffentlicht: (2025)
von: Rodríguez-Muñoz, Adrián, et al.
Veröffentlicht: (2025)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
VB: Visibility Benchmark for Visibility and Perspective Reasoning in Images
von: Tripathi, Neil
Veröffentlicht: (2026)
von: Tripathi, Neil
Veröffentlicht: (2026)
MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Certified Circuits: Stability Guarantees for Mechanistic Circuits
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
VSA4VQA: Scaling a Vector Symbolic Architecture to Visual Question Answering on Natural Images
von: Penzkofer, Anna, et al.
Veröffentlicht: (2024)
von: Penzkofer, Anna, et al.
Veröffentlicht: (2024)
FACE: Faithful Automatic Concept Extraction
von: Bhusal, Dipkamal, et al.
Veröffentlicht: (2025)
von: Bhusal, Dipkamal, et al.
Veröffentlicht: (2025)
Automatic dental superimposition of 3D intraorals and 2D photographs for human identification
von: Villegas-Yeguas, Antonio D., et al.
Veröffentlicht: (2026)
von: Villegas-Yeguas, Antonio D., et al.
Veröffentlicht: (2026)
Classifier-to-Bias: Toward Unsupervised Automatic Bias Detection for Visual Classifiers
von: Guimard, Quentin, et al.
Veröffentlicht: (2025)
von: Guimard, Quentin, et al.
Veröffentlicht: (2025)
Automatic Medical Report Generation: Methods and Applications
von: Guo, Li, et al.
Veröffentlicht: (2024)
von: Guo, Li, et al.
Veröffentlicht: (2024)
Alignment-Aware and Reliability-Gated Multimodal Fusion for Unmanned Aerial Vehicle Detection Across Heterogeneous Thermal-Visual Sensors
von: Jahan, Ishrat, et al.
Veröffentlicht: (2026)
von: Jahan, Ishrat, et al.
Veröffentlicht: (2026)
Ontology-Guided Diffusion for Zero-Shot Visual Sim2Real Transfer
von: Youssef, Mohamed, et al.
Veröffentlicht: (2026)
von: Youssef, Mohamed, et al.
Veröffentlicht: (2026)
Generalized Class Discovery in Instance Segmentation
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
Automatized Self-Supervised Learning for Skin Lesion Screening
von: Useini, Vullnet, et al.
Veröffentlicht: (2023)
von: Useini, Vullnet, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Line of Sight: On Linear Representations in VLLMs
von: Rajaram, Achyuta, et al.
Veröffentlicht: (2025) -
A Multimodal Automated Interpretability Agent
von: Shaham, Tamar Rott, et al.
Veröffentlicht: (2024) -
Nearest Neighbor Normalization Improves Multimodal Retrieval
von: Chowdhury, Neil, et al.
Veröffentlicht: (2024) -
CoE: Chain-of-Explanation via Automatic Visual Concept Circuit Description and Polysemanticity Quantification
von: Yu, Wenlong, et al.
Veröffentlicht: (2025) -
Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations
von: Kwon, Dahee, et al.
Veröffentlicht: (2025)