Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video
Fuente:
arXiv
Guardado en:
| Autores principales: | Joseph, Sonia, Suresh, Praneet, Hufe, Lorenz, Stevinson, Edward, Graham, Robert, Vadi, Yash, Bzdok, Danilo, Lapuschkin, Sebastian, Sharkey, Lee, Richards, Blake Aaron |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Steering CLIP's vision transformer with sparse autoencoders
por: Joseph, Sonia, et al.
Publicado: (2025)
por: Joseph, Sonia, et al.
Publicado: (2025)
From Noise to Narrative: Tracing the Origins of Hallucinations in Transformers
por: Suresh, Praneet, et al.
Publicado: (2025)
por: Suresh, Praneet, et al.
Publicado: (2025)
Quantifying LLM Attention-Head Stability: Implications for Circuit Universality
por: Bali, Karan, et al.
Publicado: (2026)
por: Bali, Karan, et al.
Publicado: (2026)
Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP
por: Hufe, Lorenz, et al.
Publicado: (2025)
por: Hufe, Lorenz, et al.
Publicado: (2025)
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
por: Dreyer, Maximilian, et al.
Publicado: (2025)
por: Dreyer, Maximilian, et al.
Publicado: (2025)
The Uncanny Valley: A Comprehensive Analysis of Diffusion Models
por: Ghanem, Karam, et al.
Publicado: (2024)
por: Ghanem, Karam, et al.
Publicado: (2024)
Estimating Unknown Population Sizes Using the Hypergeometric Distribution
por: Hodgson, Liam, et al.
Publicado: (2024)
por: Hodgson, Liam, et al.
Publicado: (2024)
Open Problems in Mechanistic Interpretability
por: Sharkey, Lee, et al.
Publicado: (2025)
por: Sharkey, Lee, et al.
Publicado: (2025)
Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
por: Das, Sourya Dipta, et al.
Publicado: (2024)
por: Das, Sourya Dipta, et al.
Publicado: (2024)
ContextBench: Modifying Contexts for Targeted Latent Activation
por: Graham, Robert, et al.
Publicado: (2025)
por: Graham, Robert, et al.
Publicado: (2025)
Towards the AI Historian: Agentic Information Extraction from Primary Sources
por: Hufe, Lorenz, et al.
Publicado: (2026)
por: Hufe, Lorenz, et al.
Publicado: (2026)
Cultural Heritage in International Economic Law
por: Vadi, Valentina
Publicado: (2024)
por: Vadi, Valentina
Publicado: (2024)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
por: Bareeva, Dilyara, et al.
Publicado: (2024)
por: Bareeva, Dilyara, et al.
Publicado: (2024)
Unsupervised Out-of-Distribution Dialect Detection with Mahalanobis Distance
por: Das, Sourya Dipta, et al.
Publicado: (2023)
por: Das, Sourya Dipta, et al.
Publicado: (2023)
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition
por: Braun, Dan, et al.
Publicado: (2025)
por: Braun, Dan, et al.
Publicado: (2025)
Adversarial Attacks Leverage Interference Between Features in Superposition
por: Stevinson, Edward, et al.
Publicado: (2025)
por: Stevinson, Edward, et al.
Publicado: (2025)
Prismas
Publicado: (2012)
Publicado: (2012)
Prisma
Publicado: (2003)
Publicado: (2003)
Granulomatous mural folliculitis and cytotoxic interface dermatitis in a pygmy goat associated with ovine herpesvirus‐2 and systemic lesions of malignant catarrhal fever
por: Peter Richards‐Rios, et al.
Publicado: (2025)
por: Peter Richards‐Rios, et al.
Publicado: (2025)
Demographic Dataset on Race, Ethnicity, Age and Sex in Neuromuscular Disease Studies (2004-2024)
por: Fontanelli, Lorenzo, et al.
Publicado: (2025)
por: Fontanelli, Lorenzo, et al.
Publicado: (2025)
Interpreting Physics in Video World Models
por: Joseph, Sonia, et al.
Publicado: (2026)
por: Joseph, Sonia, et al.
Publicado: (2026)
SIREN: An Open Source Neutrino Injection Toolkit
por: Schneider, Austin, et al.
Publicado: (2024)
por: Schneider, Austin, et al.
Publicado: (2024)
Prisma Tecnológico
Publicado: (2021)
Publicado: (2021)
Prisma Social
Publicado: (2012)
Publicado: (2012)
Prisma Jurídico
Publicado: (2019)
Publicado: (2019)
3D-Speaker-Toolkit: An Open-Source Toolkit for Multimodal Speaker Verification and Diarization
por: Chen, Yafeng, et al.
Publicado: (2024)
por: Chen, Yafeng, et al.
Publicado: (2024)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
por: Pan, Leyi, et al.
Publicado: (2024)
por: Pan, Leyi, et al.
Publicado: (2024)
SocialPulse: An Open-Source Subreddit Sensemaking Toolkit
por: Birkelbach, Stephanie, et al.
Publicado: (2026)
por: Birkelbach, Stephanie, et al.
Publicado: (2026)
BERT vs GPT for financial engineering
por: Sharkey, Edward, et al.
Publicado: (2024)
por: Sharkey, Edward, et al.
Publicado: (2024)
PyEncode: An Open-Source Library for Structured Quantum State Preparation
por: Suresh, Krishnan, et al.
Publicado: (2026)
por: Suresh, Krishnan, et al.
Publicado: (2026)
Mitochondria‐nucleus crosstalk characterizes Alzheimer's disease across 1,5 million brain cells
por: Chloé Savignac, et al.
Publicado: (2025)
por: Chloé Savignac, et al.
Publicado: (2025)
Qualitätsmessung als Prisma
Publicado: (2024)
Publicado: (2024)
Declaración para Prismas
por: Charles A. Hale
Publicado: (2007)
por: Charles A. Hale
Publicado: (2007)
PGLearn -- An Open-Source Learning Toolkit for Optimal Power Flow
por: Klamkin, Michael, et al.
Publicado: (2025)
por: Klamkin, Michael, et al.
Publicado: (2025)
Amphion: An Open-Source Audio, Music and Speech Generation Toolkit
por: Zhang, Xueyao, et al.
Publicado: (2023)
por: Zhang, Xueyao, et al.
Publicado: (2023)
Groupy: An Open‐Source Toolkit for Molecular Simulation and Property Calculation
por: Ruichen Liu, et al.
Publicado: (2024)
por: Ruichen Liu, et al.
Publicado: (2024)
Discovering a Zeta Map Algorithm on Dyck Paths via Mechanistic Interpretability
por: Huang, Xiaoyu, et al.
Publicado: (2026)
por: Huang, Xiaoyu, et al.
Publicado: (2026)
Mechanistically Interpreting Compression in Vision-Language Models
por: Elluru, Veeraraju, et al.
Publicado: (2026)
por: Elluru, Veeraraju, et al.
Publicado: (2026)
Brain Age Prediction: Deep Models Need a Hand to Generalize
por: Reza Rajabli, et al.
Publicado: (2025)
por: Reza Rajabli, et al.
Publicado: (2025)
OPEN-THEATRE: An Open-Source Toolkit for LLM-based Interactive Drama
por: Xu, Tianyang, et al.
Publicado: (2025)
por: Xu, Tianyang, et al.
Publicado: (2025)
Ejemplares similares
-
Steering CLIP's vision transformer with sparse autoencoders
por: Joseph, Sonia, et al.
Publicado: (2025) -
From Noise to Narrative: Tracing the Origins of Hallucinations in Transformers
por: Suresh, Praneet, et al.
Publicado: (2025) -
Quantifying LLM Attention-Head Stability: Implications for Circuit Universality
por: Bali, Karan, et al.
Publicado: (2026) -
Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP
por: Hufe, Lorenz, et al.
Publicado: (2025) -
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
por: Dreyer, Maximilian, et al.
Publicado: (2025)