Gespeichert in:
| 1. Verfasser: | |
|---|---|
| Format: | Recurso digital |
| Sprache: | Englisch |
| Veröffentlicht: |
Zenodo
2025
|
| Online-Zugang: | https://doi.org/10.5281/zenodo.17741981 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
| _version_ | 1866901519223226368 |
|---|---|
| author | Rios-Sialer, Ian |
| author_facet | Rios-Sialer, Ian |
| contents | <p><em>Category-Theoretic Wanderings into Interpretability</em> is a piece of <strong>technical autotheory</strong> that asks how we can use category theory to frame interpretability. It queers the ecologies of knowledge, employing abstract mathematical language to discuss both intimate and technical frameworks. The work writes about love addiction, faithfulness in Anthropic's Circuit Tracing experiments, philosophical questions of meaning, and invites the field of AI Safety to feel their way through opacity.</p> |
| format | Recurso digital |
| id | zenodo_https___doi_org_10_5281_zenodo_17741981 |
| institution | Zenodo |
| language | eng |
| publishDate | 2025 |
| publisher | Zenodo |
| record_format | zenodo |
| spellingShingle | Category-Theoretic Wanderings into Interpretability Rios-Sialer, Ian <p><em>Category-Theoretic Wanderings into Interpretability</em> is a piece of <strong>technical autotheory</strong> that asks how we can use category theory to frame interpretability. It queers the ecologies of knowledge, employing abstract mathematical language to discuss both intimate and technical frameworks. The work writes about love addiction, faithfulness in Anthropic's Circuit Tracing experiments, philosophical questions of meaning, and invites the field of AI Safety to feel their way through opacity.</p> |
| title | Category-Theoretic Wanderings into Interpretability |
| url | https://doi.org/10.5281/zenodo.17741981 |