Gespeichert in:
| Hauptverfasser: | Nielsen, Beatrix M. G., Marconato, Emanuele, Gresele, Luigi, Dittadi, Andrea, Buchholz, Simon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.15438 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Does Closeness in Distribution Imply Representational Similarity? An Identifiability Perspective
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2025)
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2025)
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
Relational Linear Properties in Language Models: An Empirical Investigation
von: Valer, Giovanni, et al.
Veröffentlicht: (2026)
von: Valer, Giovanni, et al.
Veröffentlicht: (2026)
Causal Component Analysis
von: Wendong, Liang, et al.
Veröffentlicht: (2023)
von: Wendong, Liang, et al.
Veröffentlicht: (2023)
Concise and Logically Consistent Conformal Sets for Neuro-Symbolic Concept-Based Models
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2026)
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2026)
What is causal about causal models and representations?
von: Jørgensen, Frederik Hytting, et al.
Veröffentlicht: (2025)
von: Jørgensen, Frederik Hytting, et al.
Veröffentlicht: (2025)
Shortcuts and Identifiability in Concept-based Models from a Neuro-Symbolic Lens
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2025)
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2025)
DiffEnc: Variational Diffusion with a Learned Encoder
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2023)
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2023)
BEARS Make Neuro-Symbolic Models Aware of their Reasoning Shortcuts
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
Sparse Shift Autoencoders for Identifying Concepts from Large Language Model Activations
von: Joshi, Shruti, et al.
Veröffentlicht: (2025)
von: Joshi, Shruti, et al.
Veröffentlicht: (2025)
A Neuro-Symbolic Benchmark Suite for Concept Quality and Reasoning Shortcuts
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2024)
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2024)
Rank-Aware Spectral Bounds on Attention Logits for Stable Low-Precision Training
von: Emadi, Seyed Morteza
Veröffentlicht: (2026)
von: Emadi, Seyed Morteza
Veröffentlicht: (2026)
Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
von: Rajendran, Goutham, et al.
Veröffentlicht: (2024)
von: Rajendran, Goutham, et al.
Veröffentlicht: (2024)
From Logits to Hierarchies: Hierarchical Clustering made Simple
von: Palumbo, Emanuele, et al.
Veröffentlicht: (2024)
von: Palumbo, Emanuele, et al.
Veröffentlicht: (2024)
Harmonizing Multi-Objective LLM Unlearning via Unified Domain Representation and Bidirectional Logit Distillation
von: Zhong, Yisheng, et al.
Veröffentlicht: (2026)
von: Zhong, Yisheng, et al.
Veröffentlicht: (2026)
Measuring Time-Series Dataset Similarity using Wasserstein Distance
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
Logit Dynamics in Softmax Policy Gradient Methods
von: Li, Yingru
Veröffentlicht: (2025)
von: Li, Yingru
Veröffentlicht: (2025)
SMART: Relation-Aware Learning of Geometric Representations for Knowledge Graphs
von: Amouzouvi, Kossi, et al.
Veröffentlicht: (2025)
von: Amouzouvi, Kossi, et al.
Veröffentlicht: (2025)
Spectral Logit Sculpting: Adaptive Low-Rank Logit Transformation for Controlled Text Generation
von: Li, Jin, et al.
Veröffentlicht: (2025)
von: Li, Jin, et al.
Veröffentlicht: (2025)
Logit Distillation on Manifolds: Mapping by Learning
von: Yang, Yiru, et al.
Veröffentlicht: (2026)
von: Yang, Yiru, et al.
Veröffentlicht: (2026)
Algorithmic causal structure emerging through compression
von: Wendong, Liang, et al.
Veröffentlicht: (2025)
von: Wendong, Liang, et al.
Veröffentlicht: (2025)
Peak-Controlled Logits Poisoning Attack in Federated Distillation
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction
von: Lu, Shengyao, et al.
Veröffentlicht: (2025)
von: Lu, Shengyao, et al.
Veröffentlicht: (2025)
Molecular Graph Representation Learning via Structural Similarity Information
von: Yao, Chengyu, et al.
Veröffentlicht: (2024)
von: Yao, Chengyu, et al.
Veröffentlicht: (2024)
What Representational Similarity Measures Imply about Decodable Information
von: Harvey, Sarah E., et al.
Veröffentlicht: (2024)
von: Harvey, Sarah E., et al.
Veröffentlicht: (2024)
An Explainable Multi-Task Similarity Measure: Integrating Accumulated Local Effects and Weighted Fréchet Distance
von: Hidalgo, Pablo, et al.
Veröffentlicht: (2026)
von: Hidalgo, Pablo, et al.
Veröffentlicht: (2026)
Auxiliary Reward Generation with Transition Distance Representation Learning
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
SCALA: Split Federated Learning with Concatenated Activations and Logit Adjustments
von: Yang, Jiarong, et al.
Veröffentlicht: (2024)
von: Yang, Jiarong, et al.
Veröffentlicht: (2024)
From Projection to Prediction: Beyond Logits for Scalable Language Models
von: Dong, Jianbing, et al.
Veröffentlicht: (2025)
von: Dong, Jianbing, et al.
Veröffentlicht: (2025)
CLadder: Assessing Causal Reasoning in Language Models
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
From Molecules to Mixtures: Learning Representations of Olfactory Mixture Similarity using Inductive Biases
von: Tom, Gary, et al.
Veröffentlicht: (2025)
von: Tom, Gary, et al.
Veröffentlicht: (2025)
Early Detection of Multidrug Resistance Using Multivariate Time Series Analysis and Interpretable Patient-Similarity Representations
von: Escudero-Arnanz, Óscar, et al.
Veröffentlicht: (2025)
von: Escudero-Arnanz, Óscar, et al.
Veröffentlicht: (2025)
Formalising the Logit Shift Induced by LoRA: A Technical Note
von: Shi, Xiang, et al.
Veröffentlicht: (2026)
von: Shi, Xiang, et al.
Veröffentlicht: (2026)
The Triangle of Similarity: A Multi-Faceted Framework for Comparing Neural Network Representations
von: Sirikova, Olha, et al.
Veröffentlicht: (2026)
von: Sirikova, Olha, et al.
Veröffentlicht: (2026)
Multiclass Local Calibration with the Jensen-Shannon Distance
von: Barbera, Cesare, et al.
Veröffentlicht: (2025)
von: Barbera, Cesare, et al.
Veröffentlicht: (2025)
SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion
von: Hu, Zizhao, et al.
Veröffentlicht: (2026)
von: Hu, Zizhao, et al.
Veröffentlicht: (2026)
Sharpness-Aware Minimization in Logit Space Efficiently Enhances Direct Preference Optimization
von: Luo, Haocheng, et al.
Veröffentlicht: (2026)
von: Luo, Haocheng, et al.
Veröffentlicht: (2026)
An Adversarial Example for Direct Logit Attribution: Memory Management in GELU-4L
von: Janiak, Jett, et al.
Veröffentlicht: (2023)
von: Janiak, Jett, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
When Does Closeness in Distribution Imply Representational Similarity? An Identifiability Perspective
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2025) -
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024) -
Relational Linear Properties in Language Models: An Empirical Investigation
von: Valer, Giovanni, et al.
Veröffentlicht: (2026) -
Causal Component Analysis
von: Wendong, Liang, et al.
Veröffentlicht: (2023) -
Concise and Logically Consistent Conformal Sets for Neuro-Symbolic Concept-Based Models
von: Bortolotti, Samuele, et al.
Veröffentlicht: (2026)