Gespeichert in:
| Hauptverfasser: | Piotrowski, Mateusz, Riechers, Paul M., Filan, Daniel, Shai, Adam S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.01954 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neural networks leverage nominally quantum and post-quantum representations
von: Riechers, Paul M., et al.
Veröffentlicht: (2025)
von: Riechers, Paul M., et al.
Veröffentlicht: (2025)
Transformers represent belief state geometry in their residual stream
von: Shai, Adam S., et al.
Veröffentlicht: (2024)
von: Shai, Adam S., et al.
Veröffentlicht: (2024)
Rank-1 LoRAs Encode Interpretable Reasoning Signals
von: Ward, Jake, et al.
Veröffentlicht: (2025)
von: Ward, Jake, et al.
Veröffentlicht: (2025)
Next-token pretraining implies in-context learning
von: Riechers, Paul M., et al.
Veröffentlicht: (2025)
von: Riechers, Paul M., et al.
Veröffentlicht: (2025)
Transformers learn factored representations
von: Shai, Adam, et al.
Veröffentlicht: (2026)
von: Shai, Adam, et al.
Veröffentlicht: (2026)
Geometry and Dynamics of LayerNorm
von: Riechers, Paul M.
Veröffentlicht: (2024)
von: Riechers, Paul M.
Veröffentlicht: (2024)
In-context learning agents are asymmetric belief updaters
von: Schubert, Johannes A., et al.
Veröffentlicht: (2024)
von: Schubert, Johannes A., et al.
Veröffentlicht: (2024)
Positive concave deep equilibrium models
von: Gabor, Mateusz, et al.
Veröffentlicht: (2024)
von: Gabor, Mateusz, et al.
Veröffentlicht: (2024)
Fixed points of nonnegative neural networks
von: Piotrowski, Tomasz J., et al.
Veröffentlicht: (2021)
von: Piotrowski, Tomasz J., et al.
Veröffentlicht: (2021)
Concept-based explainability for an EEG transformer model
von: Gjølbye, Anders, et al.
Veröffentlicht: (2023)
von: Gjølbye, Anders, et al.
Veröffentlicht: (2023)
An explainable transformer circuit for compositional generalization
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
A multi-criteria approach for selecting an explanation from the set of counterfactuals produced by an ensemble of explainers
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
Learning the greatest common divisor: explaining transformer predictions
von: Charton, François
Veröffentlicht: (2023)
von: Charton, François
Veröffentlicht: (2023)
Accurate estimation of feature importance faithfulness for tree models
von: Gajewski, Mateusz, et al.
Veröffentlicht: (2024)
von: Gajewski, Mateusz, et al.
Veröffentlicht: (2024)
survex: an R package for explaining machine learning survival models
von: Spytek, Mikołaj, et al.
Veröffentlicht: (2023)
von: Spytek, Mikołaj, et al.
Veröffentlicht: (2023)
LLMs are not (consistently) Bayesian: Quantifying internal (in)consistencies of LLMs' probabilistic beliefs
von: Chen, Chacha, et al.
Veröffentlicht: (2026)
von: Chen, Chacha, et al.
Veröffentlicht: (2026)
MolPILE -- large-scale, diverse dataset for molecular representation learning
von: Adamczyk, Jakub, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jakub, et al.
Veröffentlicht: (2025)
Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning
von: Surdej, Rafał, et al.
Veröffentlicht: (2025)
von: Surdej, Rafał, et al.
Veröffentlicht: (2025)
Robust Conformal Prediction Using Privileged Information
von: Feldman, Shai, et al.
Veröffentlicht: (2024)
von: Feldman, Shai, et al.
Veröffentlicht: (2024)
How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation
von: Feldman, Shai, et al.
Veröffentlicht: (2026)
von: Feldman, Shai, et al.
Veröffentlicht: (2026)
Deep Dreams Are Made of This: Visualizing Monosemantic Features in Diffusion Models
von: Szokalski, Adam, et al.
Veröffentlicht: (2026)
von: Szokalski, Adam, et al.
Veröffentlicht: (2026)
Baseflow identification via explainable AI with Kolmogorov-Arnold networks
von: Liu, Chuyang, et al.
Veröffentlicht: (2024)
von: Liu, Chuyang, et al.
Veröffentlicht: (2024)
Do LLMs have core beliefs?
von: Sokol, Anna, et al.
Veröffentlicht: (2026)
von: Sokol, Anna, et al.
Veröffentlicht: (2026)
Statistical and structural identifiability in representation learning
von: Nelson, Walter, et al.
Veröffentlicht: (2026)
von: Nelson, Walter, et al.
Veröffentlicht: (2026)
Learning from positive and unlabeled examples -Finite size sample bounds
von: Mansouri, Farnam, et al.
Veröffentlicht: (2025)
von: Mansouri, Farnam, et al.
Veröffentlicht: (2025)
Online Learning with Improving Agents: Multiclass, Budgeted Agents and Bandit Learners
von: Ashkezari, Sajad, et al.
Veröffentlicht: (2026)
von: Ashkezari, Sajad, et al.
Veröffentlicht: (2026)
Multinomial belief networks for healthcare data
von: Donker, H. C., et al.
Veröffentlicht: (2023)
von: Donker, H. C., et al.
Veröffentlicht: (2023)
LaB-GATr: geometric algebra transformers for large biomedical surface and volume meshes
von: Suk, Julian, et al.
Veröffentlicht: (2024)
von: Suk, Julian, et al.
Veröffentlicht: (2024)
DeepShare: Sharing ReLU Across Channels and Layers for Efficient Private Inference
von: Bornfeld, Yonathan, et al.
Veröffentlicht: (2025)
von: Bornfeld, Yonathan, et al.
Veröffentlicht: (2025)
Learning to Score
von: Kriger, Yogev, et al.
Veröffentlicht: (2025)
von: Kriger, Yogev, et al.
Veröffentlicht: (2025)
Global explainability of a deep abstaining classifier
von: Dhaubhadel, Sayera, et al.
Veröffentlicht: (2025)
von: Dhaubhadel, Sayera, et al.
Veröffentlicht: (2025)
Contrastive representations of high-dimensional, structured treatments
von: Andreu, Oriol Corcoll, et al.
Veröffentlicht: (2024)
von: Andreu, Oriol Corcoll, et al.
Veröffentlicht: (2024)
VARSHAP: Addressing Global Dependency Problems in Explainable AI with Variance-Based Local Feature Attribution
von: Gajewski, Mateusz, et al.
Veröffentlicht: (2025)
von: Gajewski, Mateusz, et al.
Veröffentlicht: (2025)
Amortized Causal Discovery with Prior-Fitted Networks
von: Sypniewski, Mateusz, et al.
Veröffentlicht: (2025)
von: Sypniewski, Mateusz, et al.
Veröffentlicht: (2025)
Conformal Prediction with Corrupted Labels: Uncertain Imputation and Robust Re-weighting
von: Feldman, Shai, et al.
Veröffentlicht: (2025)
von: Feldman, Shai, et al.
Veröffentlicht: (2025)
A Novel Data-Dependent Learning Paradigm for Large Hypothesis Classes
von: Pour, Alireza F., et al.
Veröffentlicht: (2025)
von: Pour, Alireza F., et al.
Veröffentlicht: (2025)
Federated style aware transformer aggregation of representations
von: Jeon, Mincheol, et al.
Veröffentlicht: (2025)
von: Jeon, Mincheol, et al.
Veröffentlicht: (2025)
Evaluating representation learning on the protein structure universe
von: Jamasb, Arian R., et al.
Veröffentlicht: (2024)
von: Jamasb, Arian R., et al.
Veröffentlicht: (2024)
Universal crystal material property prediction via multi-view geometric fusion in graph transformers
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
Characterizing higher-order representations through generative diffusion models explains human decoded neurofeedback performance
von: Asrari, Hojjat Azimi, et al.
Veröffentlicht: (2025)
von: Asrari, Hojjat Azimi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neural networks leverage nominally quantum and post-quantum representations
von: Riechers, Paul M., et al.
Veröffentlicht: (2025) -
Transformers represent belief state geometry in their residual stream
von: Shai, Adam S., et al.
Veröffentlicht: (2024) -
Rank-1 LoRAs Encode Interpretable Reasoning Signals
von: Ward, Jake, et al.
Veröffentlicht: (2025) -
Next-token pretraining implies in-context learning
von: Riechers, Paul M., et al.
Veröffentlicht: (2025) -
Transformers learn factored representations
von: Shai, Adam, et al.
Veröffentlicht: (2026)