Local Inconsistency Resolution: The Interplay between Attention and Control in Probabilistic Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Richardson, Oliver E., Samiei, Mandana, Shakerinava, Mehran, Viviano, Joseph D., Kabid, Abdessamad El, Parviz, Ali, Bengio, Yoshua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Approaching the Harm of Gradient Attacks While Only Flipping Labels
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025)
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025)
Multiscale Neural PDE Surrogates for Prediction and Downscaling: Application to Ocean Currents
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025)
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025)
torchgfn: A PyTorch GFlowNet library
von: Viviano, Joseph D., et al.
Veröffentlicht: (2023)
von: Viviano, Joseph D., et al.
Veröffentlicht: (2023)
Newman's theorem via Carathéodory
von: Li, Yaqiao, et al.
Veröffentlicht: (2024)
von: Li, Yaqiao, et al.
Veröffentlicht: (2024)
Discrete Probabilistic Inference as Control in Multi-path Environments
von: Deleu, Tristan, et al.
Veröffentlicht: (2024)
von: Deleu, Tristan, et al.
Veröffentlicht: (2024)
Beyond Scalar Rewards: An Axiomatic Framework for Lexicographic MDPs
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2025)
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2025)
Action abstractions for amortized sampling
von: Boussif, Oussama, et al.
Veröffentlicht: (2024)
von: Boussif, Oussama, et al.
Veröffentlicht: (2024)
A scalable gene network model of regulatory dynamics in single cells
von: Bertin, Paul, et al.
Veröffentlicht: (2025)
von: Bertin, Paul, et al.
Veröffentlicht: (2025)
Machine learning and information theory concepts towards an AI Mathematician
von: Bengio, Yoshua, et al.
Veröffentlicht: (2024)
von: Bengio, Yoshua, et al.
Veröffentlicht: (2024)
Language models recognize dropout and Gaussian noise applied to their activations
von: Fornasiere, Damiano, et al.
Veröffentlicht: (2026)
von: Fornasiere, Damiano, et al.
Veröffentlicht: (2026)
The Expressive Limits of Diagonal SSMs for State-Tracking
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2026)
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2026)
Baking Symmetry into GFlowNets
von: Ma, George, et al.
Veröffentlicht: (2024)
von: Ma, George, et al.
Veröffentlicht: (2024)
Local Search GFlowNets
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
Tree Cross Attention
von: Feng, Leo, et al.
Veröffentlicht: (2023)
von: Feng, Leo, et al.
Veröffentlicht: (2023)
Weight-Sharing Regularization
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2023)
von: Shakerinava, Mehran, et al.
Veröffentlicht: (2023)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
Attention as an RNN
von: Feng, Leo, et al.
Veröffentlicht: (2024)
von: Feng, Leo, et al.
Veröffentlicht: (2024)
Probabilistic Tiny Recursive Model
von: Sghaier, Amin, et al.
Veröffentlicht: (2026)
von: Sghaier, Amin, et al.
Veröffentlicht: (2026)
Memory Efficient Neural Processes via Constant Memory Attention Block
von: Feng, Leo, et al.
Veröffentlicht: (2023)
von: Feng, Leo, et al.
Veröffentlicht: (2023)
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2025)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2025)
The Role of Symmetry in Optimizing Overparameterized Networks
von: Sareen, Kusha, et al.
Veröffentlicht: (2026)
von: Sareen, Kusha, et al.
Veröffentlicht: (2026)
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
On Generalization for Generative Flow Networks
von: Krichel, Anas, et al.
Veröffentlicht: (2024)
von: Krichel, Anas, et al.
Veröffentlicht: (2024)
Interventional Causal Representation Learning
von: Ahuja, Kartik, et al.
Veröffentlicht: (2022)
von: Ahuja, Kartik, et al.
Veröffentlicht: (2022)
A Complexity-Based Theory of Compositionality
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
Visual symbolic mechanisms: Emergent symbol processing in vision language models
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
Relative Trajectory Balance is equivalent to Trust-PCL
von: Deleu, Tristan, et al.
Veröffentlicht: (2025)
von: Deleu, Tristan, et al.
Veröffentlicht: (2025)
Fast Monte Carlo Tree Diffusion: 100x Speedup via Parallel Sparse Planning
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
In-Context Parametric Inference: Point or Distribution Estimators?
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
Generalized Optical Theorem for Structured Neutron Beams and Consequences for Forward-Transmission Null Tests of Time-Reversal Invariance
von: Samiei, Sepehr
Veröffentlicht: (2026)
von: Samiei, Sepehr
Veröffentlicht: (2026)
GFlowNet Foundations
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
Policy Targeting under Network Interference
von: Viviano, Davide
Veröffentlicht: (2019)
von: Viviano, Davide
Veröffentlicht: (2019)
a¿Quién gana en Irak? --
von: Viviano, Frank
Veröffentlicht: (2006)
von: Viviano, Frank
Veröffentlicht: (2006)
Experimental Design under Network Interference
von: Viviano, Davide
Veröffentlicht: (2020)
von: Viviano, Davide
Veröffentlicht: (2020)
Estimation of a Gas Diffusion Coefficient by Fitting Molecular Dynamics Trajectories to Finite-Difference Simulations
von: Viviano, Isaac
Veröffentlicht: (2025)
von: Viviano, Isaac
Veröffentlicht: (2025)
Can Safety Fine-Tuning Be More Principled? Lessons Learned from Cybersecurity
von: Williams-King, David, et al.
Veröffentlicht: (2025)
von: Williams-King, David, et al.
Veröffentlicht: (2025)
RL, but don't do anything I wouldn't do
von: Cohen, Michael K., et al.
Veröffentlicht: (2024)
von: Cohen, Michael K., et al.
Veröffentlicht: (2024)
Parity Requires Unified Input Dependence and Negative Eigenvalues in SSMs
von: Khavari, Behnoush, et al.
Veröffentlicht: (2025)
von: Khavari, Behnoush, et al.
Veröffentlicht: (2025)
Monte Carlo Tree Diffusion for System 2 Planning
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
Learning What Matters: Steering Diffusion via Spectrally Anisotropic Forward Noise
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Approaching the Harm of Gradient Attacks While Only Flipping Labels
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025) -
Multiscale Neural PDE Surrogates for Prediction and Downscaling: Application to Ocean Currents
von: El-Kabid, Abdessamad, et al.
Veröffentlicht: (2025) -
torchgfn: A PyTorch GFlowNet library
von: Viviano, Joseph D., et al.
Veröffentlicht: (2023) -
Newman's theorem via Carathéodory
von: Li, Yaqiao, et al.
Veröffentlicht: (2024) -
Discrete Probabilistic Inference as Control in Multi-path Environments
von: Deleu, Tristan, et al.
Veröffentlicht: (2024)