Localising Dropout Variance in Twin Networks
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Doyle, Cooper |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Your Absorbing Discrete Diffusion Secretly Models the Bayesian Posterior
von: Doyle, Cooper
Veröffentlicht: (2025)
von: Doyle, Cooper
Veröffentlicht: (2025)
Tracing the Path to Grokking: Embeddings, Dropout, and Network Activation
von: Salah, Ahmed, et al.
Veröffentlicht: (2025)
von: Salah, Ahmed, et al.
Veröffentlicht: (2025)
NeuFair: Neural Network Fairness Repair with Dropout
von: Dasu, Vishnu Asutosh, et al.
Veröffentlicht: (2024)
von: Dasu, Vishnu Asutosh, et al.
Veröffentlicht: (2024)
Activation by Interval-wise Dropout: A Simple Way to Prevent Neural Networks from Plasticity Loss
von: Park, Sangyeon, et al.
Veröffentlicht: (2025)
von: Park, Sangyeon, et al.
Veröffentlicht: (2025)
Learning Adapter Rank via Symmetry Breaking
von: Doyle, Cooper, et al.
Veröffentlicht: (2025)
von: Doyle, Cooper, et al.
Veröffentlicht: (2025)
A Combinatorial Theory of Dropout: Subnetworks, Graph Geometry, and Generalization
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2025)
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2025)
Toward Efficient Influence Function: Dropout as a Compression Tool
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
LoRA Dropout as a Sparsity Regularizer for Overfitting Control
von: Lin, Yang, et al.
Veröffentlicht: (2024)
von: Lin, Yang, et al.
Veröffentlicht: (2024)
Scale-Dropout: Estimating Uncertainty in Deep Neural Networks Using Stochastic Scale
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2023)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2023)
Dropout Robustness and Cognitive Profiling of Transformer Models via Stochastic Inference
von: Caiado, Antônio Junior Alves, et al.
Veröffentlicht: (2026)
von: Caiado, Antônio Junior Alves, et al.
Veröffentlicht: (2026)
GNN-VPA: A Variance-Preserving Aggregation Strategy for Graph Neural Networks
von: Schneckenreiter, Lisa, et al.
Veröffentlicht: (2024)
von: Schneckenreiter, Lisa, et al.
Veröffentlicht: (2024)
A Mathematical Framework for Temporal Modeling and Counterfactual Policy Simulation of Student Dropout
von: da Silva, Rafael, et al.
Veröffentlicht: (2026)
von: da Silva, Rafael, et al.
Veröffentlicht: (2026)
Quantifying Variance in Evaluation Benchmarks
von: Madaan, Lovish, et al.
Veröffentlicht: (2024)
von: Madaan, Lovish, et al.
Veröffentlicht: (2024)
Math Takes Two: A test for emergent mathematical reasoning in communication
von: Cooper, Michael, et al.
Veröffentlicht: (2026)
von: Cooper, Michael, et al.
Veröffentlicht: (2026)
Arbitrariness and Social Prediction: The Confounding Role of Variance in Fair Classification
von: Cooper, A. Feder, et al.
Veröffentlicht: (2023)
von: Cooper, A. Feder, et al.
Veröffentlicht: (2023)
An Empirical Study of Fault Localisation Techniques for Deep Learning
von: Humbatova, Nargiz, et al.
Veröffentlicht: (2024)
von: Humbatova, Nargiz, et al.
Veröffentlicht: (2024)
On Vanishing Variance in Transformer Length Generalization
von: Li, Ruining, et al.
Veröffentlicht: (2025)
von: Li, Ruining, et al.
Veröffentlicht: (2025)
On Variance Reduction in Learning Mean Flows
von: Lu, Juanwu, et al.
Veröffentlicht: (2026)
von: Lu, Juanwu, et al.
Veröffentlicht: (2026)
Dynamic Dropout: Leveraging Conway's Game of Life for Neural Networks Regularization
von: Freire-Obregón, David, et al.
Veröffentlicht: (2025)
von: Freire-Obregón, David, et al.
Veröffentlicht: (2025)
Scaling Learning based Policy Optimization for Temporal Logic Tasks by Controller Network Dropout
von: Hashemi, Navid, et al.
Veröffentlicht: (2024)
von: Hashemi, Navid, et al.
Veröffentlicht: (2024)
Uncertainty Estimation using Variance-Gated Distributions
von: Gillis, H. Martin, et al.
Veröffentlicht: (2025)
von: Gillis, H. Martin, et al.
Veröffentlicht: (2025)
DynamicGate MLP Conditional Computation via Learned Structural Dropout and Input Dependent Gating for Functional Plasticity
von: Choi, Yong Il
Veröffentlicht: (2026)
von: Choi, Yong Il
Veröffentlicht: (2026)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
von: He, Jiafan, et al.
Veröffentlicht: (2025)
von: He, Jiafan, et al.
Veröffentlicht: (2025)
Adaptive Variance-Penalized Continual Learning with Fisher Regularization
von: Sarkar, Krisanu
Veröffentlicht: (2025)
von: Sarkar, Krisanu
Veröffentlicht: (2025)
Dual Randomized Smoothing: Beyond Global Noise Variance
von: Sun, Chenhao, et al.
Veröffentlicht: (2025)
von: Sun, Chenhao, et al.
Veröffentlicht: (2025)
A Variance Minimization Approach to Temporal-Difference Learning
von: Chen, Xingguo, et al.
Veröffentlicht: (2024)
von: Chen, Xingguo, et al.
Veröffentlicht: (2024)
DeltaDQ: Ultra-High Delta Compression for Fine-Tuned LLMs via Group-wise Dropout and Separate Quantization
von: Jiang, Yanfeng, et al.
Veröffentlicht: (2024)
von: Jiang, Yanfeng, et al.
Veröffentlicht: (2024)
General Uncertainty Estimation with Delta Variances
von: Schmitt, Simon, et al.
Veröffentlicht: (2025)
von: Schmitt, Simon, et al.
Veröffentlicht: (2025)
A Framework for Fair Evaluation of Variance-Aware Bandit Algorithms
von: Wolf, Elise
Veröffentlicht: (2025)
von: Wolf, Elise
Veröffentlicht: (2025)
Variance-Reduction Guidance: Sampling Trajectory Optimization for Diffusion Models
von: Xu, Shifeng, et al.
Veröffentlicht: (2025)
von: Xu, Shifeng, et al.
Veröffentlicht: (2025)
A Variance-Reduced Cubic-Regularized Newton for Policy Optimization
von: Sun, Cheng, et al.
Veröffentlicht: (2025)
von: Sun, Cheng, et al.
Veröffentlicht: (2025)
Resource-Constrained Affect Modelling via Variance Regularisation Pruning
von: Pinitas, Kosmas, et al.
Veröffentlicht: (2026)
von: Pinitas, Kosmas, et al.
Veröffentlicht: (2026)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Minimum Variance Unbiased N:M Sparsity for the Neural Gradients
von: Chmiel, Brian, et al.
Veröffentlicht: (2022)
von: Chmiel, Brian, et al.
Veröffentlicht: (2022)
Unsupervised Disentanglement of Content and Style via Variance-Invariance Constraints
von: Wu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Wu, Yuxuan, et al.
Veröffentlicht: (2024)
Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMs
von: Huang, Luke J., et al.
Veröffentlicht: (2026)
von: Huang, Luke J., et al.
Veröffentlicht: (2026)
Ratio-Variance Regularized Policy Optimization for Efficient LLM Fine-tuning
von: Luo, Yu, et al.
Veröffentlicht: (2026)
von: Luo, Yu, et al.
Veröffentlicht: (2026)
Low Variance Off-policy Evaluation with State-based Importance Sampling
von: Bossens, David M., et al.
Veröffentlicht: (2022)
von: Bossens, David M., et al.
Veröffentlicht: (2022)
Model-Based Epistemic Variance of Values for Risk-Aware Policy Optimization
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
Stochastic Variance-Reduced Iterative Hard Thresholding in Graph Sparsity Optimization
von: Fox, Derek, et al.
Veröffentlicht: (2024)
von: Fox, Derek, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Your Absorbing Discrete Diffusion Secretly Models the Bayesian Posterior
von: Doyle, Cooper
Veröffentlicht: (2025) -
Tracing the Path to Grokking: Embeddings, Dropout, and Network Activation
von: Salah, Ahmed, et al.
Veröffentlicht: (2025) -
NeuFair: Neural Network Fairness Repair with Dropout
von: Dasu, Vishnu Asutosh, et al.
Veröffentlicht: (2024) -
Activation by Interval-wise Dropout: A Simple Way to Prevent Neural Networks from Plasticity Loss
von: Park, Sangyeon, et al.
Veröffentlicht: (2025) -
Learning Adapter Rank via Symmetry Breaking
von: Doyle, Cooper, et al.
Veröffentlicht: (2025)