Softmax gradient policy for variance minimization and risk-averse multi armed bandits
Fuente:
arXiv
Saved in:
| Main Author: | Turinici, Gabriel |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Huber-energy measure quantization
by: Turinici, Gabriel
Published: (2022)
by: Turinici, Gabriel
Published: (2022)
Convergence of a L2 regularized Policy Gradient Algorithm for the Multi Armed Bandit
by: Anita, Stefana, et al.
Published: (2024)
by: Anita, Stefana, et al.
Published: (2024)
Optimal time sampling in physics-informed neural networks
by: Turinici, Gabriel
Published: (2024)
by: Turinici, Gabriel
Published: (2024)
Onflow: a model free, online portfolio allocation algorithm robust to transaction fees
by: Turinici, Gabriel, et al.
Published: (2023)
by: Turinici, Gabriel, et al.
Published: (2023)
Regime-Aware Time Weighting for Physics-Informed Neural Networks
by: Turinici, Gabriel
Published: (2024)
by: Turinici, Gabriel
Published: (2024)
Intrinsic and Extrinsic Organized Attention: Softmax Invariance and Network Sparsity
by: Fasina, Oluwadamilola, et al.
Published: (2025)
by: Fasina, Oluwadamilola, et al.
Published: (2025)
Precision autotuning for linear solvers via contextual bandit-based RL
by: Carson, Erin, et al.
Published: (2026)
by: Carson, Erin, et al.
Published: (2026)
MBExplainer: Multilevel bandit-based explanations for downstream models with augmented graph embeddings
by: Golgoon, Ashkan, et al.
Published: (2024)
by: Golgoon, Ashkan, et al.
Published: (2024)
Assessment of the January 2025 Los Angeles County wildfires: A multi-modal analysis of impact, response, and population exposure
by: Seydi, Seyd Teymoor
Published: (2025)
by: Seydi, Seyd Teymoor
Published: (2025)
Learning to Discover Iterative Spectral Algorithms
by: Liu, Zihang, et al.
Published: (2026)
by: Liu, Zihang, et al.
Published: (2026)
ProFlow: Zero-Shot Physics-Consistent Sampling via Proximal Flow Guidance
by: Yu, Zichao, et al.
Published: (2026)
by: Yu, Zichao, et al.
Published: (2026)
Autoregression-Free Neural Operators for Time-Dependent PDEs
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Core-Halo Decomposition: Decentralizing Large-Scale Fixed-Point Problems
by: Haixiang, et al.
Published: (2026)
by: Haixiang, et al.
Published: (2026)
Leveraging Gauge Freedom for Learning Non-Gradient Population Dynamics of Stochastic Systems
by: Berman, Jules, et al.
Published: (2026)
by: Berman, Jules, et al.
Published: (2026)
AutoNumerics: An Autonomous, PDE-Agnostic Multi-Agent Pipeline for Scientific Computing
by: Du, Jianda, et al.
Published: (2026)
by: Du, Jianda, et al.
Published: (2026)
Kinetic-based regularization: Learning spatial derivatives and PDE applications
by: Ganguly, Abhisek, et al.
Published: (2026)
by: Ganguly, Abhisek, et al.
Published: (2026)
V-ABFT: Variance-Based Adaptive Threshold for Fault-Tolerant Matrix Multiplication in Mixed-Precision Deep Learning
by: Gao, Yiheng, et al.
Published: (2026)
by: Gao, Yiheng, et al.
Published: (2026)
IFNSO: Iteration-Free Newton-Schulz Orthogonalization
by: Hu, Chen, et al.
Published: (2026)
by: Hu, Chen, et al.
Published: (2026)
Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models
by: Islam, Chashi Mahiul, et al.
Published: (2026)
by: Islam, Chashi Mahiul, et al.
Published: (2026)
CATO: Charted Attention for Neural PDE Operators
by: Cheng, Chun-Wun, et al.
Published: (2026)
by: Cheng, Chun-Wun, et al.
Published: (2026)
FlashSinkhorn: IO-Aware Entropic Optimal Transport on GPU
by: Ye, Felix X. -F., et al.
Published: (2026)
by: Ye, Felix X. -F., et al.
Published: (2026)
Shape-informed cardiac mechanics surrogates in data-scarce regimes via geometric encoding and generative augmentation
by: Carrara, Davide, et al.
Published: (2026)
by: Carrara, Davide, et al.
Published: (2026)
FI-KAN: Fractal Interpolation Kolmogorov-Arnold Networks
by: N'guessan, Gnankan Landry Regis
Published: (2026)
by: N'guessan, Gnankan Landry Regis
Published: (2026)
Low-Rank Compression of Pretrained Models via Randomized Subspace Iteration
by: Pourkamali-Anaraki, Farhad
Published: (2026)
by: Pourkamali-Anaraki, Farhad
Published: (2026)
Graph-Instructed Neural Networks for parametric problems with varying boundary conditions
by: Della Santa, Francesco, et al.
Published: (2026)
by: Della Santa, Francesco, et al.
Published: (2026)
A Practical Approach to Causal Inference over Time
by: Cinquini, Martina, et al.
Published: (2024)
by: Cinquini, Martina, et al.
Published: (2024)
A Mathematical Guide to Operator Learning
by: Boullé, Nicolas, et al.
Published: (2023)
by: Boullé, Nicolas, et al.
Published: (2023)
Truncated Matrix Completion - An Empirical Study
by: Naik, Rishhabh, et al.
Published: (2025)
by: Naik, Rishhabh, et al.
Published: (2025)
BWLer: Barycentric Weight Layer Elucidates a Precision-Conditioning Tradeoff for PINNs
by: Liu, Jerry, et al.
Published: (2025)
by: Liu, Jerry, et al.
Published: (2025)
Random weights of DNNs and emergence of fixed points
by: Berlyand, L., et al.
Published: (2025)
by: Berlyand, L., et al.
Published: (2025)
Deriving Transformer Architectures as Implicit Multinomial Regression
by: Actor, Jonas A., et al.
Published: (2025)
by: Actor, Jonas A., et al.
Published: (2025)
Sparse $L^1$-Autoencoders for Scientific Data Compression
by: Chung, Matthias, et al.
Published: (2024)
by: Chung, Matthias, et al.
Published: (2024)
Graph Neural Networks for Emulation of Finite-Element Ice Dynamics in Greenland and Antarctic Ice Sheets
by: Koo, Younghyun, et al.
Published: (2024)
by: Koo, Younghyun, et al.
Published: (2024)
Neural Operators with Localized Integral and Differential Kernels
by: Liu-Schiaffini, Miguel, et al.
Published: (2024)
by: Liu-Schiaffini, Miguel, et al.
Published: (2024)
HomPINNs: homotopy physics-informed neural networks for solving the inverse problems of nonlinear differential equations with multiple solutions
by: Zheng, Haoyang, et al.
Published: (2023)
by: Zheng, Haoyang, et al.
Published: (2023)
Generalized Orders of Magnitude for Scalable, Parallel, High-Dynamic-Range Computation
by: Heinsen, Franz A., et al.
Published: (2025)
by: Heinsen, Franz A., et al.
Published: (2025)
Critical Sampling for Robust Evolution Operator Learning of Unknown Dynamical Systems
by: Zhang, Ce, et al.
Published: (2023)
by: Zhang, Ce, et al.
Published: (2023)
AttNS: Attention-Inspired Numerical Solving For Limited Data Scenarios
by: Huang, Zhongzhan, et al.
Published: (2023)
by: Huang, Zhongzhan, et al.
Published: (2023)
Paired Autoencoders for Likelihood-free Estimation in Inverse Problems
by: Chung, Matthias, et al.
Published: (2024)
by: Chung, Matthias, et al.
Published: (2024)
Operator learning without the adjoint
by: Boullé, Nicolas, et al.
Published: (2024)
by: Boullé, Nicolas, et al.
Published: (2024)
Similar Items
-
Huber-energy measure quantization
by: Turinici, Gabriel
Published: (2022) -
Convergence of a L2 regularized Policy Gradient Algorithm for the Multi Armed Bandit
by: Anita, Stefana, et al.
Published: (2024) -
Optimal time sampling in physics-informed neural networks
by: Turinici, Gabriel
Published: (2024) -
Onflow: a model free, online portfolio allocation algorithm robust to transaction fees
by: Turinici, Gabriel, et al.
Published: (2023) -
Regime-Aware Time Weighting for Physics-Informed Neural Networks
by: Turinici, Gabriel
Published: (2024)