Policy Newton methods for Distortion Riskmetrics
Fuente:
arXiv
Saved in:
| Main Authors: | Pachal, Soumen, Maniyar, Mizhaan Prajit, A, Prashanth L. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generalized Random Direction Newton Algorithms for Stochastic Optimization
by: Pachal, Soumen, et al.
Published: (2026)
by: Pachal, Soumen, et al.
Published: (2026)
Risk-sensitive reinforcement learning using expectiles, shortfall risk and optimized certainty equivalent risk
by: Gupte, Sumedh, et al.
Published: (2026)
by: Gupte, Sumedh, et al.
Published: (2026)
Public Transit Arrival Prediction: a Seq2Seq RNN Approach
by: Bhutani, Nancy, et al.
Published: (2022)
by: Bhutani, Nancy, et al.
Published: (2022)
Policy Gradient Methods for Distortion Risk Measures
by: Vijayan, Nithia, et al.
Published: (2021)
by: Vijayan, Nithia, et al.
Published: (2021)
Risk Sharing, Measuring Variability, and Distortion Riskmetrics
by: Jean‐Gabriel Lauzier, et al.
Published: (2025)
by: Jean‐Gabriel Lauzier, et al.
Published: (2025)
Transformers can do Bayesian Clustering
by: Bhaskaran, Prajit, et al.
Published: (2025)
by: Bhaskaran, Prajit, et al.
Published: (2025)
X-Node: Self-Explanation is All We Need
by: Sengupta, Prajit, et al.
Published: (2025)
by: Sengupta, Prajit, et al.
Published: (2025)
A policy gradient approach for optimization of smooth risk measures
by: Vijayan, Nithia, et al.
Published: (2022)
by: Vijayan, Nithia, et al.
Published: (2022)
Smoothed functional-based gradient algorithms for off-policy reinforcement learning: A non-asymptotic viewpoint
by: Vijayan, Nithia, et al.
Published: (2021)
by: Vijayan, Nithia, et al.
Published: (2021)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
by: Iwaki, Ryo, et al.
Published: (2026)
by: Iwaki, Ryo, et al.
Published: (2026)
Optimizing Shortfall Risk Metric for Learning Regression Models
by: Ramaswamy, Harish G., et al.
Published: (2025)
by: Ramaswamy, Harish G., et al.
Published: (2025)
Adaptive pruning-based Newton's method for distributed learning
by: Chen, Shuzhen, et al.
Published: (2023)
by: Chen, Shuzhen, et al.
Published: (2023)
Risk Estimation in a Markov Cost Process: Lower and Upper Bounds
by: Thoppe, Gugan, et al.
Published: (2023)
by: Thoppe, Gugan, et al.
Published: (2023)
Concentration Bounds for Optimized Certainty Equivalent Risk Estimation
by: Ghosh, Ayon, et al.
Published: (2024)
by: Ghosh, Ayon, et al.
Published: (2024)
FireGNN: Neuro-Symbolic Graph Neural Networks with Trainable Fuzzy Rules for Interpretable Medical Image Classification
by: Sengupta, Prajit, et al.
Published: (2025)
by: Sengupta, Prajit, et al.
Published: (2025)
Policy Newton Algorithm in Reproducing Kernel Hilbert Space
by: Zhang, Yixian, et al.
Published: (2025)
by: Zhang, Yixian, et al.
Published: (2025)
Nonlinear discretizations and Newton's method: characterizing stationary points of regression objectives
by: Rowan, Conor
Published: (2025)
by: Rowan, Conor
Published: (2025)
Optimization of utility-based shortfall risk: A non-asymptotic viewpoint
by: Gupte, Sumedh, et al.
Published: (2023)
by: Gupte, Sumedh, et al.
Published: (2023)
A Variance-Reduced Cubic-Regularized Newton for Policy Optimization
by: Sun, Cheng, et al.
Published: (2025)
by: Sun, Cheng, et al.
Published: (2025)
A Finite-Sample Analysis of an Actor-Critic Algorithm for Mean-Variance Optimization in a Discounted MDP
by: Sangadi, Tejaram, et al.
Published: (2024)
by: Sangadi, Tejaram, et al.
Published: (2024)
Minimum mean-squared error estimation with bandit feedback
by: Ghosh, Ayon, et al.
Published: (2022)
by: Ghosh, Ayon, et al.
Published: (2022)
Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs
by: Thoppe, Gugan, et al.
Published: (2026)
by: Thoppe, Gugan, et al.
Published: (2026)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Gradient descent with generalized Newton's method
by: Bu, Zhiqi, et al.
Published: (2024)
by: Bu, Zhiqi, et al.
Published: (2024)
Newton-CG methods for nonconvex unconstrained optimization with Hölder continuous Hessian
by: He, Chuan, et al.
Published: (2023)
by: He, Chuan, et al.
Published: (2023)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026)
by: Rajendran, Prajit T, et al.
Published: (2026)
Reinforced Model Predictive Control via Trust-Region Quasi-Newton Policy Optimization
by: Brandner, Dean, et al.
Published: (2024)
by: Brandner, Dean, et al.
Published: (2024)
Preference-centric Bandits: Optimality of Mixtures and Regret-efficient Algorithms
by: Tatlı, Meltem, et al.
Published: (2025)
by: Tatlı, Meltem, et al.
Published: (2025)
Risk-sensitive Bandits: Arm Mixture Optimality and Regret-efficient Algorithms
by: Tatlı, Meltem, et al.
Published: (2025)
by: Tatlı, Meltem, et al.
Published: (2025)
Gauss-Newton Unlearning for the LLM Era
by: McKinney, Lev, et al.
Published: (2026)
by: McKinney, Lev, et al.
Published: (2026)
Markov Chain Variance Estimation: A Stochastic Approximation Approach
by: Agrawal, Shubhada, et al.
Published: (2024)
by: Agrawal, Shubhada, et al.
Published: (2024)
OASI: Objective-Aware Surrogate Initialization for Multi-Objective Bayesian Optimization in TinyML Keyword Spotting
by: Garai, Soumen, et al.
Published: (2025)
by: Garai, Soumen, et al.
Published: (2025)
Error whitening: Why Gauss-Newton outperforms Newton
by: McKay, Maricela Best, et al.
Published: (2026)
by: McKay, Maricela Best, et al.
Published: (2026)
Policy gradient methods for ordinal policies
by: Weinberger, Simón, et al.
Published: (2025)
by: Weinberger, Simón, et al.
Published: (2025)
SAND: One-Shot Feature Selection with Additive Noise Distortion
by: Pad, Pedram, et al.
Published: (2025)
by: Pad, Pedram, et al.
Published: (2025)
A fast neural hybrid Newton solver adapted to implicit methods for nonlinear dynamics
by: Jin, Tianyu, et al.
Published: (2024)
by: Jin, Tianyu, et al.
Published: (2024)
Model-Free $δ$-Policy Iteration Based on Damped Newton Method for Nonlinear Continuous-Time H$\infty$ Tracking Control
by: Wang, Qi
Published: (2024)
by: Wang, Qi
Published: (2024)
Frugality in second-order optimization: floating-point approximations for Newton's method
by: Carrino, Giuseppe, et al.
Published: (2025)
by: Carrino, Giuseppe, et al.
Published: (2025)
The Rate-Distortion-Polysemanticity Tradeoff in SAEs
by: Mencattini, Tommaso, et al.
Published: (2026)
by: Mencattini, Tommaso, et al.
Published: (2026)
A Constrained BA Algorithm for Rate-Distortion and Distortion-Rate Functions
by: Chen, Lingyi, et al.
Published: (2023)
by: Chen, Lingyi, et al.
Published: (2023)
Similar Items
-
Generalized Random Direction Newton Algorithms for Stochastic Optimization
by: Pachal, Soumen, et al.
Published: (2026) -
Risk-sensitive reinforcement learning using expectiles, shortfall risk and optimized certainty equivalent risk
by: Gupte, Sumedh, et al.
Published: (2026) -
Public Transit Arrival Prediction: a Seq2Seq RNN Approach
by: Bhutani, Nancy, et al.
Published: (2022) -
Policy Gradient Methods for Distortion Risk Measures
by: Vijayan, Nithia, et al.
Published: (2021) -
Risk Sharing, Measuring Variability, and Distortion Riskmetrics
by: Jean‐Gabriel Lauzier, et al.
Published: (2025)