Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
Fuente:
arXiv
Salvato in:
| Autore principale: | Jia, Yanwei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploratory Randomization for Discrete-Time Linear Exponential Quadratic Gaussian (LEQG) Problem
di: Lleo, Sebastien, et al.
Pubblicazione: (2025)
di: Lleo, Sebastien, et al.
Pubblicazione: (2025)
Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study
di: Huang, Yilie, et al.
Pubblicazione: (2024)
di: Huang, Yilie, et al.
Pubblicazione: (2024)
Exploratory Randomization for Discrete-Time Risk-Sensitive Benchmarked Investment Management with Reinforcement Learning
di: Lleo, Sebastien, et al.
Pubblicazione: (2026)
di: Lleo, Sebastien, et al.
Pubblicazione: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
di: Mehta, Prashant, et al.
Pubblicazione: (2026)
di: Mehta, Prashant, et al.
Pubblicazione: (2026)
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
di: Sakha, Masoud S., et al.
Pubblicazione: (2026)
di: Sakha, Masoud S., et al.
Pubblicazione: (2026)
Portfolio Optimization in a Market with Hidden Gaussian Drift and Randomly Arriving Expert Opinions: Modeling and Theoretical Results
di: Gabih, Abdelali, et al.
Pubblicazione: (2023)
di: Gabih, Abdelali, et al.
Pubblicazione: (2023)
Smart leverage? Rethinking the role of Leveraged Exchange Traded Funds in constructing portfolios to beat a benchmark
di: van Staden, Pieter, et al.
Pubblicazione: (2024)
di: van Staden, Pieter, et al.
Pubblicazione: (2024)
Power Utility Maximization with Expert Opinions at Fixed Arrival Times in a Market with Hidden Gaussian Drift
di: Gabih, Abdelali, et al.
Pubblicazione: (2023)
di: Gabih, Abdelali, et al.
Pubblicazione: (2023)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
di: Luo, Sheng, et al.
Pubblicazione: (2024)
di: Luo, Sheng, et al.
Pubblicazione: (2024)
Reflected stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
di: Liu, Lu, et al.
Pubblicazione: (2025)
di: Liu, Lu, et al.
Pubblicazione: (2025)
Optimal Feedback Control in Social Networks in a McKean-Vlasov-Friedkin-Johnsen System
di: Pramanik, Paramahansa
Pubblicazione: (2025)
di: Pramanik, Paramahansa
Pubblicazione: (2025)
Open-loop and closed-loop solvabilities for zero-sum stochastic linear quadratic differential games of Markovian regime switching system
di: Wu, Fan, et al.
Pubblicazione: (2024)
di: Wu, Fan, et al.
Pubblicazione: (2024)
Stochastic linear-quadratic differential game with Markovian jumps in an infinite horizon
di: Wu, Fan, et al.
Pubblicazione: (2024)
di: Wu, Fan, et al.
Pubblicazione: (2024)
Money-Back Tontines for Retirement Decumulation: Neural-Network Optimization under Systematic Longevity Risk
di: Orozco, German Nova, et al.
Pubblicazione: (2026)
di: Orozco, German Nova, et al.
Pubblicazione: (2026)
An optimal level of Stubbornness to win a soccer match
di: Pramanik, Paramahansa
Pubblicazione: (2025)
di: Pramanik, Paramahansa
Pubblicazione: (2025)
Path integral control under McKean-Vlasov dynamics
di: Bennett, Timothy
Pubblicazione: (2024)
di: Bennett, Timothy
Pubblicazione: (2024)
Robust optimal investment and consumption strategies with portfolio constraints and stochastic environment
di: Garces, Len Patrick Dominic M., et al.
Pubblicazione: (2024)
di: Garces, Len Patrick Dominic M., et al.
Pubblicazione: (2024)
Policy Gradient for Continuous-Time Mean-Field Control
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
Optimal dividend payout with path-dependent drawdown constraint
di: Guan, Chonghu, et al.
Pubblicazione: (2023)
di: Guan, Chonghu, et al.
Pubblicazione: (2023)
Dynamic Weight Optimization for Double Linear Policy: A Stochastic Model Predictive Control Approach
di: Hong, Tan Chin, et al.
Pubblicazione: (2026)
di: Hong, Tan Chin, et al.
Pubblicazione: (2026)
Reinforcement Learning, Optimal Control, and Bayesian Filtering in Data Assimilation
di: Hammoud, Abed
Pubblicazione: (2026)
di: Hammoud, Abed
Pubblicazione: (2026)
Sample Complexity of Policy Gradient for Log-Growth Control
di: Pan, Qiuhua, et al.
Pubblicazione: (2026)
di: Pan, Qiuhua, et al.
Pubblicazione: (2026)
Feedback strategies in the market with uncertainties
di: Issah, Mustapha Nyenye
Pubblicazione: (2024)
di: Issah, Mustapha Nyenye
Pubblicazione: (2024)
Optimal consumption under loss-averse multiplicative habit-formation preferences
di: Angoshtari, Bahman, et al.
Pubblicazione: (2024)
di: Angoshtari, Bahman, et al.
Pubblicazione: (2024)
Entropy Regularization under Bayesian Drift Uncertainty
di: Au, Andy
Pubblicazione: (2026)
di: Au, Andy
Pubblicazione: (2026)
Exploratory Mean-Variance with Jumps: An Equilibrium Approach
di: Chen, Yuling Max, et al.
Pubblicazione: (2025)
di: Chen, Yuling Max, et al.
Pubblicazione: (2025)
Continuous time Stochastic optimal control under discrete time partial observations
di: Bayer, Christian, et al.
Pubblicazione: (2024)
di: Bayer, Christian, et al.
Pubblicazione: (2024)
Robust Trading in a Generalized Lattice Market
di: Hsieh, Chung-Han, et al.
Pubblicazione: (2023)
di: Hsieh, Chung-Han, et al.
Pubblicazione: (2023)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
di: Zheng, Yaowei, et al.
Pubblicazione: (2026)
di: Zheng, Yaowei, et al.
Pubblicazione: (2026)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
di: Bai, Yitao, et al.
Pubblicazione: (2026)
di: Bai, Yitao, et al.
Pubblicazione: (2026)
Stochastic Control with Signatures
di: Bank, P., et al.
Pubblicazione: (2024)
di: Bank, P., et al.
Pubblicazione: (2024)
Dividend ratcheting and capital injection under the Cramér-Lundberg model: Strong solution and optimal strategy
di: Guan, Chonghu, et al.
Pubblicazione: (2026)
di: Guan, Chonghu, et al.
Pubblicazione: (2026)
Optimal two-parameter portfolio management strategy with transaction costs
di: Ma, Chutian, et al.
Pubblicazione: (2024)
di: Ma, Chutian, et al.
Pubblicazione: (2024)
A measure-valued HJB perspective on Bayesian optimal adaptive control
di: Cox, Alexander M. G., et al.
Pubblicazione: (2025)
di: Cox, Alexander M. G., et al.
Pubblicazione: (2025)
Optimal control of SDEs with merely measurable drift: an HJB approach
di: Du, Kai, et al.
Pubblicazione: (2025)
di: Du, Kai, et al.
Pubblicazione: (2025)
On the Rate of Gaussian Approximation for Linear Regression Problems
di: Khusainov, Marat, et al.
Pubblicazione: (2025)
di: Khusainov, Marat, et al.
Pubblicazione: (2025)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
An Optimal-Control Approach to Infinite-Horizon Restless Bandits: Achieving Asymptotic Optimality with Minimal Assumptions
di: YAN, Chen
Pubblicazione: (2024)
di: YAN, Chen
Pubblicazione: (2024)
Convergence of Neural Network Policies for Risk--Reward Optimization
di: Chen, Chang, et al.
Pubblicazione: (2026)
di: Chen, Chang, et al.
Pubblicazione: (2026)
Robust Probability Hypothesis Density Filtering: Theory and Algorithms
di: Lei, Ming, et al.
Pubblicazione: (2025)
di: Lei, Ming, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Exploratory Randomization for Discrete-Time Linear Exponential Quadratic Gaussian (LEQG) Problem
di: Lleo, Sebastien, et al.
Pubblicazione: (2025) -
Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study
di: Huang, Yilie, et al.
Pubblicazione: (2024) -
Exploratory Randomization for Discrete-Time Risk-Sensitive Benchmarked Investment Management with Reinforcement Learning
di: Lleo, Sebastien, et al.
Pubblicazione: (2026) -
Optimistic Training and Convergence of Q-Learning -- Extended Version
di: Mehta, Prashant, et al.
Pubblicazione: (2026) -
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
di: Sakha, Masoud S., et al.
Pubblicazione: (2026)