MetaCURL: Non-stationary Concave Utility Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Moreno, Bianca Marin, Brégère, Margaux, Gaillard, Pierre, Oudjane, Nadia |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Episodic Convex Reinforcement Learning
by: Moreno, Bianca Marin, et al.
Published: (2025)
by: Moreno, Bianca Marin, et al.
Published: (2025)
Online Markov Decision Processes with Terminal Law Constraints
by: Moreno, Bianca Marin, et al.
Published: (2026)
by: Moreno, Bianca Marin, et al.
Published: (2026)
Randomized Midpoint Method for Log-Concave Sampling under Constraints
by: Yu, Yifeng, et al.
Published: (2024)
by: Yu, Yifeng, et al.
Published: (2024)
Zeroth-Order Sampling Methods for Non-Log-Concave Distributions: Alleviating Metastability by Denoising Diffusion
by: He, Ye, et al.
Published: (2024)
by: He, Ye, et al.
Published: (2024)
Revenue Maximization Under Sequential Price Competition Via The Estimation Of s-Concave Demand Functions
by: Bracale, Daniele, et al.
Published: (2025)
by: Bracale, Daniele, et al.
Published: (2025)
Poisson Midpoint Method for Log Concave Sampling: Beyond the Strong Error Lower Bounds
by: Srinivasan, Rishikesh, et al.
Published: (2025)
by: Srinivasan, Rishikesh, et al.
Published: (2025)
Estimating stationary mass, frequency by frequency
by: Nakul, Milind, et al.
Published: (2025)
by: Nakul, Milind, et al.
Published: (2025)
Online Learning Approach for Survival Analysis
by: Fernandez, Camila, et al.
Published: (2024)
by: Fernandez, Camila, et al.
Published: (2024)
Structured Prediction in Online Learning
by: Boudart, Pierre, et al.
Published: (2024)
by: Boudart, Pierre, et al.
Published: (2024)
Bayesian Inference with Shaped Deep Non-linear MLPs
by: Hanin, Boris, et al.
Published: (2026)
by: Hanin, Boris, et al.
Published: (2026)
Non-Asymptotic Analysis of Data Augmentation for Precision Matrix Estimation
by: Morisset, Lucas, et al.
Published: (2025)
by: Morisset, Lucas, et al.
Published: (2025)
Learning with Expected Signatures: Theory and Applications
by: Lucchese, Lorenzo, et al.
Published: (2025)
by: Lucchese, Lorenzo, et al.
Published: (2025)
On Universality of Non-Separable Approximate Message Passing Algorithms
by: Lovig, Max, et al.
Published: (2025)
by: Lovig, Max, et al.
Published: (2025)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
by: Boudart, Pierre, et al.
Published: (2025)
by: Boudart, Pierre, et al.
Published: (2025)
Minimax Rates for Learning Pairwise Interactions in Attention-Style Models
by: Zucker, Shai, et al.
Published: (2025)
by: Zucker, Shai, et al.
Published: (2025)
Forecast collapse of transformer-based models under squared loss in financial time series
by: Andreoletti, Pierre
Published: (2026)
by: Andreoletti, Pierre
Published: (2026)
Optimal Exact Recovery in Semi-Supervised Learning: A Study of Spectral Methods and Graph Convolutional Networks
by: Wang, Hai-Xiao, et al.
Published: (2024)
by: Wang, Hai-Xiao, et al.
Published: (2024)
Non-asymptotic estimates for accelerated high order Langevin Monte Carlo algorithms
by: Neufeld, Ariel, et al.
Published: (2024)
by: Neufeld, Ariel, et al.
Published: (2024)
Information-Theoretic Limits and Strong Consistency on Binary Non-uniform Hypergraph Stochastic Block Models
by: Wang, Hai-Xiao
Published: (2023)
by: Wang, Hai-Xiao
Published: (2023)
Statistical limits of correlation detection in trees
by: Ganassali, Luca, et al.
Published: (2022)
by: Ganassali, Luca, et al.
Published: (2022)
Sampling from multimodal distributions with warm starts: Non-asymptotic bounds for the Reweighted Annealed Leap-Point Sampler
by: Lee, Holden, et al.
Published: (2025)
by: Lee, Holden, et al.
Published: (2025)
Concave Statistical Utility Maximization Bandits via Influence-Function Gradients
by: Carrasco, Matías, et al.
Published: (2026)
by: Carrasco, Matías, et al.
Published: (2026)
Probabilistic Inference and Learning with Stein's Method
by: Liu, Qiang, et al.
Published: (2026)
by: Liu, Qiang, et al.
Published: (2026)
Non-ergodic inference for stationary-increment harmonizable stable processes
by: Hoang, Ly Viet, et al.
Published: (2024)
by: Hoang, Ly Viet, et al.
Published: (2024)
Simple Relative Deviation Bounds for Covariance and Gram Matrices
by: Barzilai, Daniel, et al.
Published: (2024)
by: Barzilai, Daniel, et al.
Published: (2024)
Convergence Bounds for Sequential Monte Carlo on Multimodal Distributions using Soft Decomposition
by: Lee, Holden, et al.
Published: (2024)
by: Lee, Holden, et al.
Published: (2024)
Optimization, Isoperimetric Inequalities, and Sampling via Lyapunov Potentials
by: Chen, August Y., et al.
Published: (2024)
by: Chen, August Y., et al.
Published: (2024)
Multivariate Gaussian Approximation for Random Forest via Region-based Stabilization
by: Shi, Zhaoyang, et al.
Published: (2024)
by: Shi, Zhaoyang, et al.
Published: (2024)
A new approach for imprecise probabilities
by: Basili, Marcello, et al.
Published: (2024)
by: Basili, Marcello, et al.
Published: (2024)
Sharp bounds on aggregate expert error
by: Kontorovich, Aryeh, et al.
Published: (2024)
by: Kontorovich, Aryeh, et al.
Published: (2024)
Universality of Kernel Random Matrices and Kernel Regression in the Quadratic Regime
by: Pandit, Parthe, et al.
Published: (2024)
by: Pandit, Parthe, et al.
Published: (2024)
Dynamic Structural Causal Models
by: Boeken, Philip, et al.
Published: (2024)
by: Boeken, Philip, et al.
Published: (2024)
Central Limit Theorem for Bayesian Neural Network trained with Variational Inference
by: Descours, Arnaud, et al.
Published: (2024)
by: Descours, Arnaud, et al.
Published: (2024)
WHOMP: Optimizing Randomized Controlled Trials via Wasserstein Homogeneity
by: Xu, Shizhou, et al.
Published: (2024)
by: Xu, Shizhou, et al.
Published: (2024)
Analysis of a multi-target linear shrinkage covariance estimator
by: Oriol, Benoit
Published: (2024)
by: Oriol, Benoit
Published: (2024)
Exponential tilting of subweibull distributions
by: Townes, F. William
Published: (2024)
by: Townes, F. William
Published: (2024)
Asymptotic spectrum of weighted sample covariance: another proof of spectrum convergence
by: Oriol, Benoit
Published: (2024)
by: Oriol, Benoit
Published: (2024)
Improved Finite-Particle Convergence Rates for Stein Variational Gradient Descent
by: Banerjee, Sayan, et al.
Published: (2024)
by: Banerjee, Sayan, et al.
Published: (2024)
Are Bayesian networks typically faithful?
by: Boeken, Philip, et al.
Published: (2024)
by: Boeken, Philip, et al.
Published: (2024)
An Analysis of Elo Rating Systems via Markov Chains
by: Olesker-Taylor, Sam, et al.
Published: (2024)
by: Olesker-Taylor, Sam, et al.
Published: (2024)
Similar Items
-
Online Episodic Convex Reinforcement Learning
by: Moreno, Bianca Marin, et al.
Published: (2025) -
Online Markov Decision Processes with Terminal Law Constraints
by: Moreno, Bianca Marin, et al.
Published: (2026) -
Randomized Midpoint Method for Log-Concave Sampling under Constraints
by: Yu, Yifeng, et al.
Published: (2024) -
Zeroth-Order Sampling Methods for Non-Log-Concave Distributions: Alleviating Metastability by Denoising Diffusion
by: He, Ye, et al.
Published: (2024) -
Revenue Maximization Under Sequential Price Competition Via The Estimation Of s-Concave Demand Functions
by: Bracale, Daniele, et al.
Published: (2025)