Policy Newton Algorithm in Reproducing Kernel Hilbert Space
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yixian, Tang, Huaze, Wang, Chao, Ding, Wenbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
by: Zhang, Yixian, et al.
Published: (2025)
by: Zhang, Yixian, et al.
Published: (2025)
Learning Reconstructive Embeddings in Reproducing Kernel Hilbert Spaces via the Representer Theorem
by: Feito-Casares, Enrique, et al.
Published: (2026)
by: Feito-Casares, Enrique, et al.
Published: (2026)
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026)
by: Zixian, Wang
Published: (2026)
Group Orthogonalized Policy Optimization:Group Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026)
by: Zixian, Wang
Published: (2026)
Quantum Policy Gradient in Reproducing Kernel Hilbert Space
by: Bossens, David M., et al.
Published: (2024)
by: Bossens, David M., et al.
Published: (2024)
Beyond Expected Return: Accounting for Policy Reproducibility when Evaluating Reinforcement Learning Algorithms
by: Flageat, Manon, et al.
Published: (2023)
by: Flageat, Manon, et al.
Published: (2023)
Controlling False Discovery in Arbitrarily Structured Hypothesis Spaces via Reproducing Kernels
by: Perets, Binyamin, et al.
Published: (2026)
by: Perets, Binyamin, et al.
Published: (2026)
A Variance-Reduced Cubic-Regularized Newton for Policy Optimization
by: Sun, Cheng, et al.
Published: (2025)
by: Sun, Cheng, et al.
Published: (2025)
RUMAD: Reinforcement-Unifying Multi-Agent Debate
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
From Sorting Algorithms to Scalable Kernels: Bayesian Optimization in High-Dimensional Permutation Spaces
by: Xie, Zikai, et al.
Published: (2025)
by: Xie, Zikai, et al.
Published: (2025)
Policy Gradient with Kernel Quadrature
by: Hayakawa, Satoshi, et al.
Published: (2023)
by: Hayakawa, Satoshi, et al.
Published: (2023)
Foundation Policies with Hilbert Representations
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
DualWeaver: Synergistic Feature Weaving Surrogates for Multivariate Forecasting with Univariate Time Series Foundation Models
by: Li, Jinpeng, et al.
Published: (2026)
by: Li, Jinpeng, et al.
Published: (2026)
Function-Space Empirical Bayes Regularisation with Large Vision-Language Model Priors
by: Hao, Pengcheng, et al.
Published: (2026)
by: Hao, Pengcheng, et al.
Published: (2026)
Algorithms for Learning Kernels Based on Centered Alignment
by: Cortes, Corinna, et al.
Published: (2012)
by: Cortes, Corinna, et al.
Published: (2012)
Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies
by: Lee, Haanvid, et al.
Published: (2024)
by: Lee, Haanvid, et al.
Published: (2024)
Quotient-Space Diffusion Models
by: Xu, Yixian, et al.
Published: (2026)
by: Xu, Yixian, et al.
Published: (2026)
On the Convergence of Irregular Sampling in Reproducing Kernel Hilbert Spaces
by: Iske, Armin
Published: (2025)
by: Iske, Armin
Published: (2025)
Learning of Hamiltonian Dynamics with Reproducing Kernel Hilbert Spaces
by: Smith, Torbjørn, et al.
Published: (2023)
by: Smith, Torbjørn, et al.
Published: (2023)
Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods
by: Störck, Felix, et al.
Published: (2025)
by: Störck, Felix, et al.
Published: (2025)
IFNSO: Iteration-Free Newton-Schulz Orthogonalization
by: Hu, Chen, et al.
Published: (2026)
by: Hu, Chen, et al.
Published: (2026)
Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
Kernel VICReg for Self-Supervised Learning in Reproducing Kernel Hilbert Space
by: Sepanj, M. Hadi, et al.
Published: (2025)
by: Sepanj, M. Hadi, et al.
Published: (2025)
Reproducibility study of FairAC
by: de Jong, Gijs, et al.
Published: (2024)
by: de Jong, Gijs, et al.
Published: (2024)
Efficient On-Policy Reinforcement Learning via Exploration of Sparse Parameter Space
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
On Quasi-Localized Dual Pairs in Reproducing Kernel Hilbert Spaces
by: Harbrecht, Helmut, et al.
Published: (2024)
by: Harbrecht, Helmut, et al.
Published: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
by: Vakili, Sattar, et al.
Published: (2024)
by: Vakili, Sattar, et al.
Published: (2024)
On the Replicability and Reproducibility of Deep Learning in Software Engineering
by: Liu, Chao, et al.
Published: (2020)
by: Liu, Chao, et al.
Published: (2020)
ExO-PPO: an Extended Off-policy Proximal Policy Optimization Algorithm
by: Wang, Hanyong, et al.
Published: (2026)
by: Wang, Hanyong, et al.
Published: (2026)
A Geometry-Aware Algorithm to Learn Hierarchical Embeddings in Hyperbolic Space
by: Wang, Zhangyu, et al.
Published: (2024)
by: Wang, Zhangyu, et al.
Published: (2024)
Learning in Feature Spaces via Coupled Covariances: Asymmetric Kernel SVD and Nyström method
by: Tao, Qinghua, et al.
Published: (2024)
by: Tao, Qinghua, et al.
Published: (2024)
Gradient Multi-Normalization for Stateless and Scalable LLM Training
by: Scetbon, Meyer, et al.
Published: (2025)
by: Scetbon, Meyer, et al.
Published: (2025)
Towards Efficient Optimizer Design for LLM via Structured Fisher Approximation with a Low-Rank Extension
by: Gong, Wenbo, et al.
Published: (2025)
by: Gong, Wenbo, et al.
Published: (2025)
SWAN: SGD with Normalization and Whitening Enables Stateless LLM Training
by: Ma, Chao, et al.
Published: (2024)
by: Ma, Chao, et al.
Published: (2024)
Single-stream Policy Optimization
by: Xu, Zhongwen, et al.
Published: (2025)
by: Xu, Zhongwen, et al.
Published: (2025)
Scalable Policy-Based RL Algorithms for POMDPs
by: Anjarlekar, Ameya, et al.
Published: (2025)
by: Anjarlekar, Ameya, et al.
Published: (2025)
MASP: Scalable GNN-based Planning for Multi-Agent Navigation
by: Yang, Xinyi, et al.
Published: (2023)
by: Yang, Xinyi, et al.
Published: (2023)
MobileKernelBench: Can LLMs Write Efficient Kernels for Mobile Devices?
by: Zou, Xingze, et al.
Published: (2026)
by: Zou, Xingze, et al.
Published: (2026)
Statistical Analysis of Policy Space Compression Problem
by: Molaei, Majid, et al.
Published: (2024)
by: Molaei, Majid, et al.
Published: (2024)
Similar Items
-
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
by: Zhang, Yixian, et al.
Published: (2025) -
Learning Reconstructive Embeddings in Reproducing Kernel Hilbert Spaces via the Representer Theorem
by: Feito-Casares, Enrique, et al.
Published: (2026) -
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026) -
Group Orthogonalized Policy Optimization:Group Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026) -
Quantum Policy Gradient in Reproducing Kernel Hilbert Space
by: Bossens, David M., et al.
Published: (2024)