Saved in:
| Main Author: | Yang, Lingyi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.20538 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural Controlled Differential Equations with Quantum Hidden Evolutions
by: Yang, Lingyi, et al.
Published: (2024)
by: Yang, Lingyi, et al.
Published: (2024)
Universal hidden monotonic trend estimation with contrastive learning
by: Pineau, Edouard, et al.
Published: (2022)
by: Pineau, Edouard, et al.
Published: (2022)
Regularized Q-learning
by: Lim, Han-Dong, et al.
Published: (2022)
by: Lim, Han-Dong, et al.
Published: (2022)
On the representation and learning of monotone triangular transport maps
by: Baptista, Ricardo, et al.
Published: (2020)
by: Baptista, Ricardo, et al.
Published: (2020)
HyperQ-Opt: Q-learning for Hyperparameter Optimization
by: Hasan, Md. Tarek
Published: (2024)
by: Hasan, Md. Tarek
Published: (2024)
VA-learning as a more efficient alternative to Q-learning
by: Tang, Yunhao, et al.
Published: (2023)
by: Tang, Yunhao, et al.
Published: (2023)
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
by: Zhang, Jing, et al.
Published: (2024)
by: Zhang, Jing, et al.
Published: (2024)
Transfer Q-learning
by: Chen, Elynn, et al.
Published: (2022)
by: Chen, Elynn, et al.
Published: (2022)
Q-learning with Posterior Sampling
by: Agrawal, Priyank, et al.
Published: (2025)
by: Agrawal, Priyank, et al.
Published: (2025)
Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning
by: Manda, Kausthubh, et al.
Published: (2025)
by: Manda, Kausthubh, et al.
Published: (2025)
Generalizing while preserving monotonicity in comparison-based preference learning models
by: Fageot, Julien, et al.
Published: (2025)
by: Fageot, Julien, et al.
Published: (2025)
Optimistic Q-learning for average reward and episodic reinforcement learning
by: Agrawal, Priyank, et al.
Published: (2024)
by: Agrawal, Priyank, et al.
Published: (2024)
Faithful Embeddings of Irregular and Asynchronous Data for Online Log-NCDEs
by: Walker, Benjamin, et al.
Published: (2026)
by: Walker, Benjamin, et al.
Published: (2026)
Deep Double Q-learning
by: Nagarajan, Prabhat, et al.
Published: (2025)
by: Nagarajan, Prabhat, et al.
Published: (2025)
Gaussian Approximation for Asynchronous Q-learning
by: Rubtsov, Artemy, et al.
Published: (2026)
by: Rubtsov, Artemy, et al.
Published: (2026)
Top-$K$ ranking with a monotone adversary
by: Yang, Yuepeng, et al.
Published: (2024)
by: Yang, Yuepeng, et al.
Published: (2024)
Gap-Dependent Bounds for Federated $Q$-learning
by: Zhang, Haochen, et al.
Published: (2025)
by: Zhang, Haochen, et al.
Published: (2025)
Counterfactual identifiability beyond global monotonicity: non-monotone triangular structural causal models
by: Tan, Pengcheng, et al.
Published: (2026)
by: Tan, Pengcheng, et al.
Published: (2026)
Structured Linear CDEs: Maximally Expressive and Parallel-in-Time Sequence Models
by: Walker, Benjamin, et al.
Published: (2025)
by: Walker, Benjamin, et al.
Published: (2025)
Q-learning with Adjoint Matching
by: Li, Qiyang, et al.
Published: (2026)
by: Li, Qiyang, et al.
Published: (2026)
World Model-Enabled Causal Digital Twins for Semantic Communications in Physical AI Systems
by: Wang, Lingyi, et al.
Published: (2026)
by: Wang, Lingyi, et al.
Published: (2026)
MinMaxMin $Q$-learning
by: Soffair, Nitsan, et al.
Published: (2024)
by: Soffair, Nitsan, et al.
Published: (2024)
Is Q-learning an Ill-posed Problem?
by: Wissmann, Philipp, et al.
Published: (2025)
by: Wissmann, Philipp, et al.
Published: (2025)
Periodic agent-state based Q-learning for POMDPs
by: Sinha, Amit, et al.
Published: (2024)
by: Sinha, Amit, et al.
Published: (2024)
Maximum entropy GFlowNets with soft Q-learning
by: Mohammadpour, Sobhan, et al.
Published: (2023)
by: Mohammadpour, Sobhan, et al.
Published: (2023)
Lexicographic optimization-based approaches to learning a representative model for multi-criteria sorting with non-monotonic criteria
by: Zhang, Zhen, et al.
Published: (2024)
by: Zhang, Zhen, et al.
Published: (2024)
Near-Equivalent Q-learning Policies for Dynamic Treatment Regimes
by: Yazzourh, Sophia, et al.
Published: (2026)
by: Yazzourh, Sophia, et al.
Published: (2026)
Convergence of regularized agent-state-based Q-learning in POMDPs
by: Sinha, Amit, et al.
Published: (2025)
by: Sinha, Amit, et al.
Published: (2025)
On Gaussian approximation for entropy-regularized Q-learning with function approximation
by: Rubtsov, Artemy, et al.
Published: (2026)
by: Rubtsov, Artemy, et al.
Published: (2026)
Q-learning with temporal memory to navigate turbulence
by: Rando, Marco, et al.
Published: (2024)
by: Rando, Marco, et al.
Published: (2024)
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024)
by: Schmitt-Förster, Peter, et al.
Published: (2024)
Stabilizing Extreme Q-learning by Maclaurin Expansion
by: Omura, Motoki, et al.
Published: (2024)
by: Omura, Motoki, et al.
Published: (2024)
Asymptotic Analysis of Sample-averaged Q-learning
by: Panda, Saunak Kumar, et al.
Published: (2024)
by: Panda, Saunak Kumar, et al.
Published: (2024)
Tensor-Efficient High-Dimensional Q-learning
by: Wu, Junyi, et al.
Published: (2025)
by: Wu, Junyi, et al.
Published: (2025)
A general learning scheme for classical and quantum Ising machines
by: Schmid, Ludwig, et al.
Published: (2023)
by: Schmid, Ludwig, et al.
Published: (2023)
DMWM: Dual-Mind World Model with Long-Term Imagination
by: Wang, Lingyi, et al.
Published: (2025)
by: Wang, Lingyi, et al.
Published: (2025)
Dual-Mind World Models: A General Framework for Learning in Dynamic Wireless Networks
by: Wang, Lingyi, et al.
Published: (2025)
by: Wang, Lingyi, et al.
Published: (2025)
Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization
by: Bonnerjee, Soham, et al.
Published: (2026)
by: Bonnerjee, Soham, et al.
Published: (2026)
Contextualized Hybrid Ensemble Q-learning: Learning Fast with Control Priors
by: Cramer, Emma, et al.
Published: (2024)
by: Cramer, Emma, et al.
Published: (2024)
Q-learning for Quantile MDPs: A Decomposition, Performance, and Convergence Analysis
by: Hau, Jia Lin, et al.
Published: (2024)
by: Hau, Jia Lin, et al.
Published: (2024)
Similar Items
-
Neural Controlled Differential Equations with Quantum Hidden Evolutions
by: Yang, Lingyi, et al.
Published: (2024) -
Universal hidden monotonic trend estimation with contrastive learning
by: Pineau, Edouard, et al.
Published: (2022) -
Regularized Q-learning
by: Lim, Han-Dong, et al.
Published: (2022) -
On the representation and learning of monotone triangular transport maps
by: Baptista, Ricardo, et al.
Published: (2020) -
HyperQ-Opt: Q-learning for Hyperparameter Optimization
by: Hasan, Md. Tarek
Published: (2024)