Delayed homomorphic reinforcement learning for environments with delayed feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Jongsoo, Kim, Jangwon, Han, Soohee |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning via Conservative Agent for Environments with Random Delays
by: Lee, Jongsoo, et al.
Published: (2025)
by: Lee, Jongsoo, et al.
Published: (2025)
Deep reinforcement learning for irrigation scheduling using high-dimensional sensor feedback
by: Saikai, Yuji, et al.
Published: (2023)
by: Saikai, Yuji, et al.
Published: (2023)
Learning to summarize user information for personalized reinforcement learning from human feedback
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
An efficient deep reinforcement learning environment for flexible job-shop scheduling
by: Wu, Xinquan, et al.
Published: (2025)
by: Wu, Xinquan, et al.
Published: (2025)
Using reinforcement learning to probe the role of feedback in skill acquisition
by: Terpin, Antonio, et al.
Published: (2025)
by: Terpin, Antonio, et al.
Published: (2025)
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)
by: Lee, Sunbowen, et al.
Published: (2025)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Curriculum reinforcement learning with measurable task representation learning
by: Wen, Yongyan, et al.
Published: (2026)
by: Wen, Yongyan, et al.
Published: (2026)
RedactOR: An LLM-Powered Framework for Automatic Clinical Data De-Identification
by: Singh, Praphul, et al.
Published: (2025)
by: Singh, Praphul, et al.
Published: (2025)
Interpretable experiential learning based on state history and global feedback
by: Kolonin, Anton
Published: (2026)
by: Kolonin, Anton
Published: (2026)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Deep reinforcement learning with time-scale invariant memory
by: Kabir, Md Rysul, et al.
Published: (2024)
by: Kabir, Md Rysul, et al.
Published: (2024)
Offline reinforcement learning for job-shop scheduling problems
by: Echeverria, Imanol, et al.
Published: (2024)
by: Echeverria, Imanol, et al.
Published: (2024)
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025)
by: Kadurha, David Krame, et al.
Published: (2025)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
by: Yeom, Junghyuk, et al.
Published: (2024)
by: Yeom, Junghyuk, et al.
Published: (2024)
Not all tokens are needed(NAT): token efficient reinforcement learning
by: Sang, Hejian, et al.
Published: (2026)
by: Sang, Hejian, et al.
Published: (2026)
Leveraging weights signals -- Predicting and improving generalizability in reinforcement learning
by: Moulin, Olivier, et al.
Published: (2025)
by: Moulin, Olivier, et al.
Published: (2025)
Dynamic feature selection in medical predictive monitoring by reinforcement learning
by: Chen, Yutong, et al.
Published: (2024)
by: Chen, Yutong, et al.
Published: (2024)
Economic span selection of bridge based on deep reinforcement learning
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Improving out-of-distribution generalization in graphs via hierarchical semantic environments
by: Piao, Yinhua, et al.
Published: (2024)
by: Piao, Yinhua, et al.
Published: (2024)
Task diversity produces systematic transfer but inhibits continual reinforcement learning
by: Seth, Purab, et al.
Published: (2026)
by: Seth, Purab, et al.
Published: (2026)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving
by: Qu, Yansong, et al.
Published: (2026)
by: Qu, Yansong, et al.
Published: (2026)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
by: Saghafian, Armin, et al.
Published: (2024)
by: Saghafian, Armin, et al.
Published: (2024)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
by: Hutson, Miles, et al.
Published: (2024)
by: Hutson, Miles, et al.
Published: (2024)
An advantage based policy transfer algorithm for reinforcement learning with measures of transferability
by: Alam, Md Ferdous, et al.
Published: (2023)
by: Alam, Md Ferdous, et al.
Published: (2023)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
by: Kobayashi, Seijin, et al.
Published: (2025)
by: Kobayashi, Seijin, et al.
Published: (2025)
Survey on reinforcement learning for language processing
by: Uc-Cetina, Victor, et al.
Published: (2021)
by: Uc-Cetina, Victor, et al.
Published: (2021)
Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs
by: Gan, Xingwei, et al.
Published: (2026)
by: Gan, Xingwei, et al.
Published: (2026)
CHARME: A chain-based reinforcement learning approach for the minor embedding problem
by: Ngo, Hoang M., et al.
Published: (2024)
by: Ngo, Hoang M., et al.
Published: (2024)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Designing an efficient and equitable humanitarian supply chain dynamically via reinforcement learning
by: Jin, Weijia
Published: (2025)
by: Jin, Weijia
Published: (2025)
Convergence of a model-free entropy-regularized inverse reinforcement learning algorithm
by: Renard, Titouan, et al.
Published: (2024)
by: Renard, Titouan, et al.
Published: (2024)
Maximum diffusion reinforcement learning
by: Berrueta, Thomas A., et al.
Published: (2023)
by: Berrueta, Thomas A., et al.
Published: (2023)
Simulation-based reinforcement learning for real-world autonomous driving
by: Osiński, Błażej, et al.
Published: (2019)
by: Osiński, Błażej, et al.
Published: (2019)
Designing a double deep reinforcement learning selection tool for resilient demand prediction
by: Benziane, Bilel Abderrahmane, et al.
Published: (2026)
by: Benziane, Bilel Abderrahmane, et al.
Published: (2026)
Estimating unknown parameters in differential equations with a reinforcement learning based PSO method
by: Sun, Wenkui, et al.
Published: (2024)
by: Sun, Wenkui, et al.
Published: (2024)
Current applications and potential future directions of reinforcement learning-based Digital Twins in agriculture
by: Goldenits, Georg, et al.
Published: (2024)
by: Goldenits, Georg, et al.
Published: (2024)
Bridging the phenotype-target gap for molecular generation via multi-objective reinforcement learning
by: Guo, Haotian, et al.
Published: (2025)
by: Guo, Haotian, et al.
Published: (2025)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Similar Items
-
Reinforcement Learning via Conservative Agent for Environments with Random Delays
by: Lee, Jongsoo, et al.
Published: (2025) -
Deep reinforcement learning for irrigation scheduling using high-dimensional sensor feedback
by: Saikai, Yuji, et al.
Published: (2023) -
Learning to summarize user information for personalized reinforcement learning from human feedback
by: Nam, Hyunji, et al.
Published: (2025) -
An efficient deep reinforcement learning environment for flexible job-shop scheduling
by: Wu, Xinquan, et al.
Published: (2025) -
Using reinforcement learning to probe the role of feedback in skill acquisition
by: Terpin, Antonio, et al.
Published: (2025)