Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Surdej, Rafał, Bortkiewicz, Michał, Lewandowski, Alex, Ostaszewski, Mateusz, Lyle, Clare |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Continually by Spectral Regularization
von: Lewandowski, Alex, et al.
Veröffentlicht: (2024)
von: Lewandowski, Alex, et al.
Veröffentlicht: (2024)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem
von: Wołczyk, Maciej, et al.
Veröffentlicht: (2024)
von: Wołczyk, Maciej, et al.
Veröffentlicht: (2024)
A Case for Validation Buffer in Pessimistic Actor-Critic
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
On the Role of Iterative Computation in Reinforcement Learning
von: Ghugare, Raj, et al.
Veröffentlicht: (2026)
von: Ghugare, Raj, et al.
Veröffentlicht: (2026)
Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
Adaptive Rational Activations to Boost Deep Reinforcement Learning
von: Delfosse, Quentin, et al.
Veröffentlicht: (2021)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2021)
Contrastive Representations for Temporal Reasoning
von: Ziarko, Alicja, et al.
Veröffentlicht: (2025)
von: Ziarko, Alicja, et al.
Veröffentlicht: (2025)
Frequency and Generalisation of Periodic Activation Functions in Reinforcement Learning
von: Mavor-Parker, Augustine N., et al.
Veröffentlicht: (2024)
von: Mavor-Parker, Augustine N., et al.
Veröffentlicht: (2024)
Rational Neural Networks have Expressivity Advantages
von: Tang, Maosen, et al.
Veröffentlicht: (2026)
von: Tang, Maosen, et al.
Veröffentlicht: (2026)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
Weight Clipping for Deep Continual and Reinforcement Learning
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
Reinforcement Teaching
von: Muslimani, Calarina, et al.
Veröffentlicht: (2022)
von: Muslimani, Calarina, et al.
Veröffentlicht: (2022)
Plastic Learning with Deep Fourier Features
von: Lewandowski, Alex, et al.
Veröffentlicht: (2024)
von: Lewandowski, Alex, et al.
Veröffentlicht: (2024)
On The Expressivity of Objective-Specification Formalisms in Reinforcement Learning
von: Subramani, Rohan, et al.
Veröffentlicht: (2023)
von: Subramani, Rohan, et al.
Veröffentlicht: (2023)
PhysGym: Benchmarking LLMs in Interactive Physics Discovery with Controlled Priors
von: Chen, Yimeng, et al.
Veröffentlicht: (2025)
von: Chen, Yimeng, et al.
Veröffentlicht: (2025)
Unpacking Softmax: How Temperature Drives Representation Collapse, Compression, and Generalization
von: Masarczyk, Wojciech, et al.
Veröffentlicht: (2025)
von: Masarczyk, Wojciech, et al.
Veröffentlicht: (2025)
Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model
von: Rowland, Mark, et al.
Veröffentlicht: (2024)
von: Rowland, Mark, et al.
Veröffentlicht: (2024)
Learning Expressive Random Feature Models via Parametrized Activations
von: Ma, Zailin, et al.
Veröffentlicht: (2024)
von: Ma, Zailin, et al.
Veröffentlicht: (2024)
Distributionally Robust Constrained Reinforcement Learning under Strong Duality
von: Zhang, Zhengfei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengfei, et al.
Veröffentlicht: (2024)
Safe and Balanced: A Framework for Constrained Multi-Objective Reinforcement Learning
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
Reduced storage direct tensor ring decomposition for convolutional neural networks compression
von: Gabor, Mateusz, et al.
Veröffentlicht: (2024)
von: Gabor, Mateusz, et al.
Veröffentlicht: (2024)
Learning to Forget: Continual Learning with Adaptive Weight Decay
von: Ramesh, Aditya A., et al.
Veröffentlicht: (2026)
von: Ramesh, Aditya A., et al.
Veröffentlicht: (2026)
EXPO: Stable Reinforcement Learning with Expressive Policies
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
Level Generation with Constrained Expressive Range
von: Bazzaz, Mahsa, et al.
Veröffentlicht: (2025)
von: Bazzaz, Mahsa, et al.
Veröffentlicht: (2025)
UAMDP: Uncertainty-Aware Markov Decision Process for Risk-Constrained Reinforcement Learning from Probabilistic Forecasts
von: Koren, Michal, et al.
Veröffentlicht: (2025)
von: Koren, Michal, et al.
Veröffentlicht: (2025)
On the Expressiveness of Rational ReLU Neural Networks With Bounded Depth
von: Averkov, Gennadiy, et al.
Veröffentlicht: (2025)
von: Averkov, Gennadiy, et al.
Veröffentlicht: (2025)
Rationality Measurement and Theory for Reinforcement Learning Agents
von: Qian, Kejiang, et al.
Veröffentlicht: (2026)
von: Qian, Kejiang, et al.
Veröffentlicht: (2026)
Accelerating Goal-Conditioned RL Algorithms and Research
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2024)
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2024)
What Can Grokking Teach Us About Learning Under Nonstationarity?
von: Lyle, Clare, et al.
Veröffentlicht: (2025)
von: Lyle, Clare, et al.
Veröffentlicht: (2025)
Federated Learning by Utility-Constrained Stochastic Aggregation for Improving Rational Participation
von: Yashwanth, M, et al.
Veröffentlicht: (2026)
von: Yashwanth, M, et al.
Veröffentlicht: (2026)
What Ails Generative Structure-based Drug Design: Expressivity is Too Little or Too Much?
von: Karczewski, Rafał, et al.
Veröffentlicht: (2024)
von: Karczewski, Rafał, et al.
Veröffentlicht: (2024)
Robust off-policy Reinforcement Learning via Soft Constrained Adversary
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2024)
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2024)
Implicit Language Models are RNNs: Balancing Parallelization and Expressivity
von: Schöne, Mark, et al.
Veröffentlicht: (2025)
von: Schöne, Mark, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Efficient Toxicity Detection in Competitive Online Video Games
von: Morrier, Jacob, et al.
Veröffentlicht: (2025)
von: Morrier, Jacob, et al.
Veröffentlicht: (2025)
Expressive Value Learning for Scalable Offline Reinforcement Learning
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
Rectified Robust Policy Optimization for Model-Uncertain Constrained Reinforcement Learning without Strong Duality
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
TCRL: Temporal-Coupled Adversarial Training for Robust Constrained Reinforcement Learning in Worst-Case Scenarios
von: Xu, Wentao, et al.
Veröffentlicht: (2026)
von: Xu, Wentao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning Continually by Spectral Regularization
von: Lewandowski, Alex, et al.
Veröffentlicht: (2024) -
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025) -
Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
von: Nauman, Michal, et al.
Veröffentlicht: (2024) -
Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem
von: Wołczyk, Maciej, et al.
Veröffentlicht: (2024) -
A Case for Validation Buffer in Pessimistic Actor-Critic
von: Nauman, Michal, et al.
Veröffentlicht: (2024)