Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
Fuente:
arXiv
Saved in:
| Main Authors: | Sorstkins, Andrejs, Tariq, Omer, Bilal, Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025)
by: Sorstkins, Andrejs
Published: (2025)
Quantum-Inspired Reinforcement Learning for Secure and Sustainable AIoT-Driven Supply Chain Systems
by: Dastagir, Muhammad Bilal Akram, et al.
Published: (2026)
by: Dastagir, Muhammad Bilal Akram, et al.
Published: (2026)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
Diagnostics of cognitive failures in multi-agent expert systems using dynamic evaluation protocols and subsequent mutation of the processing context
by: Sorstkins, Andrejs, et al.
Published: (2025)
by: Sorstkins, Andrejs, et al.
Published: (2025)
Rollback-Free Stable Brick Structures Generation
by: Xu, Chenhui, et al.
Published: (2026)
by: Xu, Chenhui, et al.
Published: (2026)
Personalized Reinforcement Learning with a Budget of Policies
by: Ivanov, Dmitry, et al.
Published: (2024)
by: Ivanov, Dmitry, et al.
Published: (2024)
General Preference Reinforcement Learning
by: Umer, Muhammad, et al.
Published: (2026)
by: Umer, Muhammad, et al.
Published: (2026)
NOS-Gate: Queue-Aware Streaming IDS for Consumer Gateways under Timing-Controlled Evasion
by: Bilal, Muhammad, et al.
Published: (2026)
by: Bilal, Muhammad, et al.
Published: (2026)
Learning Markov State Abstractions for Deep Reinforcement Learning
by: Allen, Cameron, et al.
Published: (2021)
by: Allen, Cameron, et al.
Published: (2021)
Learning-Driven Exploration for Reinforcement Learning
by: Usama, Muhammad, et al.
Published: (2019)
by: Usama, Muhammad, et al.
Published: (2019)
Unveiling the Role of Expert Guidance: A Comparative Analysis of User-centered Imitation Learning and Traditional Reinforcement Learning
by: Gomaa, Amr, et al.
Published: (2024)
by: Gomaa, Amr, et al.
Published: (2024)
Offline Imitation Learning with Model-based Reverse Augmentation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Meta-Reinforcement Learning for Fast and Data-Efficient Spectrum Allocation in Dynamic Wireless Networks
by: Giwa, Oluwaseyi, et al.
Published: (2025)
by: Giwa, Oluwaseyi, et al.
Published: (2025)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2025)
by: Tiwari, Saket, et al.
Published: (2025)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
by: Corrado, Nicholas E., et al.
Published: (2023)
by: Corrado, Nicholas E., et al.
Published: (2023)
Toward Adaptive Reasoning in Large Language Models with Thought Rollback
by: Chen, Sijia, et al.
Published: (2024)
by: Chen, Sijia, et al.
Published: (2024)
Augmenting Offline Reinforcement Learning with State-only Interactions
by: Li, Shangzhe, et al.
Published: (2024)
by: Li, Shangzhe, et al.
Published: (2024)
Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning
by: Trumpp, Raphael, et al.
Published: (2026)
by: Trumpp, Raphael, et al.
Published: (2026)
Federated Hierarchical Reinforcement Learning for Adaptive Traffic Signal Control
by: Fu, Yongjie, et al.
Published: (2025)
by: Fu, Yongjie, et al.
Published: (2025)
ADLight: A Universal Approach of Traffic Signal Control with Augmented Data Using Reinforcement Learning
by: Wang, Maonan, et al.
Published: (2022)
by: Wang, Maonan, et al.
Published: (2022)
Time Reversal Symmetry for Efficient Robotic Manipulations in Deep Reinforcement Learning
by: Jiang, Yunpeng, et al.
Published: (2025)
by: Jiang, Yunpeng, et al.
Published: (2025)
DARLING: Detection Augmented Reinforcement Learning with Non-Stationary Guarantees
by: Gerogiannis, Argyrios, et al.
Published: (2026)
by: Gerogiannis, Argyrios, et al.
Published: (2026)
CARE: Decoding Time Safety Alignment via Rollback and Introspection Intervention
by: Hu, Xiaomeng, et al.
Published: (2025)
by: Hu, Xiaomeng, et al.
Published: (2025)
Heterogeneous Knowledge for Augmented Modular Reinforcement Learning
by: Wolf, Lorenz, et al.
Published: (2023)
by: Wolf, Lorenz, et al.
Published: (2023)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)
by: Sharma, Abhishek, et al.
Published: (2023)
Reverse Forward Curriculum Learning for Extreme Sample and Demonstration Efficiency in Reinforcement Learning
by: Tao, Stone, et al.
Published: (2024)
by: Tao, Stone, et al.
Published: (2024)
Wireless Channel Aware Data Augmentation Methods for Deep Learning-Based Indoor Localization
by: Serbetci, Omer Gokalp, et al.
Published: (2024)
by: Serbetci, Omer Gokalp, et al.
Published: (2024)
Trajectory-Level Data Augmentation for Offline Reinforcement Learning
by: Schmähling, Tobias, et al.
Published: (2026)
by: Schmähling, Tobias, et al.
Published: (2026)
Neural CDEs as Correctors for Learned Time Series Models
by: Shahid, Muhammad Bilal, et al.
Published: (2025)
by: Shahid, Muhammad Bilal, et al.
Published: (2025)
Quantum-Efficient Reinforcement Learning Solutions for Last-Mile On-Demand Delivery
by: Moosavi, Farzan, et al.
Published: (2025)
by: Moosavi, Farzan, et al.
Published: (2025)
Evaluating the Robustness of Reinforcement Learning based Adaptive Traffic Signal Control
by: Kwesiga, Dickens, et al.
Published: (2026)
by: Kwesiga, Dickens, et al.
Published: (2026)
Cloning Ideology and Style using Deep Learning
by: Beg, Omer, et al.
Published: (2022)
by: Beg, Omer, et al.
Published: (2022)
Bi-GRU Based Deception Detection using EEG Signals
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
Stable Hadamard Memory: Revitalizing Memory-Augmented Agents for Reinforcement Learning
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Diffusion-based Episodes Augmentation for Offline Multi-Agent Reinforcement Learning
by: Oh, Jihwan, et al.
Published: (2024)
by: Oh, Jihwan, et al.
Published: (2024)
Spectral Embedding via Chebyshev Bases for Robust DeepONet Approximation
by: Abid, Muhammad, et al.
Published: (2025)
by: Abid, Muhammad, et al.
Published: (2025)
Heuristic Transformer: Belief Augmented In-Context Reinforcement Learning
by: Dippel, Oliver, et al.
Published: (2025)
by: Dippel, Oliver, et al.
Published: (2025)
Distributional Reinforcement Learning with Information Bottleneck for Uncertainty-Aware DRAM Equalization
by: Usama, Muhammad, et al.
Published: (2026)
by: Usama, Muhammad, et al.
Published: (2026)
Similar Items
-
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025) -
Quantum-Inspired Reinforcement Learning for Secure and Sustainable AIoT-Driven Supply Chain Systems
by: Dastagir, Muhammad Bilal Akram, et al.
Published: (2026) -
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025) -
Diagnostics of cognitive failures in multi-agent expert systems using dynamic evaluation protocols and subsequent mutation of the processing context
by: Sorstkins, Andrejs, et al.
Published: (2025) -
Rollback-Free Stable Brick Structures Generation
by: Xu, Chenhui, et al.
Published: (2026)