Safe Continual Reinforcement Learning Methods for Nonstationary Environments. Towards a Survey of the State of the Art
Fuente:
arXiv
Salvato in:
| Autore principale: | Tomashevskiy, Timofey |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Adaptive Latent-Space Constraints in Personalized Federated Learning
di: Ayromlou, Sana, et al.
Pubblicazione: (2025)
di: Ayromlou, Sana, et al.
Pubblicazione: (2025)
GLL: A Differentiable Graph Learning Layer for Neural Networks
di: Brown, Jason, et al.
Pubblicazione: (2024)
di: Brown, Jason, et al.
Pubblicazione: (2024)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
di: Shafieinejad, Masoumeh, et al.
Pubblicazione: (2026)
di: Shafieinejad, Masoumeh, et al.
Pubblicazione: (2026)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
Matryoshka Policy Gradient for Entropy-Regularized RL: Convergence and Global Optimality
di: Ged, François, et al.
Pubblicazione: (2023)
di: Ged, François, et al.
Pubblicazione: (2023)
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
di: Tomashevskiy, Timofey
Pubblicazione: (2026)
di: Tomashevskiy, Timofey
Pubblicazione: (2026)
CellARC: Measuring Intelligence with Cellular Automata
di: Lžičař, Miroslav
Pubblicazione: (2025)
di: Lžičař, Miroslav
Pubblicazione: (2025)
Recent Advances in Data-Driven Business Process Management
di: Ackermann, Lars, et al.
Pubblicazione: (2024)
di: Ackermann, Lars, et al.
Pubblicazione: (2024)
The Normalized Difference Layer: A Differentiable Spectral Index Formulation for Deep Learning
di: Lotfi, Ali, et al.
Pubblicazione: (2026)
di: Lotfi, Ali, et al.
Pubblicazione: (2026)
Convolutional Neural Networks Can (Meta-)Learn the Same-Different Relation
di: Gupta, Max, et al.
Pubblicazione: (2025)
di: Gupta, Max, et al.
Pubblicazione: (2025)
Attention Please: What Transformer Models Really Learn for Process Prediction
di: Käppel, Martin, et al.
Pubblicazione: (2024)
di: Käppel, Martin, et al.
Pubblicazione: (2024)
Thinking Machines: Mathematical Reasoning in the Age of LLMs
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
JAM: Controllable and Responsible Text Generation via Causal Reasoning and Latent Vector Manipulation
di: Huang, Yingbing, et al.
Pubblicazione: (2025)
di: Huang, Yingbing, et al.
Pubblicazione: (2025)
Graph Transformers: A Survey
di: Shehzad, Ahsan, et al.
Pubblicazione: (2024)
di: Shehzad, Ahsan, et al.
Pubblicazione: (2024)
Position: Mechanistic Interpretability Must Disclose Identification Assumptions for Causal Claims
di: Lin, Zezheng, et al.
Pubblicazione: (2026)
di: Lin, Zezheng, et al.
Pubblicazione: (2026)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
di: Platzer, André
Pubblicazione: (2024)
di: Platzer, André
Pubblicazione: (2024)
The Matching Principle: A Geometric Theory of Loss Functions for Nuisance-Robust Representation Learning
di: Rajput, Vishal
Pubblicazione: (2026)
di: Rajput, Vishal
Pubblicazione: (2026)
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
di: Jing, Yuheng, et al.
Pubblicazione: (2026)
di: Jing, Yuheng, et al.
Pubblicazione: (2026)
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
di: Klačan, Ján, et al.
Pubblicazione: (2026)
di: Klačan, Ján, et al.
Pubblicazione: (2026)
Task-Aligned Self-Supervised Learning for Medical Image Analysis: A Systematic Review and Practical Design Guidelines
di: Wimalasiri, Chathura
Pubblicazione: (2026)
di: Wimalasiri, Chathura
Pubblicazione: (2026)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
di: Cao, Deyu, et al.
Pubblicazione: (2025)
di: Cao, Deyu, et al.
Pubblicazione: (2025)
Emergence of Goal-Directed Behaviors via Active Inference with Self-Prior
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
Complex Facial Expression Recognition Using Deep Knowledge Distillation of Basic Features
di: Maiden, Angus, et al.
Pubblicazione: (2023)
di: Maiden, Angus, et al.
Pubblicazione: (2023)
Neural Encoding for Image Recall: Human-Like Memory
di: Foussereau, Virgile, et al.
Pubblicazione: (2024)
di: Foussereau, Virgile, et al.
Pubblicazione: (2024)
Machine Learning as Performative Materialist Practice: Thirteen Theses on the Epistemology, Methodology, and Politics of Applied ML
di: De Unánue, Adolfo, et al.
Pubblicazione: (2026)
di: De Unánue, Adolfo, et al.
Pubblicazione: (2026)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
di: Atamuradov, Sanjar
Pubblicazione: (2025)
di: Atamuradov, Sanjar
Pubblicazione: (2025)
Threat Detection in Social Media Networks Using Machine Learning Based Network Analysis
di: Agrawal, Aditi Sanjay
Pubblicazione: (2026)
di: Agrawal, Aditi Sanjay
Pubblicazione: (2026)
Pattern Recognition Tasks with Personalized Federated Learning
di: Rahman, Md. Arifur, et al.
Pubblicazione: (2026)
di: Rahman, Md. Arifur, et al.
Pubblicazione: (2026)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
di: Bian, Yijun, et al.
Pubblicazione: (2024)
di: Bian, Yijun, et al.
Pubblicazione: (2024)
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
di: Ho, Siu Hang, et al.
Pubblicazione: (2025)
di: Ho, Siu Hang, et al.
Pubblicazione: (2025)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
di: Wang, Zhen, et al.
Pubblicazione: (2025)
di: Wang, Zhen, et al.
Pubblicazione: (2025)
CaLoRAify: Calorie Estimation with Visual-Text Pairing and LoRA-Driven Visual Language Models
di: Yao, Dongyu, et al.
Pubblicazione: (2024)
di: Yao, Dongyu, et al.
Pubblicazione: (2024)
Attention-based graph neural networks: a survey
di: Sun, Chengcheng, et al.
Pubblicazione: (2026)
di: Sun, Chengcheng, et al.
Pubblicazione: (2026)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
di: Lomasto, Luigi, et al.
Pubblicazione: (2026)
di: Lomasto, Luigi, et al.
Pubblicazione: (2026)
Hazard-Responsive Digital Twin for Climate-Driven Urban Resilience and Equity
di: Shen, Zhenglai, et al.
Pubblicazione: (2025)
di: Shen, Zhenglai, et al.
Pubblicazione: (2025)
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
di: Brännvall, Rickard, et al.
Pubblicazione: (2023)
di: Brännvall, Rickard, et al.
Pubblicazione: (2023)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
di: Zhou, Xueyang, et al.
Pubblicazione: (2025)
di: Zhou, Xueyang, et al.
Pubblicazione: (2025)
NeuBTF: Neural fields for BTF encoding and transfer
di: Rodriguez-Pardo, Carlos, et al.
Pubblicazione: (2023)
di: Rodriguez-Pardo, Carlos, et al.
Pubblicazione: (2023)
Single-image Reflectance and Transmittance Estimation from Any Flatbed Scanner
di: Rodriguez-Pardo, Carlos, et al.
Pubblicazione: (2025)
di: Rodriguez-Pardo, Carlos, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
di: Guo, Dongxin, et al.
Pubblicazione: (2026) -
Adaptive Latent-Space Constraints in Personalized Federated Learning
di: Ayromlou, Sana, et al.
Pubblicazione: (2025) -
GLL: A Differentiable Graph Learning Layer for Neural Networks
di: Brown, Jason, et al.
Pubblicazione: (2024) -
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
di: Shafieinejad, Masoumeh, et al.
Pubblicazione: (2026) -
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)