Gespeichert in:
| Hauptverfasser: | Umili, Elena, Argenziano, Francesco, Capobianco, Roberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2408.08677 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepDFA: Injecting Temporal Logic in Deep Learning for Sequential Subsymbolic Applications
von: Umili, Elena, et al.
Veröffentlicht: (2026)
von: Umili, Elena, et al.
Veröffentlicht: (2026)
DeepDFA: Automata Learning through Neural Probabilistic Relaxations
von: Umili, Elena, et al.
Veröffentlicht: (2024)
von: Umili, Elena, et al.
Veröffentlicht: (2024)
Fully Learnable Neural Reward Machines
von: Dewidar, Hazem, et al.
Veröffentlicht: (2025)
von: Dewidar, Hazem, et al.
Veröffentlicht: (2025)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
von: Pannacci, Matteo, et al.
Veröffentlicht: (2026)
von: Pannacci, Matteo, et al.
Veröffentlicht: (2026)
Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
Numeric Reward Machines
von: Levina, Kristina, et al.
Veröffentlicht: (2024)
von: Levina, Kristina, et al.
Veröffentlicht: (2024)
Don't Push the Button! Exploring Data Leakage Risks in Machine Learning and Transfer Learning
von: Apicella, Andrea, et al.
Veröffentlicht: (2024)
von: Apicella, Andrea, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Symbolic Reward Machines
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Stochastic Reward Machines
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
Learning Robust Reward Machines from Noisy Labels
von: Parac, Roko, et al.
Veröffentlicht: (2024)
von: Parac, Roko, et al.
Veröffentlicht: (2024)
Provably Efficient Exploration in Reward Machines with Low Regret
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
Maximally Permissive Reward Machines
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
von: Levina, Kristina, et al.
Veröffentlicht: (2026)
von: Levina, Kristina, et al.
Veröffentlicht: (2026)
Pushdown Reward Machines for Reinforcement Learning
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
Mechanistic Neural Networks for Scientific Machine Learning
von: Pervez, Adeel, et al.
Veröffentlicht: (2024)
von: Pervez, Adeel, et al.
Veröffentlicht: (2024)
Expressive Temporal Specifications for Reward Monitoring
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
von: Azran, Guy, et al.
Veröffentlicht: (2023)
von: Azran, Guy, et al.
Veröffentlicht: (2023)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
Few-shot Steerable Alignment: Adapting Rewards and LLM Policies with Neural Processes
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2024)
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2024)
Defining and Monitoring Complex Robot Activities via LLMs and Symbolic Reasoning
von: Argenziano, Francesco, et al.
Veröffentlicht: (2025)
von: Argenziano, Francesco, et al.
Veröffentlicht: (2025)
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
von: Maura-Rivero, Roberto-Rafael, et al.
Veröffentlicht: (2025)
von: Maura-Rivero, Roberto-Rafael, et al.
Veröffentlicht: (2025)
Inferring Reward Machines and Transition Machines from Partially Observable Markov Decision Processes
von: Wu, Yuly, et al.
Veröffentlicht: (2025)
von: Wu, Yuly, et al.
Veröffentlicht: (2025)
Attention-Based Reward Shaping for Sparse and Delayed Rewards
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
Reward Hacking Mitigation using Verifiable Composite Rewards
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
Repairing Reward Functions with Feedback to Mitigate Reward Hacking
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
Intrinsic Reward Policy Optimization for Sparse-Reward Environments
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
Multi-Agent Reinforcement Learning with a Hierarchy of Reward Machines
von: Zheng, Xuejing, et al.
Veröffentlicht: (2024)
von: Zheng, Xuejing, et al.
Veröffentlicht: (2024)
Neuro-Symbolic Predictive Process Monitoring
von: Mezini, Axel, et al.
Veröffentlicht: (2025)
von: Mezini, Axel, et al.
Veröffentlicht: (2025)
Reward Centering
von: Naik, Abhishek, et al.
Veröffentlicht: (2024)
von: Naik, Abhishek, et al.
Veröffentlicht: (2024)
Recursive Inference Machines for Neural Reasoning
von: Komisarczyk, Mieszko, et al.
Veröffentlicht: (2026)
von: Komisarczyk, Mieszko, et al.
Veröffentlicht: (2026)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
SemiReward: A General Reward Model for Semi-supervised Learning
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
von: Chen, Yang, et al.
Veröffentlicht: (2025)
von: Chen, Yang, et al.
Veröffentlicht: (2025)
Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
von: Wang, Chaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Chaoqi, et al.
Veröffentlicht: (2025)
Tiered Reward: Designing Rewards for Specification and Fast Learning of Desired Behavior
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
Rethinking Rubric Generation for Improving LLM Judge and Reward Modeling for Open-ended Tasks
von: Shen, William F., et al.
Veröffentlicht: (2026)
von: Shen, William F., et al.
Veröffentlicht: (2026)
Neural Network Conversion of Machine Learning Pipelines
von: Sung, Man-Ling, et al.
Veröffentlicht: (2026)
von: Sung, Man-Ling, et al.
Veröffentlicht: (2026)
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DeepDFA: Injecting Temporal Logic in Deep Learning for Sequential Subsymbolic Applications
von: Umili, Elena, et al.
Veröffentlicht: (2026) -
DeepDFA: Automata Learning through Neural Probabilistic Relaxations
von: Umili, Elena, et al.
Veröffentlicht: (2024) -
Fully Learnable Neural Reward Machines
von: Dewidar, Hazem, et al.
Veröffentlicht: (2025) -
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
von: Pannacci, Matteo, et al.
Veröffentlicht: (2026) -
Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)