Zero-Shot Instruction Following in RL via Structured LTL Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jackermeier, Mathias, Giuri, Mattia, Cloete, Jacques, Abate, Alessandro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Instruction Following in RL via Structured LTL Representations
von: Giuri, Mattia, et al.
Veröffentlicht: (2025)
von: Giuri, Mattia, et al.
Veröffentlicht: (2025)
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2024)
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2024)
PlatoLTL: Learning to Generalize Across Symbols in LTL Instructions for Multi-Task RL
von: Cloete, Jacques, et al.
Veröffentlicht: (2026)
von: Cloete, Jacques, et al.
Veröffentlicht: (2026)
Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions
von: Abate, Alessandro, et al.
Veröffentlicht: (2026)
von: Abate, Alessandro, et al.
Veröffentlicht: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
von: Pannacci, Matteo, et al.
Veröffentlicht: (2026)
von: Pannacci, Matteo, et al.
Veröffentlicht: (2026)
SPoRt -- Safe Policy Ratio: Certified Training and Deployment of Task Policies in Model-Free RL
von: Cloete, Jacques, et al.
Veröffentlicht: (2025)
von: Cloete, Jacques, et al.
Veröffentlicht: (2025)
Dual Box Embeddings for the Description Logic EL++
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2023)
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2023)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
Partial Identifiability and Misspecification in Inverse Reinforcement Learning
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
Partial Identifiability in Inverse Reinforcement Learning For Agents With Non-Exponential Discounting
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following
von: Yang, Brian, et al.
Veröffentlicht: (2024)
von: Yang, Brian, et al.
Veröffentlicht: (2024)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
von: Nazir, Mohammad Saif, et al.
Veröffentlicht: (2025)
von: Nazir, Mohammad Saif, et al.
Veröffentlicht: (2025)
Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL
von: Kim, Sung-Hyun, et al.
Veröffentlicht: (2025)
von: Kim, Sung-Hyun, et al.
Veröffentlicht: (2025)
Zero-Shot Robustification of Zero-Shot Models
von: Adila, Dyah, et al.
Veröffentlicht: (2023)
von: Adila, Dyah, et al.
Veröffentlicht: (2023)
Regret-Free Reinforcement Learning for LTL Specifications
von: Majumdar, Rupak, et al.
Veröffentlicht: (2024)
von: Majumdar, Rupak, et al.
Veröffentlicht: (2024)
Efficient Solution and Learning of Robust Factored MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
Neural Proofs for Sound Verification and Control of Complex Systems
von: Abate, Alessandro
Veröffentlicht: (2025)
von: Abate, Alessandro
Veröffentlicht: (2025)
Zero-Shot Reinforcement Learning via Function Encoders
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
Temporal-Difference Variational Continual Learning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Reinforcement Learning with LTL and $ω$-Regular Objectives via Optimality-Preserving Translation to Average Rewards
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
On Zero-Shot Reinforcement Learning
von: Jeen, Scott
Veröffentlicht: (2025)
von: Jeen, Scott
Veröffentlicht: (2025)
ZeroShotOpt: Towards Zero-Shot Pretrained Models for Efficient Black-Box Optimization
von: Meindl, Jamison, et al.
Veröffentlicht: (2025)
von: Meindl, Jamison, et al.
Veröffentlicht: (2025)
Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols
von: Griffin, Charlie, et al.
Veröffentlicht: (2024)
von: Griffin, Charlie, et al.
Veröffentlicht: (2024)
Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
von: Zeng, Yirong, et al.
Veröffentlicht: (2025)
von: Zeng, Yirong, et al.
Veröffentlicht: (2025)
Zero-Shot Quantization via Weight-Space Arithmetic
von: Solombrino, Daniele, et al.
Veröffentlicht: (2026)
von: Solombrino, Daniele, et al.
Veröffentlicht: (2026)
Bus-Conditioned Zero-Shot Trajectory Generation via Task Arithmetic
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
Improving Instruction-Following in Language Models through Activation Steering
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
von: Frans, Kevin, et al.
Veröffentlicht: (2024)
von: Frans, Kevin, et al.
Veröffentlicht: (2024)
Zero-Shot Cyclic Peptide Design via Composable Geometric Constraints
von: Jiang, Dapeng, et al.
Veröffentlicht: (2025)
von: Jiang, Dapeng, et al.
Veröffentlicht: (2025)
TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations
von: Bae, Junik, et al.
Veröffentlicht: (2024)
von: Bae, Junik, et al.
Veröffentlicht: (2024)
GLIDE-RL: Grounded Language Instruction through DEmonstration in RL
von: Kharyal, Chaitanya, et al.
Veröffentlicht: (2024)
von: Kharyal, Chaitanya, et al.
Veröffentlicht: (2024)
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
von: Höpner, Niklas, et al.
Veröffentlicht: (2025)
von: Höpner, Niklas, et al.
Veröffentlicht: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
Zero-Shot Off-Policy Learning
von: Asadulaev, Arip, et al.
Veröffentlicht: (2026)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2026)
Certifiably Robust Policies for Uncertain Parametric Environments
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2024)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Zero-Shot Instruction Following in RL via Structured LTL Representations
von: Giuri, Mattia, et al.
Veröffentlicht: (2025) -
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2024) -
PlatoLTL: Learning to Generalize Across Symbols in LTL Instructions for Multi-Task RL
von: Cloete, Jacques, et al.
Veröffentlicht: (2026) -
Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions
von: Abate, Alessandro, et al.
Veröffentlicht: (2026) -
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)