Zero-Shot Instruction Following in RL via Structured LTL Representations
Fuente:
arXiv
Guardado en:
| Autores principales: | Giuri, Mattia, Jackermeier, Mathias, Abate, Alessandro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zero-Shot Instruction Following in RL via Structured LTL Representations
por: Jackermeier, Mathias, et al.
Publicado: (2026)
por: Jackermeier, Mathias, et al.
Publicado: (2026)
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
por: Jackermeier, Mathias, et al.
Publicado: (2024)
por: Jackermeier, Mathias, et al.
Publicado: (2024)
PlatoLTL: Learning to Generalize Across Symbols in LTL Instructions for Multi-Task RL
por: Cloete, Jacques, et al.
Publicado: (2026)
por: Cloete, Jacques, et al.
Publicado: (2026)
Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions
por: Abate, Alessandro, et al.
Publicado: (2026)
por: Abate, Alessandro, et al.
Publicado: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
por: Schnitzer, Yannik, et al.
Publicado: (2026)
por: Schnitzer, Yannik, et al.
Publicado: (2026)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
por: Pannacci, Matteo, et al.
Publicado: (2026)
por: Pannacci, Matteo, et al.
Publicado: (2026)
Dual Box Embeddings for the Description Logic EL++
por: Jackermeier, Mathias, et al.
Publicado: (2023)
por: Jackermeier, Mathias, et al.
Publicado: (2023)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
Partial Identifiability and Misspecification in Inverse Reinforcement Learning
por: Skalse, Joar, et al.
Publicado: (2024)
por: Skalse, Joar, et al.
Publicado: (2024)
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks
por: Skalse, Joar, et al.
Publicado: (2024)
por: Skalse, Joar, et al.
Publicado: (2024)
Partial Identifiability in Inverse Reinforcement Learning For Agents With Non-Exponential Discounting
por: Skalse, Joar, et al.
Publicado: (2024)
por: Skalse, Joar, et al.
Publicado: (2024)
Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following
por: Yang, Brian, et al.
Publicado: (2024)
por: Yang, Brian, et al.
Publicado: (2024)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL
por: Kim, Sung-Hyun, et al.
Publicado: (2025)
por: Kim, Sung-Hyun, et al.
Publicado: (2025)
Zero-Shot Robustification of Zero-Shot Models
por: Adila, Dyah, et al.
Publicado: (2023)
por: Adila, Dyah, et al.
Publicado: (2023)
Regret-Free Reinforcement Learning for LTL Specifications
por: Majumdar, Rupak, et al.
Publicado: (2024)
por: Majumdar, Rupak, et al.
Publicado: (2024)
Efficient Solution and Learning of Robust Factored MDPs
por: Schnitzer, Yannik, et al.
Publicado: (2025)
por: Schnitzer, Yannik, et al.
Publicado: (2025)
Neural Proofs for Sound Verification and Control of Complex Systems
por: Abate, Alessandro
Publicado: (2025)
por: Abate, Alessandro
Publicado: (2025)
Zero-Shot Reinforcement Learning via Function Encoders
por: Ingebrand, Tyler, et al.
Publicado: (2024)
por: Ingebrand, Tyler, et al.
Publicado: (2024)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
por: Melo, Luckeciano C., et al.
Publicado: (2025)
por: Melo, Luckeciano C., et al.
Publicado: (2025)
Temporal-Difference Variational Continual Learning
por: Melo, Luckeciano C., et al.
Publicado: (2024)
por: Melo, Luckeciano C., et al.
Publicado: (2024)
Reinforcement Learning with LTL and $ω$-Regular Objectives via Optimality-Preserving Translation to Average Rewards
por: Le, Xuan-Bach, et al.
Publicado: (2024)
por: Le, Xuan-Bach, et al.
Publicado: (2024)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
por: Kang, Minjae, et al.
Publicado: (2026)
por: Kang, Minjae, et al.
Publicado: (2026)
On Zero-Shot Reinforcement Learning
por: Jeen, Scott
Publicado: (2025)
por: Jeen, Scott
Publicado: (2025)
ZeroShotOpt: Towards Zero-Shot Pretrained Models for Efficient Black-Box Optimization
por: Meindl, Jamison, et al.
Publicado: (2025)
por: Meindl, Jamison, et al.
Publicado: (2025)
Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
por: Zeng, Yirong, et al.
Publicado: (2025)
por: Zeng, Yirong, et al.
Publicado: (2025)
Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols
por: Griffin, Charlie, et al.
Publicado: (2024)
por: Griffin, Charlie, et al.
Publicado: (2024)
Zero-Shot Quantization via Weight-Space Arithmetic
por: Solombrino, Daniele, et al.
Publicado: (2026)
por: Solombrino, Daniele, et al.
Publicado: (2026)
Zero-Shot Cyclic Peptide Design via Composable Geometric Constraints
por: Jiang, Dapeng, et al.
Publicado: (2025)
por: Jiang, Dapeng, et al.
Publicado: (2025)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
por: Frans, Kevin, et al.
Publicado: (2024)
por: Frans, Kevin, et al.
Publicado: (2024)
Bus-Conditioned Zero-Shot Trajectory Generation via Task Arithmetic
por: Liu, Shuai, et al.
Publicado: (2026)
por: Liu, Shuai, et al.
Publicado: (2026)
Improving Instruction-Following in Language Models through Activation Steering
por: Stolfo, Alessandro, et al.
Publicado: (2024)
por: Stolfo, Alessandro, et al.
Publicado: (2024)
TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations
por: Bae, Junik, et al.
Publicado: (2024)
por: Bae, Junik, et al.
Publicado: (2024)
GLIDE-RL: Grounded Language Instruction through DEmonstration in RL
por: Kharyal, Chaitanya, et al.
Publicado: (2024)
por: Kharyal, Chaitanya, et al.
Publicado: (2024)
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
por: Höpner, Niklas, et al.
Publicado: (2025)
por: Höpner, Niklas, et al.
Publicado: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
por: He, Bingxiang, et al.
Publicado: (2024)
por: He, Bingxiang, et al.
Publicado: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
por: Zhang, Jesse, et al.
Publicado: (2025)
por: Zhang, Jesse, et al.
Publicado: (2025)
Zero-Shot Off-Policy Learning
por: Asadulaev, Arip, et al.
Publicado: (2026)
por: Asadulaev, Arip, et al.
Publicado: (2026)
Certifiably Robust Policies for Uncertain Parametric Environments
por: Schnitzer, Yannik, et al.
Publicado: (2024)
por: Schnitzer, Yannik, et al.
Publicado: (2024)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
por: Batra, Sumeet, et al.
Publicado: (2024)
por: Batra, Sumeet, et al.
Publicado: (2024)
Ejemplares similares
-
Zero-Shot Instruction Following in RL via Structured LTL Representations
por: Jackermeier, Mathias, et al.
Publicado: (2026) -
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
por: Jackermeier, Mathias, et al.
Publicado: (2024) -
PlatoLTL: Learning to Generalize Across Symbols in LTL Instructions for Multi-Task RL
por: Cloete, Jacques, et al.
Publicado: (2026) -
Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions
por: Abate, Alessandro, et al.
Publicado: (2026) -
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
por: Schnitzer, Yannik, et al.
Publicado: (2026)