Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Volovikova, Zoya, Sorokin, Nikita, Lukashevskiy, Dmitriy, Panov, Aleksandr, Skrynnik, Alexey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Instruction Following with Goal-Conditioned Reinforcement Learning in Virtual Environments
by: Volovikova, Zoya, et al.
Published: (2024)
by: Volovikova, Zoya, et al.
Published: (2024)
CrafText Benchmark: Advancing Instruction Following in Complex Multimodal Open-Ended World
by: Volovikova, Zoya, et al.
Published: (2025)
by: Volovikova, Zoya, et al.
Published: (2025)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
by: Ivanova, Anastasiia, et al.
Published: (2025)
by: Ivanova, Anastasiia, et al.
Published: (2025)
CAMAR: Continuous Actions Multi-Agent Routing
by: Pshenitsyn, Artem, et al.
Published: (2025)
by: Pshenitsyn, Artem, et al.
Published: (2025)
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
by: Andreychuk, Anton, et al.
Published: (2024)
by: Andreychuk, Anton, et al.
Published: (2024)
Pragmatic Instruction Following and Goal Assistance via Cooperative Language-Guided Inverse Planning
by: Zhi-Xuan, Tan, et al.
Published: (2024)
by: Zhi-Xuan, Tan, et al.
Published: (2024)
Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following
by: Ren, Qingyu, et al.
Published: (2025)
by: Ren, Qingyu, et al.
Published: (2025)
Revisiting Tree Search for LLMs: Gumbel and Sequential Halving for Budget-Scalable Reasoning
by: Ugadiarov, Leonid, et al.
Published: (2026)
by: Ugadiarov, Leonid, et al.
Published: (2026)
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
by: Andreychuk, Anton, et al.
Published: (2025)
by: Andreychuk, Anton, et al.
Published: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
by: Onishchenko, Anatoly O., et al.
Published: (2025)
by: Onishchenko, Anatoly O., et al.
Published: (2025)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
VerIF: Verification Engineering for Reinforcement Learning in Instruction Following
by: Peng, Hao, et al.
Published: (2025)
by: Peng, Hao, et al.
Published: (2025)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
by: Nesterova, Maria, et al.
Published: (2026)
by: Nesterova, Maria, et al.
Published: (2026)
SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning
by: Zhao, Chenyang, et al.
Published: (2024)
by: Zhao, Chenyang, et al.
Published: (2024)
Compositional Automata Embeddings for Goal-Conditioned Reinforcement Learning
by: Yalcinkaya, Beyazit, et al.
Published: (2024)
by: Yalcinkaya, Beyazit, et al.
Published: (2024)
IDAT: A Multi-Modal Dataset and Toolkit for Building and Evaluating Interactive Task-Solving Agents
by: Mohanty, Shrestha, et al.
Published: (2024)
by: Mohanty, Shrestha, et al.
Published: (2024)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
by: Grigorev, Danil S., et al.
Published: (2025)
by: Grigorev, Danil S., et al.
Published: (2025)
Reliable Extraction of Clinical Follow-Up Instructions: A Hybrid Neural-Symbolic Pipeline
by: Laufer, Michal, et al.
Published: (2026)
by: Laufer, Michal, et al.
Published: (2026)
Planning without Search: Refining Frontier LLMs with Offline Goal-Conditioned RL
by: Hong, Joey, et al.
Published: (2025)
by: Hong, Joey, et al.
Published: (2025)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
by: Shchendrigin, Oleg, et al.
Published: (2026)
by: Shchendrigin, Oleg, et al.
Published: (2026)
Self-Review Framework for Enhancing Instruction Following Capability of LLM
by: Park, Sihyun
Published: (2025)
by: Park, Sihyun
Published: (2025)
HREF: Human Response-Guided Evaluation of Instruction Following in Language Models
by: Lyu, Xinxi, et al.
Published: (2024)
by: Lyu, Xinxi, et al.
Published: (2024)
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
by: Fei, Zhaoye, et al.
Published: (2025)
by: Fei, Zhaoye, et al.
Published: (2025)
Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction
by: Ding, Zepeng, et al.
Published: (2024)
by: Ding, Zepeng, et al.
Published: (2024)
BAP v2: An Enhanced Task Framework for Instruction Following in Minecraft Dialogues
by: Jayannavar, Prashant, et al.
Published: (2025)
by: Jayannavar, Prashant, et al.
Published: (2025)
Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction Following
by: Zhang, Kongcheng, et al.
Published: (2025)
by: Zhang, Kongcheng, et al.
Published: (2025)
Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models
by: Rubashevskii, Aleksandr, et al.
Published: (2026)
by: Rubashevskii, Aleksandr, et al.
Published: (2026)
Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task
by: Molinaro, Gaia, et al.
Published: (2026)
by: Molinaro, Gaia, et al.
Published: (2026)
Knowledge-Guided Manipulation Using Multi-Task Reinforcement Learning
by: Narendra, Aditya, et al.
Published: (2026)
by: Narendra, Aditya, et al.
Published: (2026)
Rubrics to Tokens: Bridging Response-level Rubrics and Token-level Rewards in Instruction Following Tasks
by: Xu, Tianze, et al.
Published: (2026)
by: Xu, Tianze, et al.
Published: (2026)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
by: Qu, Kaixian, et al.
Published: (2024)
by: Qu, Kaixian, et al.
Published: (2024)
The Instruction Gap: LLMs get lost in Following Instruction
by: Tripathi, Vishesh, et al.
Published: (2025)
by: Tripathi, Vishesh, et al.
Published: (2025)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
by: Skrynnik, Alexey, et al.
Published: (2024)
by: Skrynnik, Alexey, et al.
Published: (2024)
Personalized Learning Path Planning with Goal-Driven Learner State Modeling
by: Lim, Joy Jia Yin, et al.
Published: (2025)
by: Lim, Joy Jia Yin, et al.
Published: (2025)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
by: Huang, Hui, et al.
Published: (2025)
by: Huang, Hui, et al.
Published: (2025)
Learning to Better Search with Language Models via Guided Reinforced Self-Training
by: Moon, Seungyong, et al.
Published: (2024)
by: Moon, Seungyong, et al.
Published: (2024)
MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following
by: Lou, Renze, et al.
Published: (2023)
by: Lou, Renze, et al.
Published: (2023)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
Training with Pseudo-Code for Instruction Following
by: Kumar, Prince, et al.
Published: (2025)
by: Kumar, Prince, et al.
Published: (2025)
WildIFEval: Instruction Following in the Wild
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Similar Items
-
Instruction Following with Goal-Conditioned Reinforcement Learning in Virtual Environments
by: Volovikova, Zoya, et al.
Published: (2024) -
CrafText Benchmark: Advancing Instruction Following in Complex Multimodal Open-Ended World
by: Volovikova, Zoya, et al.
Published: (2025) -
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
by: Ivanova, Anastasiia, et al.
Published: (2025) -
CAMAR: Continuous Actions Multi-Agent Routing
by: Pshenitsyn, Artem, et al.
Published: (2025) -
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
by: Andreychuk, Anton, et al.
Published: (2024)