How Far Are LLMs from Symbolic Planners? An NLP-Based Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Armony, Ma'ayan, Meroño-Peñuela, Albert, Canal, Gerard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Schema Generation for Large Knowledge Graphs Using Large Language Models
by: Zhang, Bohui, et al.
Published: (2025)
by: Zhang, Bohui, et al.
Published: (2025)
PathE: Leveraging Entity-Agnostic Paths for Parameter-Efficient Knowledge Graph Embeddings
by: Reklos, Ioannis, et al.
Published: (2025)
by: Reklos, Ioannis, et al.
Published: (2025)
Towards Responsible AI Music: an Investigation of Trustworthy Features for Creative Systems
by: de Berardinis, Jacopo, et al.
Published: (2025)
by: de Berardinis, Jacopo, et al.
Published: (2025)
How Far Can Pretrained LLMs Go in Symbolic Music? Controlled Comparisons of Supervised and Preference-based Adaptation
by: Kumar, Deepak, et al.
Published: (2026)
by: Kumar, Deepak, et al.
Published: (2026)
LLMs for Relational Reasoning: How Far are We?
by: Li, Zhiming, et al.
Published: (2024)
by: Li, Zhiming, et al.
Published: (2024)
Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation
by: Chen, Yongchao, et al.
Published: (2025)
by: Chen, Yongchao, et al.
Published: (2025)
How Far are VLMs from Visual Spatial Intelligence? A Benchmark-Driven Perspective
by: Yu, Songsong, et al.
Published: (2025)
by: Yu, Songsong, et al.
Published: (2025)
How Far are LLMs from Real Search? A Comprehensive Study on Efficiency, Completeness, and Inherent Capabilities
by: Lin, Minhua, et al.
Published: (2025)
by: Lin, Minhua, et al.
Published: (2025)
How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use
by: Lin, Minhua, et al.
Published: (2026)
by: Lin, Minhua, et al.
Published: (2026)
Large Language Models for Predictive Analysis: How Far Are They?
by: Chen, Qin, et al.
Published: (2025)
by: Chen, Qin, et al.
Published: (2025)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
by: Huang, Jen-tse, et al.
Published: (2024)
by: Huang, Jen-tse, et al.
Published: (2024)
How Far Are AI Scientists from Changing the World?
by: Xie, Qiujie, et al.
Published: (2025)
by: Xie, Qiujie, et al.
Published: (2025)
Improving Ontology Requirements Engineering with OntoChat and Participatory Prompting
by: Zhao, Yihang, et al.
Published: (2024)
by: Zhao, Yihang, et al.
Published: (2024)
HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?
by: Yu, Fangchen, et al.
Published: (2025)
by: Yu, Fangchen, et al.
Published: (2025)
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
by: Zhu, Wang, et al.
Published: (2024)
by: Zhu, Wang, et al.
Published: (2024)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
NSP: A Neuro-Symbolic Natural Language Navigational Planner
by: English, William, et al.
Published: (2024)
by: English, William, et al.
Published: (2024)
How Far is Video Generation from World Model: A Physical Law Perspective
by: Kang, Bingyi, et al.
Published: (2024)
by: Kang, Bingyi, et al.
Published: (2024)
How Far Are We from True Unlearnability?
by: Ye, Kai, et al.
Published: (2025)
by: Ye, Kai, et al.
Published: (2025)
Proving Olympiad Inequalities by Synergizing LLMs and Symbolic Reasoning
by: Li, Zenan, et al.
Published: (2025)
by: Li, Zenan, et al.
Published: (2025)
TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?
by: Xiao, Jieting, et al.
Published: (2026)
by: Xiao, Jieting, et al.
Published: (2026)
How Far Are We From AGI: Are LLMs All We Need?
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
Artifact for TOSEM paper: Exploring Development Methods for Reactive Synthesis Specifications
by: Ma'ayan, Dor, et al.
Published: (2025)
by: Ma'ayan, Dor, et al.
Published: (2025)
How Far Are We from Optimal Reasoning Efficiency?
by: Gao, Jiaxuan, et al.
Published: (2025)
by: Gao, Jiaxuan, et al.
Published: (2025)
Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
by: Zheng, Haizhong, et al.
Published: (2025)
by: Zheng, Haizhong, et al.
Published: (2025)
VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?
by: Liu, Junpeng, et al.
Published: (2024)
by: Liu, Junpeng, et al.
Published: (2024)
Planner-R1: Reward Shaping Enables Efficient Agentic RL with Smaller LLMs
by: Zhu, Siyu, et al.
Published: (2025)
by: Zhu, Siyu, et al.
Published: (2025)
The AI Hippocampus: How Far are We From Human Memory?
by: Jia, Zixia, et al.
Published: (2026)
by: Jia, Zixia, et al.
Published: (2026)
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions?
by: Wachowiak, Lennart, et al.
Published: (2024)
by: Wachowiak, Lennart, et al.
Published: (2024)
Deep Learning-Based Identification of Inconsistent Method Names: How Far Are We?
by: Wang, Taiming, et al.
Published: (2025)
by: Wang, Taiming, et al.
Published: (2025)
OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?
by: Li, Yifei, et al.
Published: (2025)
by: Li, Yifei, et al.
Published: (2025)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
by: Srinivasan, Sudarshan, et al.
Published: (2024)
by: Srinivasan, Sudarshan, et al.
Published: (2024)
DSBench: How Far Are Data Science Agents from Becoming Data Science Experts?
by: Jing, Liqiang, et al.
Published: (2024)
by: Jing, Liqiang, et al.
Published: (2024)
Perspectives in Play: A Multi-Perspective Approach for More Inclusive NLP Systems
by: Muscato, Benedetta, et al.
Published: (2025)
by: Muscato, Benedetta, et al.
Published: (2025)
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs?
by: Ziomek, Juliusz, et al.
Published: (2026)
by: Ziomek, Juliusz, et al.
Published: (2026)
Neuro-Symbolic Verification on Instruction Following of LLMs
by: Su, Yiming, et al.
Published: (2026)
by: Su, Yiming, et al.
Published: (2026)
A Rule-Based Behaviour Planner for Autonomous Driving
by: Frederic, Bouchard, et al.
Published: (2024)
by: Frederic, Bouchard, et al.
Published: (2024)
CorrectionPlanner: Self-Correction Planner with Reinforcement Learning in Autonomous Driving
by: Guo, Yihong, et al.
Published: (2026)
by: Guo, Yihong, et al.
Published: (2026)
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025)
by: Gladstone, Clovis, et al.
Published: (2025)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
by: Calderon, Nitay, et al.
Published: (2024)
by: Calderon, Nitay, et al.
Published: (2024)
Similar Items
-
Schema Generation for Large Knowledge Graphs Using Large Language Models
by: Zhang, Bohui, et al.
Published: (2025) -
PathE: Leveraging Entity-Agnostic Paths for Parameter-Efficient Knowledge Graph Embeddings
by: Reklos, Ioannis, et al.
Published: (2025) -
Towards Responsible AI Music: an Investigation of Trustworthy Features for Creative Systems
by: de Berardinis, Jacopo, et al.
Published: (2025) -
How Far Can Pretrained LLMs Go in Symbolic Music? Controlled Comparisons of Supervised and Preference-based Adaptation
by: Kumar, Deepak, et al.
Published: (2026) -
LLMs for Relational Reasoning: How Far are We?
by: Li, Zhiming, et al.
Published: (2024)