Code-Driven Planning in Grid Worlds with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Aravindan, Ashwath Vaithinathan, Tang, Zhisheng, Kejriwal, Mayank |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026)
Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026)
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
An Evaluation of Estimative Uncertainty in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
A Compound AI Agent for Conversational Grant Discovery
by: Tang, Zhisheng, et al.
Published: (2026)
by: Tang, Zhisheng, et al.
Published: (2026)
Is persona enough for personality? Using ChatGPT to reconstruct an agent's latent personality from simple descriptions
by: Ji, Yongyi, et al.
Published: (2024)
by: Ji, Yongyi, et al.
Published: (2024)
Rewarding Intellectual Humility Learning When Not To Answer In Large Language Models
by: Jha, Abha, et al.
Published: (2026)
by: Jha, Abha, et al.
Published: (2026)
HALO: An Ontology for Representing and Categorizing Hallucinations in Large Language Models
by: Nananukul, Navapat, et al.
Published: (2023)
by: Nananukul, Navapat, et al.
Published: (2023)
Generating Novelty in Open-World Multi-Agent Strategic Board Games
by: Kejriwal, Mayank, et al.
Published: (2025)
by: Kejriwal, Mayank, et al.
Published: (2025)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
by: Jha, Abha, et al.
Published: (2025)
by: Jha, Abha, et al.
Published: (2025)
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference
by: Shen, Ke, et al.
Published: (2024)
by: Shen, Ke, et al.
Published: (2024)
Navigating Semantic Relations: Challenges for Language Models in Abstract Common-Sense Reasoning
by: Gawin, Cole, et al.
Published: (2025)
by: Gawin, Cole, et al.
Published: (2025)
ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations
by: Nananukul, Navapat, et al.
Published: (2026)
by: Nananukul, Navapat, et al.
Published: (2026)
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL
by: Shen, Ke, et al.
Published: (2024)
by: Shen, Ke, et al.
Published: (2024)
Do VLMs Have Bad Eyes? Diagnosing Compositional Failures via Mechanistic Interpretability
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
Structural shifts in institutional participation and collaboration within the AI arXiv preprint research ecosystem
by: Maganur, Shama, et al.
Published: (2026)
by: Maganur, Shama, et al.
Published: (2026)
An Analysis of Artificial Intelligence Adoption in NIH-Funded Research
by: Nananukul, Navapat, et al.
Published: (2026)
by: Nananukul, Navapat, et al.
Published: (2026)
Cost-Efficient Prompt Engineering for Unsupervised Entity Resolution
by: Nananukul, Navapat, et al.
Published: (2023)
by: Nananukul, Navapat, et al.
Published: (2023)
Large-Scale Multirobot Coverage Path Planning on Grids With Path Deconfliction
by: Tang, Jingtao, et al.
Published: (2024)
by: Tang, Jingtao, et al.
Published: (2024)
Vector Quantization in the Brain: Grid-like Codes in World Models
by: Peng, Xiangyuan, et al.
Published: (2025)
by: Peng, Xiangyuan, et al.
Published: (2025)
Enhance Large Language Models as Recommendation Systems with Collaborative Filtering
by: Yang, Zhisheng, et al.
Published: (2025)
by: Yang, Zhisheng, et al.
Published: (2025)
GridCodex: A RAG-Driven AI Framework for Power Grid Code Reasoning and Compliance
by: Shi, Jinquan, et al.
Published: (2025)
by: Shi, Jinquan, et al.
Published: (2025)
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
Theory Discovery in Social Networks: Automating ERGM Specification with Large Language Models
by: Sun, Yidan, et al.
Published: (2026)
by: Sun, Yidan, et al.
Published: (2026)
Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search
by: Dainese, Nicola, et al.
Published: (2024)
by: Dainese, Nicola, et al.
Published: (2024)
World-aware Planning Narratives Enhance Large Vision-Language Model Planner
by: Shi, Junhao, et al.
Published: (2025)
by: Shi, Junhao, et al.
Published: (2025)
Scalable Task Planning via Large Language Models and Structured World Representations
by: Pérez-Dattari, Rodrigo, et al.
Published: (2024)
by: Pérez-Dattari, Rodrigo, et al.
Published: (2024)
Code Driven Planning with Domain-Adaptive Critic
by: Tian, Zikang, et al.
Published: (2025)
by: Tian, Zikang, et al.
Published: (2025)
Planning-Driven Programming: A Large Language Model Programming Workflow
by: Lei, Chao, et al.
Published: (2024)
by: Lei, Chao, et al.
Published: (2024)
Large Language Model-Driven Code Compliance Checking in Building Information Modeling
by: Madireddy, Soumya, et al.
Published: (2025)
by: Madireddy, Soumya, et al.
Published: (2025)
Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents
by: Wang, Zihao, et al.
Published: (2023)
by: Wang, Zihao, et al.
Published: (2023)
Measuring and Mitigating Bias in Code Generated by Large Language Models
by: Chen, Yuxi, et al.
Published: (2026)
by: Chen, Yuxi, et al.
Published: (2026)
Planning with Reasoning using Vision Language World Model
by: Chen, Delong, et al.
Published: (2025)
by: Chen, Delong, et al.
Published: (2025)
HaM-World: Soft-Hamiltonian World Models with Selective Memory for Planning
by: Tang, Haoyun, et al.
Published: (2026)
by: Tang, Haoyun, et al.
Published: (2026)
AutoB2G: A Large Language Model-Driven Agentic Framework For Automated Building-Grid Co-Simulation
by: Zhang, Borui, et al.
Published: (2026)
by: Zhang, Borui, et al.
Published: (2026)
LOGicalThought: Logic-Based Ontological Grounding of LLMs for High-Assurance Reasoning
by: Nananukul, Navapat, et al.
Published: (2025)
by: Nananukul, Navapat, et al.
Published: (2025)
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
by: Sheth, Rajvee, et al.
Published: (2025)
by: Sheth, Rajvee, et al.
Published: (2025)
Fault Diagnosis in Power Grids with Large Language Model
by: Jing, Liu, et al.
Published: (2024)
by: Jing, Liu, et al.
Published: (2024)
Design Behaviour Codes (DBCs): A Taxonomy-Driven Layered Governance Benchmark for Large Language Models
by: Mohan, G. Madan, et al.
Published: (2026)
by: Mohan, G. Madan, et al.
Published: (2026)
Similar Items
-
Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026) -
Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2026) -
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024) -
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning
by: Tang, Zhisheng, et al.
Published: (2024) -
An Evaluation of Estimative Uncertainty in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024)