Improved Generalized Planning with LLMs through Strategy Refinement and Reflection
Fuente:
arXiv
Guardado en:
| Autores principales: | Stein, Katharina, Hodel, Nils, Fišer, Daniel, Hoffmann, Jörg, Katz, Michael, Koller, Alexander |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automating the Generation of Prompts for LLM-based Action Choice in PDDL Planning
por: Stein, Katharina, et al.
Publicado: (2023)
por: Stein, Katharina, et al.
Publicado: (2023)
On the Ability of Transformers to Verify Plans
por: Sarrof, Yash, et al.
Publicado: (2026)
por: Sarrof, Yash, et al.
Publicado: (2026)
Response: Emergent analogical reasoning in large language models
por: Hodel, Damian, et al.
Publicado: (2023)
por: Hodel, Damian, et al.
Publicado: (2023)
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
por: Huang, Yue, et al.
Publicado: (2025)
por: Huang, Yue, et al.
Publicado: (2025)
Large Language Models as Planning Domain Generators
por: Oswald, James, et al.
Publicado: (2024)
por: Oswald, James, et al.
Publicado: (2024)
Planning without Search: Refining Frontier LLMs with Offline Goal-Conditioned RL
por: Hong, Joey, et al.
Publicado: (2025)
por: Hong, Joey, et al.
Publicado: (2025)
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
por: Zhang, Hongbo, et al.
Publicado: (2025)
por: Zhang, Hongbo, et al.
Publicado: (2025)
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
por: Thillainathan, Sarubi, et al.
Publicado: (2026)
por: Thillainathan, Sarubi, et al.
Publicado: (2026)
Planning in the LLM Era: Building for Reliability and Efficiency
por: Katz, Michael, et al.
Publicado: (2026)
por: Katz, Michael, et al.
Publicado: (2026)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
por: Lee, Kyungjae, et al.
Publicado: (2024)
por: Lee, Kyungjae, et al.
Publicado: (2024)
AgentRefine: Enhancing Agent Generalization through Refinement Tuning
por: Fu, Dayuan, et al.
Publicado: (2025)
por: Fu, Dayuan, et al.
Publicado: (2025)
Reflection-Window Decoding: Text Generation with Selective Refinement
por: Tang, Zeyu, et al.
Publicado: (2025)
por: Tang, Zeyu, et al.
Publicado: (2025)
Planning Ahead with RSA: Efficient Signalling in Dynamic Environments by Projecting User Awareness across Future Timesteps
por: Das, Anwesha, et al.
Publicado: (2025)
por: Das, Anwesha, et al.
Publicado: (2025)
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation
por: Zhou, Changzhi, et al.
Publicado: (2025)
por: Zhou, Changzhi, et al.
Publicado: (2025)
GenKnowSub: Improving Modularity and Reusability of LLMs through General Knowledge Subtraction
por: Bagherifard, Mohammadtaha, et al.
Publicado: (2025)
por: Bagherifard, Mohammadtaha, et al.
Publicado: (2025)
ADaPT: As-Needed Decomposition and Planning with Language Models
por: Prasad, Archiki, et al.
Publicado: (2023)
por: Prasad, Archiki, et al.
Publicado: (2023)
Mission Impossible: Feedback-Guided Dynamic Interactive Planning for Improving Reasoning on LLMs
por: Yan, Dong, et al.
Publicado: (2025)
por: Yan, Dong, et al.
Publicado: (2025)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
por: Nguyen, Van Bach, et al.
Publicado: (2024)
por: Nguyen, Van Bach, et al.
Publicado: (2024)
MAMM-Refine: A Recipe for Improving Faithfulness in Generation with Multi-Agent Collaboration
por: Wan, David, et al.
Publicado: (2025)
por: Wan, David, et al.
Publicado: (2025)
Iterative Deployment Improves Planning Skills in LLMs
por: Corrêa, Augusto B., et al.
Publicado: (2025)
por: Corrêa, Augusto B., et al.
Publicado: (2025)
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
por: Yun, Taewon, et al.
Publicado: (2025)
por: Yun, Taewon, et al.
Publicado: (2025)
Guided Profile Generation Improves Personalization with LLMs
por: Zhang, Jiarui
Publicado: (2024)
por: Zhang, Jiarui
Publicado: (2024)
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs
por: Kumar, Somnath, et al.
Publicado: (2023)
por: Kumar, Somnath, et al.
Publicado: (2023)
Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
por: Tsai, Yu-Che, et al.
Publicado: (2025)
por: Tsai, Yu-Che, et al.
Publicado: (2025)
Generating Planning Feedback for Open-Ended Programming Exercises with LLMs
por: Demirtaş, Mehmet Arif, et al.
Publicado: (2025)
por: Demirtaş, Mehmet Arif, et al.
Publicado: (2025)
Search and Refine During Think: Facilitating Knowledge Refinement for Improved Retrieval-Augmented Reasoning
por: Shi, Yaorui, et al.
Publicado: (2025)
por: Shi, Yaorui, et al.
Publicado: (2025)
PCQPR: Proactive Conversational Question Planning with Reflection
por: Guo, Shasha, et al.
Publicado: (2024)
por: Guo, Shasha, et al.
Publicado: (2024)
Unlocking Recursive Thinking of LLMs: Alignment via Refinement
por: Zhang, Haoke, et al.
Publicado: (2025)
por: Zhang, Haoke, et al.
Publicado: (2025)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
por: Wang, Qibin, et al.
Publicado: (2025)
por: Wang, Qibin, et al.
Publicado: (2025)
LLM driven Text-to-Table Generation through Sub-Tasks Guidance and Iterative Refinement
por: C, Rajmohan, et al.
Publicado: (2025)
por: C, Rajmohan, et al.
Publicado: (2025)
Self-Improving Customer Review Response Generation Based on LLMs
por: Azov, Guy, et al.
Publicado: (2024)
por: Azov, Guy, et al.
Publicado: (2024)
KL3M Tokenizers: A Family of Domain-Specific and Character-Level Tokenizers for Legal, Financial, and Preprocessing Applications
por: Bommarito, Michael J, et al.
Publicado: (2025)
por: Bommarito, Michael J, et al.
Publicado: (2025)
Generative Floor Plan Design with LLMs via Reinforcement Learning with Verifiable Rewards
por: Lara, Luis, et al.
Publicado: (2026)
por: Lara, Luis, et al.
Publicado: (2026)
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs
por: Wachter, Jasmin, et al.
Publicado: (2025)
por: Wachter, Jasmin, et al.
Publicado: (2025)
Chasing Progress, Not Perfection: Revisiting Strategies for End-to-End LLM Plan Generation
por: Huang, Sukai, et al.
Publicado: (2024)
por: Huang, Sukai, et al.
Publicado: (2024)
Structured Thinking Matters: Improving LLMs Generalization in Causal Inference Tasks
por: Sun, Wentao, et al.
Publicado: (2025)
por: Sun, Wentao, et al.
Publicado: (2025)
PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs
por: Gungordu, Oguzhan, et al.
Publicado: (2026)
por: Gungordu, Oguzhan, et al.
Publicado: (2026)
Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations
por: Tomani, Christian, et al.
Publicado: (2024)
por: Tomani, Christian, et al.
Publicado: (2024)
Refining Salience-Aware Sparse Fine-Tuning Strategies for Language Models
por: Liu, Xinxin, et al.
Publicado: (2024)
por: Liu, Xinxin, et al.
Publicado: (2024)
Semantic Refinement with LLMs for Graph Representations
por: Thapaliya, Safal, et al.
Publicado: (2025)
por: Thapaliya, Safal, et al.
Publicado: (2025)
Ejemplares similares
-
Automating the Generation of Prompts for LLM-based Action Choice in PDDL Planning
por: Stein, Katharina, et al.
Publicado: (2023) -
On the Ability of Transformers to Verify Plans
por: Sarrof, Yash, et al.
Publicado: (2026) -
Response: Emergent analogical reasoning in large language models
por: Hodel, Damian, et al.
Publicado: (2023) -
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
por: Huang, Yue, et al.
Publicado: (2025) -
Large Language Models as Planning Domain Generators
por: Oswald, James, et al.
Publicado: (2024)