Leveraging Environment Interaction for Automated PDDL Translation and Planning with Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Mahdavi, Sadegh, Aoki, Raquel, Tang, Keyi, Cao, Yanshuai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Jump Start or False Start? A Theoretical and Empirical Evaluation of LLM-initialized Bandits
di: Bayley, Adam, et al.
Pubblicazione: (2026)
di: Bayley, Adam, et al.
Pubblicazione: (2026)
Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks
di: Patil, Gandharv, et al.
Pubblicazione: (2026)
di: Patil, Gandharv, et al.
Pubblicazione: (2026)
End-to-end PDDL Planning with Hardcoded and Dynamic Agents
di: La Malfa, Emanuele, et al.
Pubblicazione: (2025)
di: La Malfa, Emanuele, et al.
Pubblicazione: (2025)
Memorization Capacity of Multi-Head Attention in Transformers
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2023)
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2023)
Advantage Shaping as Surrogate Reward Maximization: Unifying Pass@K Policy Gradients
di: Thrampoulidis, Christos, et al.
Pubblicazione: (2025)
di: Thrampoulidis, Christos, et al.
Pubblicazione: (2025)
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2025)
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2025)
Flora: Low-Rank Adapters Are Secretly Gradient Compressors
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
di: Carta, Thomas, et al.
Pubblicazione: (2023)
di: Carta, Thomas, et al.
Pubblicazione: (2023)
Ginger: An Efficient Curvature Approximation with Linear Complexity for General Neural Networks
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
di: Hao, Yongchang, et al.
Pubblicazione: (2024)
LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits
di: Mirzaei, Amir Reza, et al.
Pubblicazione: (2025)
di: Mirzaei, Amir Reza, et al.
Pubblicazione: (2025)
Interactive and Expressive Code-Augmented Planning with Large Language Models
di: Liu, Anthony Z., et al.
Pubblicazione: (2024)
di: Liu, Anthony Z., et al.
Pubblicazione: (2024)
From Graph Diffusion to Graph Classification
di: Xian, Jia Jun Cheng, et al.
Pubblicazione: (2024)
di: Xian, Jia Jun Cheng, et al.
Pubblicazione: (2024)
Leveraging Language Models for Automated Patient Record Linkage
di: Beheshti, Mohammad, et al.
Pubblicazione: (2025)
di: Beheshti, Mohammad, et al.
Pubblicazione: (2025)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
di: Lin, Xiaohan, et al.
Pubblicazione: (2024)
di: Lin, Xiaohan, et al.
Pubblicazione: (2024)
Leveraging Foundation Language Models (FLMs) for Automated Cohort Extraction from Large EHR Databases
di: Mugambi, Purity, et al.
Pubblicazione: (2024)
di: Mugambi, Purity, et al.
Pubblicazione: (2024)
LassoFlexNet: Flexible Neural Architecture for Tabular Data
di: Lui, Kry Yik Chau, et al.
Pubblicazione: (2026)
di: Lui, Kry Yik Chau, et al.
Pubblicazione: (2026)
One Model for All Tasks: Leveraging Efficient World Models in Multi-Task Planning
di: Pu, Yuan, et al.
Pubblicazione: (2025)
di: Pu, Yuan, et al.
Pubblicazione: (2025)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
Jump Starting Bandits with LLM-Generated Prior Knowledge
di: Alamdari, Parand A., et al.
Pubblicazione: (2024)
di: Alamdari, Parand A., et al.
Pubblicazione: (2024)
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2024)
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2024)
On Large-scale Evaluation of Embedding Models for Knowledge Graph Completion
di: Shirvani-Mahdavi, Nasim, et al.
Pubblicazione: (2025)
di: Shirvani-Mahdavi, Nasim, et al.
Pubblicazione: (2025)
Regression with Large Language Models for Materials and Molecular Property Prediction
di: Jacobs, Ryan, et al.
Pubblicazione: (2024)
di: Jacobs, Ryan, et al.
Pubblicazione: (2024)
No $D_{\text{train}}$: Model-Agnostic Counterfactual Explanations Using Reinforcement Learning
di: Sun, Xiangyu, et al.
Pubblicazione: (2024)
di: Sun, Xiangyu, et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Information Verification -- an Engineering Approach
di: Hung, Nguyen Nang, et al.
Pubblicazione: (2025)
di: Hung, Nguyen Nang, et al.
Pubblicazione: (2025)
Towards Leveraging Large Language Models for Automated Medical Q&A Evaluation
di: Krolik, Jack, et al.
Pubblicazione: (2024)
di: Krolik, Jack, et al.
Pubblicazione: (2024)
Evaluating Menu OCR and Translation: A Benchmark for Aligning Human and Automated Evaluations in Large Vision-Language Models
di: Wu, Zhanglin, et al.
Pubblicazione: (2025)
di: Wu, Zhanglin, et al.
Pubblicazione: (2025)
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization
di: Tang, Zilu, et al.
Pubblicazione: (2025)
di: Tang, Zilu, et al.
Pubblicazione: (2025)
Causal Effects with Unobserved Unit Types in Interacting Human-AI Systems
di: Overman, William, et al.
Pubblicazione: (2026)
di: Overman, William, et al.
Pubblicazione: (2026)
Evolutionary Large Language Model for Automated Feature Transformation
di: Gong, Nanxu, et al.
Pubblicazione: (2024)
di: Gong, Nanxu, et al.
Pubblicazione: (2024)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
di: Shen, Keyi, et al.
Pubblicazione: (2026)
di: Shen, Keyi, et al.
Pubblicazione: (2026)
Leveraging Large Language Models for Efficient Failure Analysis in Game Development
di: Marini, Leonardo, et al.
Pubblicazione: (2024)
di: Marini, Leonardo, et al.
Pubblicazione: (2024)
LLM4FS: Leveraging Large Language Models for Feature Selection
di: Li, Jianhao, et al.
Pubblicazione: (2025)
di: Li, Jianhao, et al.
Pubblicazione: (2025)
Leveraging Large Language Models and Topic Modeling for Toxicity Classification
di: Oskouie, Haniyeh Ehsani, et al.
Pubblicazione: (2024)
di: Oskouie, Haniyeh Ehsani, et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Automated Causal Loop Diagram Generation: Enhancing System Dynamics Modeling through Curated Prompting Techniques
di: Liu, Ning-Yuan Georgia, et al.
Pubblicazione: (2025)
di: Liu, Ning-Yuan Georgia, et al.
Pubblicazione: (2025)
Assessing the Latent Automated Program Repair Capabilities of Large Language Models using Round-Trip Translation
di: Ruiz, Fernando Vallecillos, et al.
Pubblicazione: (2024)
di: Ruiz, Fernando Vallecillos, et al.
Pubblicazione: (2024)
Integrating Large Language Models in Financial Investments and Market Analysis: A Survey
di: Mahdavi, Sedigheh, et al.
Pubblicazione: (2025)
di: Mahdavi, Sedigheh, et al.
Pubblicazione: (2025)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
QoS-QoE Translation with Large Language Model
di: Yu, Yingjie, et al.
Pubblicazione: (2026)
di: Yu, Yingjie, et al.
Pubblicazione: (2026)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Jump Start or False Start? A Theoretical and Empirical Evaluation of LLM-initialized Bandits
di: Bayley, Adam, et al.
Pubblicazione: (2026) -
Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks
di: Patil, Gandharv, et al.
Pubblicazione: (2026) -
End-to-end PDDL Planning with Hardcoded and Dynamic Agents
di: La Malfa, Emanuele, et al.
Pubblicazione: (2025) -
Memorization Capacity of Multi-Head Attention in Transformers
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2023) -
Advantage Shaping as Surrogate Reward Maximization: Unifying Pass@K Policy Gradients
di: Thrampoulidis, Christos, et al.
Pubblicazione: (2025)