Searching for Programmatic Policies in Semantic Spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Moraes, Rubens O., Lelis, Levi H. S. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reclaiming the Source of Programmatic Policies: Programmatic versus Latent Spaces
by: Carvalho, Tales H., et al.
Published: (2024)
by: Carvalho, Tales H., et al.
Published: (2024)
Assessing the Interpretability of Programmatic Policies with Large Language Models
by: Bashir, Zahra, et al.
Published: (2023)
by: Bashir, Zahra, et al.
Published: (2023)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
by: Liu, Max, et al.
Published: (2024)
by: Liu, Max, et al.
Published: (2024)
InnateCoder: Learning Programmatic Options with Foundation Models
by: Moraes, Rubens O., et al.
Published: (2025)
by: Moraes, Rubens O., et al.
Published: (2025)
Levin Tree Search with Context Models
by: Orseau, Laurent, et al.
Published: (2023)
by: Orseau, Laurent, et al.
Published: (2023)
EcoSearch: A Constant-Delay Best-First Search Algorithm for Program Synthesis
by: Matricon, Théo, et al.
Published: (2024)
by: Matricon, Théo, et al.
Published: (2024)
Unveiling Options with Neural Decomposition
by: Alikhasi, Mahdi, et al.
Published: (2024)
by: Alikhasi, Mahdi, et al.
Published: (2024)
Program Semantic Inequivalence Game with Large Language Models
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2025)
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2025)
EnCompass: Enhancing Agent Programming with Search Over Program Execution Paths
by: Li, Zhening, et al.
Published: (2025)
by: Li, Zhening, et al.
Published: (2025)
Emergent Representations of Program Semantics in Language Models Trained on Programs
by: Jin, Charles, et al.
Published: (2023)
by: Jin, Charles, et al.
Published: (2023)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
by: Barone, Antonio Valerio Miceli, et al.
Published: (2026)
by: Barone, Antonio Valerio Miceli, et al.
Published: (2026)
Common Benchmarks Undervalue the Generalization Power of Programmatic Policies
by: Rajabpour, Amirhossein, et al.
Published: (2025)
by: Rajabpour, Amirhossein, et al.
Published: (2025)
BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate
by: Mazza, Arnon, et al.
Published: (2026)
by: Mazza, Arnon, et al.
Published: (2026)
Forklift: An Extensible Neural Lifter
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
Program Machine Policy: Addressing Long-Horizon Tasks by Integrating Program Synthesis and State Machines
by: Lin, Yu-An, et al.
Published: (2023)
by: Lin, Yu-An, et al.
Published: (2023)
What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces
by: Armengol-Estapé, Jordi, et al.
Published: (2025)
by: Armengol-Estapé, Jordi, et al.
Published: (2025)
ProofOptimizer: Training Language Models to Simplify Proofs without Human Demonstrations
by: Gu, Alex, et al.
Published: (2025)
by: Gu, Alex, et al.
Published: (2025)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
by: Song, Xiaoshuai, et al.
Published: (2026)
by: Song, Xiaoshuai, et al.
Published: (2026)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
by: Kohler, Hector, et al.
Published: (2024)
by: Kohler, Hector, et al.
Published: (2024)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
by: Pourreza, Mohammadreza, et al.
Published: (2025)
by: Pourreza, Mohammadreza, et al.
Published: (2025)
The Elements of Differentiable Programming
by: Blondel, Mathieu, et al.
Published: (2024)
by: Blondel, Mathieu, et al.
Published: (2024)
An LLM-Tool Compiler for Fused Parallel Function Calling
by: Singh, Simranjit, et al.
Published: (2024)
by: Singh, Simranjit, et al.
Published: (2024)
Probabilistic Programming with Programmable Variational Inference
by: Becker, McCoy R., et al.
Published: (2024)
by: Becker, McCoy R., et al.
Published: (2024)
Curriculum Learning for Small Code Language Models
by: Naïr, Marwa, et al.
Published: (2024)
by: Naïr, Marwa, et al.
Published: (2024)
SoD$^2$: Statically Optimizing Dynamic Deep Neural Network
by: Niu, Wei, et al.
Published: (2024)
by: Niu, Wei, et al.
Published: (2024)
Mirage: A Multi-Level Superoptimizer for Tensor Programs
by: Wu, Mengdi, et al.
Published: (2024)
by: Wu, Mengdi, et al.
Published: (2024)
Towards LLM-based optimization compilers. Can LLMs learn how to apply a single peephole optimization? Reasoning is all LLMs need!
by: Fang, Xiangxin, et al.
Published: (2024)
by: Fang, Xiangxin, et al.
Published: (2024)
depyf: Open the Opaque Box of PyTorch Compiler for Machine Learning Researchers
by: You, Kaichao, et al.
Published: (2024)
by: You, Kaichao, et al.
Published: (2024)
Program Synthesis using Inductive Logic Programming for the Abstraction and Reasoning Corpus
by: Rocha, Filipe Marinho, et al.
Published: (2024)
by: Rocha, Filipe Marinho, et al.
Published: (2024)
REASONING COMPILER: LLM-Guided Optimizations for Efficient Model Serving
by: Tang, Annabelle Sujun, et al.
Published: (2025)
by: Tang, Annabelle Sujun, et al.
Published: (2025)
ANCORA: Learning to Question via Manifold-Anchored Self-Play for Verifiable Reasoning
by: Yang, Chengcao
Published: (2026)
by: Yang, Chengcao
Published: (2026)
AutoPDL: Automatic Prompt Optimization for LLM Agents
by: Spiess, Claudio, et al.
Published: (2025)
by: Spiess, Claudio, et al.
Published: (2025)
DVM: A Bytecode Virtual Machine Approach for Dynamic Tensor Computation
by: Fang, Jingzhi, et al.
Published: (2026)
by: Fang, Jingzhi, et al.
Published: (2026)
PassNet: Scaling Large Language Models for Graph Compiler Pass Generation
by: Liu, Yiqun, et al.
Published: (2026)
by: Liu, Yiqun, et al.
Published: (2026)
Hexcute: A Compiler Framework for Automating Layout Synthesis in GPU Programs
by: Zhang, Xiao, et al.
Published: (2025)
by: Zhang, Xiao, et al.
Published: (2025)
Relax: Composable Abstractions for End-to-End Dynamic Machine Learning
by: Lai, Ruihang, et al.
Published: (2023)
by: Lai, Ruihang, et al.
Published: (2023)
From Reasoning to Code: GRPO Optimization for Underrepresented Languages
by: Pennino, Federico, et al.
Published: (2025)
by: Pennino, Federico, et al.
Published: (2025)
Magellan: Autonomous Discovery of Novel Compiler Optimization Heuristics with AlphaEvolve
by: Chen, Hongzheng, et al.
Published: (2026)
by: Chen, Hongzheng, et al.
Published: (2026)
Similar Items
-
Reclaiming the Source of Programmatic Policies: Programmatic versus Latent Spaces
by: Carvalho, Tales H., et al.
Published: (2024) -
Assessing the Interpretability of Programmatic Policies with Large Language Models
by: Bashir, Zahra, et al.
Published: (2023) -
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
by: Liu, Max, et al.
Published: (2024) -
InnateCoder: Learning Programmatic Options with Foundation Models
by: Moraes, Rubens O., et al.
Published: (2025) -
Levin Tree Search with Context Models
by: Orseau, Laurent, et al.
Published: (2023)