Common Benchmarks Undervalue the Generalization Power of Programmatic Policies
Fuente:
arXiv
Guardado en:
| Autores principales: | Rajabpour, Amirhossein, Aghakasiri, Kiarash, Zilles, Sandra, Lelis, Levi H. S. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reclaiming the Source of Programmatic Policies: Programmatic versus Latent Spaces
por: Carvalho, Tales H., et al.
Publicado: (2024)
por: Carvalho, Tales H., et al.
Publicado: (2024)
Searching for Programmatic Policies in Semantic Spaces
por: Moraes, Rubens O., et al.
Publicado: (2024)
por: Moraes, Rubens O., et al.
Publicado: (2024)
InnateCoder: Learning Programmatic Options with Foundation Models
por: Moraes, Rubens O., et al.
Publicado: (2025)
por: Moraes, Rubens O., et al.
Publicado: (2025)
Assessing the Interpretability of Programmatic Policies with Large Language Models
por: Bashir, Zahra, et al.
Publicado: (2023)
por: Bashir, Zahra, et al.
Publicado: (2023)
Unveiling Options with Neural Decomposition
por: Alikhasi, Mahdi, et al.
Publicado: (2024)
por: Alikhasi, Mahdi, et al.
Publicado: (2024)
Monte Carlo Tree Search in the Presence of Transition Uncertainty
por: Kohankhaki, Farnaz, et al.
Publicado: (2023)
por: Kohankhaki, Farnaz, et al.
Publicado: (2023)
Gradient-Discrepancy Acquisition for Pool-Based Active Learning
por: Khosravani, Mohamadsadegh, et al.
Publicado: (2026)
por: Khosravani, Mohamadsadegh, et al.
Publicado: (2026)
Levin Tree Search with Context Models
por: Orseau, Laurent, et al.
Publicado: (2023)
por: Orseau, Laurent, et al.
Publicado: (2023)
Survival Dynamics of Neural and Programmatic Policies in Evolutionary Reinforcement Learning
por: Roupassov-Ruiz, Anton, et al.
Publicado: (2026)
por: Roupassov-Ruiz, Anton, et al.
Publicado: (2026)
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
por: Kohler, Hector, et al.
Publicado: (2024)
por: Kohler, Hector, et al.
Publicado: (2024)
Active learning from positive and unlabeled examples
por: Mansouri, Farnam, et al.
Publicado: (2026)
por: Mansouri, Farnam, et al.
Publicado: (2026)
Multimodal LLM-assisted Evolutionary Search for Programmatic Control Policies
por: Hu, Qinglong, et al.
Publicado: (2025)
por: Hu, Qinglong, et al.
Publicado: (2025)
Gradient-Based Program Synthesis with Neurally Interpreted Languages
por: Macfarlane, Matthew V., et al.
Publicado: (2026)
por: Macfarlane, Matthew V., et al.
Publicado: (2026)
The Computational Complexity of Almost Stable Clustering with Penalties
por: Khodamoradi, Kamyar, et al.
Publicado: (2025)
por: Khodamoradi, Kamyar, et al.
Publicado: (2025)
DiPRL: Learning Discrete Programmatic Policies via Architecture Entropy Regularization
por: Hu, Chengpeng, et al.
Publicado: (2026)
por: Hu, Chengpeng, et al.
Publicado: (2026)
PolicyEvolve: Evolving Programmatic Policies by LLMs for multi-player games via Population-Based Training
por: Lv, Mingrui, et al.
Publicado: (2025)
por: Lv, Mingrui, et al.
Publicado: (2025)
Programmatic Representation Learning with Language Models
por: Poesia, Gabriel, et al.
Publicado: (2025)
por: Poesia, Gabriel, et al.
Publicado: (2025)
Objective Mispricing Detection for Shortlisting Undervalued Football Players via Market Dynamics and News Signals
por: Omejieke, Chinenye, et al.
Publicado: (2026)
por: Omejieke, Chinenye, et al.
Publicado: (2026)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
por: Liu, Max, et al.
Publicado: (2024)
por: Liu, Max, et al.
Publicado: (2024)
Formal Models of Active Learning from Contrastive Examples
por: Mansouri, Farnam, et al.
Publicado: (2025)
por: Mansouri, Farnam, et al.
Publicado: (2025)
Deep Generative Models for Subgraph Prediction
por: Mahmoudzadeh, Erfaneh, et al.
Publicado: (2024)
por: Mahmoudzadeh, Erfaneh, et al.
Publicado: (2024)
Distance-based Learning of Hypertrees
por: Fallat, Shaun, et al.
Publicado: (2025)
por: Fallat, Shaun, et al.
Publicado: (2025)
LLM as an Algorithmist: Enhancing Anomaly Detectors via Programmatic Synthesis
por: Ye, Hangting, et al.
Publicado: (2025)
por: Ye, Hangting, et al.
Publicado: (2025)
Reliable Programmatic Weak Supervision with Confidence Intervals for Label Probabilities
por: Álvarez, Verónica, et al.
Publicado: (2025)
por: Álvarez, Verónica, et al.
Publicado: (2025)
Sketch-Plan-Generalize: Learning and Planning with Neuro-Symbolic Programmatic Representations for Inductive Spatial Concepts
por: Kalithasan, Namasivayam, et al.
Publicado: (2024)
por: Kalithasan, Namasivayam, et al.
Publicado: (2024)
Leveraging Programmatically Generated Synthetic Data for Differentially Private Diffusion Training
por: Choi, Yujin, et al.
Publicado: (2024)
por: Choi, Yujin, et al.
Publicado: (2024)
GLASS: Global-Local Aggregation for Inference-time Sparsification of LLMs
por: Sattarifard, Amirmohsen, et al.
Publicado: (2025)
por: Sattarifard, Amirmohsen, et al.
Publicado: (2025)
Learning Half-Spaces from Perturbed Contrastive Examples
por: Ravari, Aryan Alavi Razavi, et al.
Publicado: (2026)
por: Ravari, Aryan Alavi Razavi, et al.
Publicado: (2026)
Programmatic Reinforcement Learning: Navigating Gridworlds
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
SPELL: Synthesis of Programmatic Edits using LLMs
por: Ramos, Daniel, et al.
Publicado: (2026)
por: Ramos, Daniel, et al.
Publicado: (2026)
Revisiting the Necessity of Graph Learning and Common Graph Benchmarks
por: Katsman, Isay, et al.
Publicado: (2024)
por: Katsman, Isay, et al.
Publicado: (2024)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
por: Reid, Cameron, et al.
Publicado: (2025)
por: Reid, Cameron, et al.
Publicado: (2025)
MORSE-500: A Programmatically Controllable Video Benchmark to Stress-Test Multimodal Reasoning
por: Cai, Zikui, et al.
Publicado: (2025)
por: Cai, Zikui, et al.
Publicado: (2025)
Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation
por: Hu, Haichen, et al.
Publicado: (2026)
por: Hu, Haichen, et al.
Publicado: (2026)
Type-Compliant Adaptation Cascades: Adapting Programmatic LM Workflows to Data
por: Lin, Chu-Cheng, et al.
Publicado: (2025)
por: Lin, Chu-Cheng, et al.
Publicado: (2025)
PoE-World: Compositional World Modeling with Products of Programmatic Experts
por: Piriyakulkij, Wasu Top, et al.
Publicado: (2025)
por: Piriyakulkij, Wasu Top, et al.
Publicado: (2025)
Scheduling That Speaks: An Interpretable Programmatic Reinforcement Learning Framework
por: Hu, Chengpeng, et al.
Publicado: (2026)
por: Hu, Chengpeng, et al.
Publicado: (2026)
CASSANDRA: Programmatic and Probabilistic Learning and Inference for Stochastic World Modeling
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2026)
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2026)
Distributed Detection of Adversarial Attacks in Multi-Agent Reinforcement Learning with Continuous Action Space
por: Kazari, Kiarash, et al.
Publicado: (2025)
por: Kazari, Kiarash, et al.
Publicado: (2025)
An Operator Splitting View of Federated Learning
por: Malekmohammadi, Saber, et al.
Publicado: (2021)
por: Malekmohammadi, Saber, et al.
Publicado: (2021)
Ejemplares similares
-
Reclaiming the Source of Programmatic Policies: Programmatic versus Latent Spaces
por: Carvalho, Tales H., et al.
Publicado: (2024) -
Searching for Programmatic Policies in Semantic Spaces
por: Moraes, Rubens O., et al.
Publicado: (2024) -
InnateCoder: Learning Programmatic Options with Foundation Models
por: Moraes, Rubens O., et al.
Publicado: (2025) -
Assessing the Interpretability of Programmatic Policies with Large Language Models
por: Bashir, Zahra, et al.
Publicado: (2023) -
Unveiling Options with Neural Decomposition
por: Alikhasi, Mahdi, et al.
Publicado: (2024)