Out-of-Distribution Generalization in the ARC-AGI Domain: Comparing Execution-Guided Neural Program Synthesis and Test-Time Fine-Tuning
Fuente:
arXiv
Saved in:
| Main Author: | Ouellette, Simon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Efficient Neurally-Guided Program Induction for ARC-AGI
by: Ouellette, Simon
Published: (2024)
by: Ouellette, Simon
Published: (2024)
Executable World Models for ARC-AGI-3 in the Era of Coding Agents
by: Rodionov, Sergey
Published: (2026)
by: Rodionov, Sergey
Published: (2026)
Neural Cellular Automata for ARC-AGI
by: Xu, Kevin, et al.
Published: (2025)
by: Xu, Kevin, et al.
Published: (2025)
ARC-AGI-2 Technical Report
by: de Oliveira, Wallyson Lemes, et al.
Published: (2026)
by: de Oliveira, Wallyson Lemes, et al.
Published: (2026)
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
by: Lehmann, Jens, et al.
Published: (2026)
by: Lehmann, Jens, et al.
Published: (2026)
Tiny Recursive Models on ARC-AGI-1: Inductive Biases, Identity Conditioning, and Test-Time Compute
by: Roye-Azar, Antonio, et al.
Published: (2025)
by: Roye-Azar, Antonio, et al.
Published: (2025)
Self-Improving Language Models for Evolutionary Program Synthesis: A Case Study on ARC-AGI
by: Pourcel, Julien, et al.
Published: (2025)
by: Pourcel, Julien, et al.
Published: (2025)
Multi-Perspective Transformers in ARC-AGI-2 Challenge
by: Talley, Caleb, et al.
Published: (2026)
by: Talley, Caleb, et al.
Published: (2026)
System 2 Reasoning for Human-AI Alignment: Generality and Adaptivity via ARC-AGI
by: Kim, Sejin, et al.
Published: (2024)
by: Kim, Sejin, et al.
Published: (2024)
Graph-Based Exploration for ARC-AGI-3 Interactive Reasoning Tasks
by: Rudakov, Evgenii, et al.
Published: (2025)
by: Rudakov, Evgenii, et al.
Published: (2025)
ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence
by: Foundation, ARC Prize
Published: (2026)
by: Foundation, ARC Prize
Published: (2026)
Distributional AGI Safety
by: Tomašev, Nenad, et al.
Published: (2025)
by: Tomašev, Nenad, et al.
Published: (2025)
ARC-AGI-2: A New Challenge for Frontier AI Reasoning Systems
by: Chollet, Francois, et al.
Published: (2025)
by: Chollet, Francois, et al.
Published: (2025)
Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2
by: Qiu, Yu-Ning, et al.
Published: (2026)
by: Qiu, Yu-Ning, et al.
Published: (2026)
Steering Out-of-Distribution Generalization with Concept Ablation Fine-Tuning
by: Casademunt, Helena, et al.
Published: (2025)
by: Casademunt, Helena, et al.
Published: (2025)
GenerationPrograms: Fine-grained Attribution with Executable Programs
by: Wan, David, et al.
Published: (2025)
by: Wan, David, et al.
Published: (2025)
Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3
by: Han, Liew Keong
Published: (2026)
by: Han, Liew Keong
Published: (2026)
MADIL: An MDL-based Framework for Efficient Program Synthesis in the ARC Benchmark
by: Ferré, Sébastien
Published: (2025)
by: Ferré, Sébastien
Published: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
by: Sakharova, Marina, et al.
Published: (2025)
by: Sakharova, Marina, et al.
Published: (2025)
Saarthi for AGI: Towards Domain-Specific General Intelligence for Formal Verification
by: Kumar, Aman, et al.
Published: (2026)
by: Kumar, Aman, et al.
Published: (2026)
Evaluating Generalization and Representation Stability in Small LMs via Prompting, Fine-Tuning and Out-of-Distribution Prompts
by: Raja, Rahul, et al.
Published: (2025)
by: Raja, Rahul, et al.
Published: (2025)
How Modality Shapes Perception and Reasoning: A Study of Error Propagation in ARC-AGI
by: Wen, Bo, et al.
Published: (2025)
by: Wen, Bo, et al.
Published: (2025)
The ARC of Progress towards AGI: A Living Survey of Abstraction and Reasoning
by: Vahdati, Sahar, et al.
Published: (2026)
by: Vahdati, Sahar, et al.
Published: (2026)
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
by: Shen, Hanwen, et al.
Published: (2026)
by: Shen, Hanwen, et al.
Published: (2026)
Surprisal-Guided Selection: Compute-Optimal Test-Time Strategies for Execution-Grounded Code Generation
by: Barnes, Jarrod
Published: (2026)
by: Barnes, Jarrod
Published: (2026)
The Relativity of AGI: Distributional Axioms, Fragility, and Undecidability
by: Majumdar, Angshul
Published: (2026)
by: Majumdar, Angshul
Published: (2026)
Enhancing Q&A with Domain-Specific Fine-Tuning and Iterative Reasoning: A Comparative Study
by: Nguyen, Zooey, et al.
Published: (2024)
by: Nguyen, Zooey, et al.
Published: (2024)
Out-of-Distribution Generalization for Neural Physics Solvers
by: Wei, Zhao, et al.
Published: (2026)
by: Wei, Zhao, et al.
Published: (2026)
Out-of-distribution generalisation is hard: evidence from ARC-like tasks
by: Dimitriadis, George, et al.
Published: (2025)
by: Dimitriadis, George, et al.
Published: (2025)
Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
Levels of AGI for Operationalizing Progress on the Path to AGI
by: Morris, Meredith Ringel, et al.
Published: (2023)
by: Morris, Meredith Ringel, et al.
Published: (2023)
AGI: Artificial General Intelligence for Education
by: Latif, Ehsan, et al.
Published: (2023)
by: Latif, Ehsan, et al.
Published: (2023)
Program Synthesis via Test-Time Transduction
by: Lee, Kang-il, et al.
Published: (2025)
by: Lee, Kang-il, et al.
Published: (2025)
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
by: Ni, Yuansheng, et al.
Published: (2025)
by: Ni, Yuansheng, et al.
Published: (2025)
Learning to Explore: Policy-Guided Outlier Synthesis for Graph Out-of-Distribution Detection
by: Sun, Li, et al.
Published: (2026)
by: Sun, Li, et al.
Published: (2026)
Generalizing Graph Neural Networks on Out-Of-Distribution Graphs
by: Fan, Shaohua, et al.
Published: (2021)
by: Fan, Shaohua, et al.
Published: (2021)
Pathways to AGI
by: Fletcher, Gordon, et al.
Published: (2026)
by: Fletcher, Gordon, et al.
Published: (2026)
Out-of-Distribution Generalization in Time Series: A Survey
by: Wu, Xin, et al.
Published: (2025)
by: Wu, Xin, et al.
Published: (2025)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Retrieval-Augmented Fine-Tuning With Preference Optimization For Visual Program Generation
by: Kang, Deokhyung, et al.
Published: (2025)
by: Kang, Deokhyung, et al.
Published: (2025)
Similar Items
-
Towards Efficient Neurally-Guided Program Induction for ARC-AGI
by: Ouellette, Simon
Published: (2024) -
Executable World Models for ARC-AGI-3 in the Era of Coding Agents
by: Rodionov, Sergey
Published: (2026) -
Neural Cellular Automata for ARC-AGI
by: Xu, Kevin, et al.
Published: (2025) -
ARC-AGI-2 Technical Report
by: de Oliveira, Wallyson Lemes, et al.
Published: (2026) -
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
by: Lehmann, Jens, et al.
Published: (2026)