Huxley-Gödel Machine: Human-Level Coding Agent Development by an Approximation of the Optimal Self-Improving Machine
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Wenyi, Piękos, Piotr, Nanbo, Li, Laakom, Firas, Chen, Yimeng, Ostaszewski, Mateusz, Zhuge, Mingchen, Schmidhuber, Jürgen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PhysGym: Benchmarking LLMs in Interactive Physics Discovery with Controlled Priors
di: Chen, Yimeng, et al.
Pubblicazione: (2025)
di: Chen, Yimeng, et al.
Pubblicazione: (2025)
FACTS: A Factored State-Space Framework For World Modelling
di: Nanbo, Li, et al.
Pubblicazione: (2024)
di: Nanbo, Li, et al.
Pubblicazione: (2024)
Fairness Overfitting in Machine Learning: An Information-Theoretic Perspective
di: Laakom, Firas, et al.
Pubblicazione: (2025)
di: Laakom, Firas, et al.
Pubblicazione: (2025)
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2026)
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2026)
Diagnosable Huxley-Gödel Machines Across ASDM Agent Levels
di: Turner, Jeremy Owen
Pubblicazione: (2025)
di: Turner, Jeremy Owen
Pubblicazione: (2025)
Mixture of Sparse Attention: Content-Based Learnable Sparse Attention via Expert-Choice Routing
di: Piękos, Piotr, et al.
Pubblicazione: (2025)
di: Piękos, Piotr, et al.
Pubblicazione: (2025)
Language Agents as Optimizable Graphs
di: Zhuge, Mingchen, et al.
Pubblicazione: (2024)
di: Zhuge, Mingchen, et al.
Pubblicazione: (2024)
Beyond Outlining: Heterogeneous Recursive Planning for Adaptive Long-form Writing with Language Models
di: Xiong, Ruibin, et al.
Pubblicazione: (2025)
di: Xiong, Ruibin, et al.
Pubblicazione: (2025)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
di: Csordás, Róbert, et al.
Pubblicazione: (2023)
di: Csordás, Róbert, et al.
Pubblicazione: (2023)
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents
di: Zhang, Jenny, et al.
Pubblicazione: (2025)
di: Zhang, Jenny, et al.
Pubblicazione: (2025)
Structural Prognostic Event Modeling for Multimodal Cancer Survival Analysis
di: Zhang, Yilan, et al.
Pubblicazione: (2025)
di: Zhang, Yilan, et al.
Pubblicazione: (2025)
Class-wise Generalization Error: an Information-Theoretic Analysis
di: Laakom, Firas, et al.
Pubblicazione: (2024)
di: Laakom, Firas, et al.
Pubblicazione: (2024)
The U. S. – Taiwan and the U. S. – Poland Non-Confrontational Asymmetry: A Comparative Analysis
di: Piotr Ostaszewski
Pubblicazione: (2016)
di: Piotr Ostaszewski
Pubblicazione: (2016)
Agent-as-a-Judge: Evaluate Agents with Agents
di: Zhuge, Mingchen, et al.
Pubblicazione: (2024)
di: Zhuge, Mingchen, et al.
Pubblicazione: (2024)
Pixel-Wise Color Constancy via Smoothness Techniques in Multi-Illuminant Scenes
di: Entok, Umut Cem, et al.
Pubblicazione: (2024)
di: Entok, Umut Cem, et al.
Pubblicazione: (2024)
Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control
di: Nauman, Michal, et al.
Pubblicazione: (2024)
di: Nauman, Michal, et al.
Pubblicazione: (2024)
A Case for Validation Buffer in Pessimistic Actor-Critic
di: Nauman, Michal, et al.
Pubblicazione: (2024)
di: Nauman, Michal, et al.
Pubblicazione: (2024)
Hyperbolic Residual Quantization: Discrete Representations for Data with Latent Hierarchies
di: Piękos, Piotr, et al.
Pubblicazione: (2025)
di: Piękos, Piotr, et al.
Pubblicazione: (2025)
SGM: A Statistical Godel Machine for Risk-Controlled Recursive Self-Modification
di: Wu, Xuening, et al.
Pubblicazione: (2025)
di: Wu, Xuening, et al.
Pubblicazione: (2025)
PDR: A Plug-and-Play Positional Decay Framework for LLM Pre-training Data Detection
di: Liu, Jinhan, et al.
Pubblicazione: (2026)
di: Liu, Jinhan, et al.
Pubblicazione: (2026)
Goldfish: Vision-Language Understanding of Arbitrarily Long Videos
di: Ataallah, Kirolos, et al.
Pubblicazione: (2024)
di: Ataallah, Kirolos, et al.
Pubblicazione: (2024)
Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
di: Nauman, Michal, et al.
Pubblicazione: (2024)
di: Nauman, Michal, et al.
Pubblicazione: (2024)
Gödel Agent: A Self-Referential Agent Framework for Recursive Self-Improvement
di: Yin, Xunjian, et al.
Pubblicazione: (2024)
di: Yin, Xunjian, et al.
Pubblicazione: (2024)
Identifying Key Nodes for the Influence Spread using a Machine Learning Approach
di: Stolarski, Mateusz, et al.
Pubblicazione: (2024)
di: Stolarski, Mateusz, et al.
Pubblicazione: (2024)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
ScholarChemQA: Unveiling the Power of Language Models in Chemical Research Question Answering
di: Chen, Xiuying, et al.
Pubblicazione: (2024)
di: Chen, Xiuying, et al.
Pubblicazione: (2024)
Towards Trustworthy Machine Learning in Production: An Overview of the Robustness in MLOps Approach
di: Bayram, Firas, et al.
Pubblicazione: (2024)
di: Bayram, Firas, et al.
Pubblicazione: (2024)
Self-Organising Neural Discrete Representation Learning à la Kohonen
di: Irie, Kazuki, et al.
Pubblicazione: (2023)
di: Irie, Kazuki, et al.
Pubblicazione: (2023)
RPRA: Predicting an LLM-Judge for Efficient but Performant Inference
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning
di: Surdej, Rafał, et al.
Pubblicazione: (2025)
di: Surdej, Rafał, et al.
Pubblicazione: (2025)
Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem
di: Wołczyk, Maciej, et al.
Pubblicazione: (2024)
di: Wołczyk, Maciej, et al.
Pubblicazione: (2024)
MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
di: Hong, Sirui, et al.
Pubblicazione: (2023)
di: Hong, Sirui, et al.
Pubblicazione: (2023)
Fast Camouflaged Object Detection via Edge-based Reversible Re-calibration Network
di: Ji, Ge-Peng, et al.
Pubblicazione: (2021)
di: Ji, Ge-Peng, et al.
Pubblicazione: (2021)
Interestingness as an Inductive Heuristic for Future Compression Progress
di: Herrmann, Vincent, et al.
Pubblicazione: (2026)
di: Herrmann, Vincent, et al.
Pubblicazione: (2026)
A Self-Improving Coding Agent
di: Robeyns, Maxime, et al.
Pubblicazione: (2025)
di: Robeyns, Maxime, et al.
Pubblicazione: (2025)
End-to-End Data Quality-Driven Framework for Machine Learning in Production Environment
di: Bayram, Firas, et al.
Pubblicazione: (2025)
di: Bayram, Firas, et al.
Pubblicazione: (2025)
Padé Approximants for Fast Radio Bursts Diffuse Dispersion Measure
di: Kalomenopoulos, Marios, et al.
Pubblicazione: (2025)
di: Kalomenopoulos, Marios, et al.
Pubblicazione: (2025)
Learned Scalable Video Coding For Humans and Machines
di: Hadizadeh, Hadi, et al.
Pubblicazione: (2023)
di: Hadizadeh, Hadi, et al.
Pubblicazione: (2023)
Strategies of Code-switching in Human-Machine Dialogs
di: Geckt, Dean, et al.
Pubblicazione: (2025)
di: Geckt, Dean, et al.
Pubblicazione: (2025)
Vibe Coding: Is Human Nature the Ghost in the Machine?
di: Knobel, Cory, et al.
Pubblicazione: (2025)
di: Knobel, Cory, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PhysGym: Benchmarking LLMs in Interactive Physics Discovery with Controlled Priors
di: Chen, Yimeng, et al.
Pubblicazione: (2025) -
FACTS: A Factored State-Space Framework For World Modelling
di: Nanbo, Li, et al.
Pubblicazione: (2024) -
Fairness Overfitting in Machine Learning: An Information-Theoretic Perspective
di: Laakom, Firas, et al.
Pubblicazione: (2025) -
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2026) -
Diagnosable Huxley-Gödel Machines Across ASDM Agent Levels
di: Turner, Jeremy Owen
Pubblicazione: (2025)