Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Grosnit, Antoine, Maraval, Alexandre, N, Refinath S, Zhao, Zichao, Doran, James, Paolo, Giuseppe, Thomas, Albert, Gonzalez, Jonas, Kumar, Abhineet, Khandelwal, Khyati, Benechehab, Abdelhakim, Cherkaoui, Hamza, El-Hili, Youssef Attia, Shao, Kun, Hao, Jianye, Yao, Jun, Kégl, Balázs, Bou-Ammar, Haitham, Wang, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
by: Hili, Youssef Attia El, et al.
Published: (2025)
by: Hili, Youssef Attia El, et al.
Published: (2025)
Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
by: Paolo, Giuseppe, et al.
Published: (2025)
by: Paolo, Giuseppe, et al.
Published: (2025)
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation
by: Benechehab, Abdelhakim, et al.
Published: (2025)
by: Benechehab, Abdelhakim, et al.
Published: (2025)
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
by: Tutunov, Rasul, et al.
Published: (2025)
by: Tutunov, Rasul, et al.
Published: (2025)
Zero-shot Model-based Reinforcement Learning using Large Language Models
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information
by: Tutnov, Rasul, et al.
Published: (2025)
by: Tutnov, Rasul, et al.
Published: (2025)
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
Why Can Large Language Models Generate Correct Chain-of-Thoughts?
by: Tutunov, Rasul, et al.
Published: (2023)
by: Tutunov, Rasul, et al.
Published: (2023)
Contextual Causal Bayesian Optimisation
by: Arsenyan, Vahan, et al.
Published: (2023)
by: Arsenyan, Vahan, et al.
Published: (2023)
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
by: Benechehab, Abdelhakim, et al.
Published: (2025)
by: Benechehab, Abdelhakim, et al.
Published: (2025)
ShortCircuit: AlphaZero-Driven Circuit Design
by: Tsaras, Dimitrios, et al.
Published: (2024)
by: Tsaras, Dimitrios, et al.
Published: (2024)
UTICA: Multi-Objective Self-Distllation Foundation Model Pretraining for Time Series Classification
by: Moakher, Yessin, et al.
Published: (2026)
by: Moakher, Yessin, et al.
Published: (2026)
Appréciation de l'état de fraîcheur des poissons marins : comparaison entre la méthode organoleptique chiffrée et le dosage de l'azote basique volatil total.
by: Attia El Hili, Hedia
Published: (1988)
by: Attia El Hili, Hedia
Published: (1988)
Pathologie des poissons marins élevés en Tunisie.
by: Attia El Hili, Hedia
Published: (1996)
by: Attia El Hili, Hedia
Published: (1996)
Effets d'une alimentation à base d'ensilage chez la daurade.
by: Attia El Hili, Hedia.
Published: (1989)
by: Attia El Hili, Hedia.
Published: (1989)
Bottlenecked Transformers: Periodic KV Cache Consolidation for Generalised Reasoning
by: Oomerjee, Adnan, et al.
Published: (2025)
by: Oomerjee, Adnan, et al.
Published: (2025)
Can LLMs predict the convergence of Stochastic Gradient Descent?
by: Zekri, Oussama, et al.
Published: (2024)
by: Zekri, Oussama, et al.
Published: (2024)
Al-Khwarizmi: Discovering Physical Laws with Foundation Models
by: Mower, Christopher E., et al.
Published: (2025)
by: Mower, Christopher E., et al.
Published: (2025)
Kolb's Experiential Learning in Action: A Curriculum for Residents
by: Kelci Butler, et al.
Published: (2026)
by: Kelci Butler, et al.
Published: (2026)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
by: Christopoulou, Fenia, et al.
Published: (2024)
by: Christopoulou, Fenia, et al.
Published: (2024)
Why the Brain Consolidates: Predictive Forgetting for Optimal Generalisation
by: Fountas, Zafeirios, et al.
Published: (2026)
by: Fountas, Zafeirios, et al.
Published: (2026)
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024)
by: Zimmer, Matthieu, et al.
Published: (2024)
Subjective Depth and Timescale Transformers: Learning Where and When to Compute
by: Wieser, Frederico, et al.
Published: (2025)
by: Wieser, Frederico, et al.
Published: (2025)
Data-driven Interpretable Hybrid Robot Dynamics
by: Mower, Christopher E., et al.
Published: (2025)
by: Mower, Christopher E., et al.
Published: (2025)
High-Dimensional Analysis of Bootstrap Ensemble Classifiers
by: Tiomoko, Malik, et al.
Published: (2025)
by: Tiomoko, Malik, et al.
Published: (2025)
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
Untangling Component Imbalance in Hybrid Linear Attention Conversion Methods
by: Benfeghoul, Martin, et al.
Published: (2025)
by: Benfeghoul, Martin, et al.
Published: (2025)
Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
by: Ji, Xiaotong, et al.
Published: (2026)
by: Ji, Xiaotong, et al.
Published: (2026)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
by: Liu, Puze, et al.
Published: (2024)
by: Liu, Puze, et al.
Published: (2024)
Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution Sharpening
by: Ji, Xiaotong, et al.
Published: (2026)
by: Ji, Xiaotong, et al.
Published: (2026)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
On Almost Surely Safe Alignment of Large Language Models at Inference-Time
by: Ji, Xiaotong, et al.
Published: (2025)
by: Ji, Xiaotong, et al.
Published: (2025)
SuRe: Surprise-Driven Prioritised Replay for Continual LLM Learning
by: Hazard, Hugo, et al.
Published: (2025)
by: Hazard, Hugo, et al.
Published: (2025)
A call for embodied AI
by: Paolo, Giuseppe, et al.
Published: (2024)
by: Paolo, Giuseppe, et al.
Published: (2024)
La religion de Constantin
by: Pierre Maraval
Published: (2013)
by: Pierre Maraval
Published: (2013)
Efficient Reinforcement Learning with Large Language Model Priors
by: Yan, Xue, et al.
Published: (2024)
by: Yan, Xue, et al.
Published: (2024)
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
by: Nguyen, Tu, et al.
Published: (2026)
by: Nguyen, Tu, et al.
Published: (2026)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
by: Bourigault, Pauline, et al.
Published: (2026)
by: Bourigault, Pauline, et al.
Published: (2026)
The $\mathbf{Y}$-Combinator for LLMs: Solving Long-Context Rot with $λ$-Calculus
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Similar Items
-
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
by: Hili, Youssef Attia El, et al.
Published: (2025) -
Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
by: Benechehab, Abdelhakim, et al.
Published: (2024) -
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
by: Paolo, Giuseppe, et al.
Published: (2025) -
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation
by: Benechehab, Abdelhakim, et al.
Published: (2025) -
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
by: Tutunov, Rasul, et al.
Published: (2025)