Memento No More: Coaching AI Agents to Master Multiple Tasks via Hints Internalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Alakuijala, Minttu, Gao, Ya, Ananov, Georgy, Kaski, Samuel, Marttinen, Pekka, Ilin, Alexander, Valpola, Harri |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Edit Knowledge, Not Just Facts via Multi-Step Reasoning over Background Stories
di: Gao, Ya, et al.
Pubblicazione: (2026)
di: Gao, Ya, et al.
Pubblicazione: (2026)
Efficient Knowledge Injection in LLMs via Self-Distillation
di: Kujanpää, Kalle, et al.
Pubblicazione: (2024)
di: Kujanpää, Kalle, et al.
Pubblicazione: (2024)
Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search
di: Dainese, Nicola, et al.
Pubblicazione: (2024)
di: Dainese, Nicola, et al.
Pubblicazione: (2024)
Recursive Decomposition with Dependencies for Generic Divide-and-Conquer Reasoning
di: Hernández-Gutiérrez, Sergio, et al.
Pubblicazione: (2025)
di: Hernández-Gutiérrez, Sergio, et al.
Pubblicazione: (2025)
Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics
di: Alakuijala, Minttu, et al.
Pubblicazione: (2024)
di: Alakuijala, Minttu, et al.
Pubblicazione: (2024)
ViPlan: A Benchmark for Visual Planning with Symbolic Predicates and Vision-Language Models
di: Merler, Matteo, et al.
Pubblicazione: (2025)
di: Merler, Matteo, et al.
Pubblicazione: (2025)
Improved Compositional Generalization by Generating Demonstrations for Meta-Learning
di: Spilsbury, Sam, et al.
Pubblicazione: (2023)
di: Spilsbury, Sam, et al.
Pubblicazione: (2023)
Tuning Qwen2.5-VL to Improve Its Web Interaction Skills
di: Yakovleva, Alexandra, et al.
Pubblicazione: (2026)
di: Yakovleva, Alexandra, et al.
Pubblicazione: (2026)
Diffusion models as probabilistic neural operators for recovering unobserved states of dynamical systems
di: Haitsiukevich, Katsiaryna, et al.
Pubblicazione: (2024)
di: Haitsiukevich, Katsiaryna, et al.
Pubblicazione: (2024)
Knowledge-augmented Graph Neural Networks with Concept-aware Attention for Adverse Drug Event Detection
di: Ji, Shaoxiong, et al.
Pubblicazione: (2023)
di: Ji, Shaoxiong, et al.
Pubblicazione: (2023)
New multimodal similarity measure for image registration via modeling local functional dependence with linear combination of learned basis functions
di: Honkamaa, Joel, et al.
Pubblicazione: (2025)
di: Honkamaa, Joel, et al.
Pubblicazione: (2025)
Towards modeling evolving longitudinal health trajectories with a transformer-based deep learning model
di: Moen, Hans, et al.
Pubblicazione: (2024)
di: Moen, Hans, et al.
Pubblicazione: (2024)
Strategies for Robust Deep Learning Based Deformable Registration
di: Honkamaa, Joel, et al.
Pubblicazione: (2025)
di: Honkamaa, Joel, et al.
Pubblicazione: (2025)
Identifiable causal inference with noisy treatment and no side information
di: Pöllänen, Antti, et al.
Pubblicazione: (2023)
di: Pöllänen, Antti, et al.
Pubblicazione: (2023)
Improving Medical Multi-modal Contrastive Learning with Expert Annotations
di: Kumar, Yogesh, et al.
Pubblicazione: (2024)
di: Kumar, Yogesh, et al.
Pubblicazione: (2024)
SITReg: Multi-resolution architecture for symmetric, inverse consistent, and topology preserving image registration
di: Honkamaa, Joel, et al.
Pubblicazione: (2023)
di: Honkamaa, Joel, et al.
Pubblicazione: (2023)
More Than Irrational: Modeling Belief-Biased Agents
di: Zhu, Yifan, et al.
Pubblicazione: (2025)
di: Zhu, Yifan, et al.
Pubblicazione: (2025)
Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
di: Zou, Zhengtao, et al.
Pubblicazione: (2025)
di: Zou, Zhengtao, et al.
Pubblicazione: (2025)
Query-Guided Self-Supervised Summarization of Nursing Notes
di: Gao, Ya, et al.
Pubblicazione: (2024)
di: Gao, Ya, et al.
Pubblicazione: (2024)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
di: Nikitin, Alexander, et al.
Pubblicazione: (2024)
di: Nikitin, Alexander, et al.
Pubblicazione: (2024)
Object-level Self-Distillation for Vision Pretraining
di: Hızlı, Çağlar, et al.
Pubblicazione: (2025)
di: Hızlı, Çağlar, et al.
Pubblicazione: (2025)
CBR-to-SQL: Rethinking Retrieval-based Text-to-SQL using Case-based Reasoning in the Healthcare Domain
di: Nguyen, Hung, et al.
Pubblicazione: (2026)
di: Nguyen, Hung, et al.
Pubblicazione: (2026)
A Bayesian method for identification of stock mixtures from molecular marker data
di: Corrander, Jukka, et al.
Pubblicazione: (2006)
di: Corrander, Jukka, et al.
Pubblicazione: (2006)
Memento-Skills: Let Agents Design Agents
di: Zhou, Huichi, et al.
Pubblicazione: (2026)
di: Zhou, Huichi, et al.
Pubblicazione: (2026)
The Legendary Saga as a Medium of Cultural Memory
di: Valpola-Walker, Alisa
Pubblicazione: (2025)
di: Valpola-Walker, Alisa
Pubblicazione: (2025)
POEMS: Product of Experts for Interpretable Multi-omic Integration using Sparse Decoding
di: Balik, Mihriban Kocak, et al.
Pubblicazione: (2025)
di: Balik, Mihriban Kocak, et al.
Pubblicazione: (2025)
MiniLingua: A Small Open-Source LLM for European Languages
di: Aksenova, Anna, et al.
Pubblicazione: (2025)
di: Aksenova, Anna, et al.
Pubblicazione: (2025)
In-Context Symbolic Regression: Leveraging Large Language Models for Function Discovery
di: Merler, Matteo, et al.
Pubblicazione: (2024)
di: Merler, Matteo, et al.
Pubblicazione: (2024)
Bayesian Meta-Learning with Expert Feedback for Task-Shift Adaptation through Causal Embeddings
di: Mäkinen, Lotta, et al.
Pubblicazione: (2026)
di: Mäkinen, Lotta, et al.
Pubblicazione: (2026)
Mémento de l'assainissement
Pubblicazione: (2019)
Pubblicazione: (2019)
La memoria fotográfica. Memento
di: Nekane Parejo
Pubblicazione: (2010)
di: Nekane Parejo
Pubblicazione: (2010)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
Non-separable Spatio-temporal Graph Kernels via SPDEs
di: Nikitin, Alexander, et al.
Pubblicazione: (2021)
di: Nikitin, Alexander, et al.
Pubblicazione: (2021)
TSGM: A Flexible Framework for Generative Modeling of Synthetic Time Series
di: Nikitin, Alexander, et al.
Pubblicazione: (2023)
di: Nikitin, Alexander, et al.
Pubblicazione: (2023)
Effects of Low‐Intensity Endurance Training on Aerobic Fitness and Risk Factors of Cardiometabolic Health in Working‐Age Adults: A Systematic Review and Meta‐Analysis
di: Olli‐Pekka Nuuttila, et al.
Pubblicazione: (2026)
di: Olli‐Pekka Nuuttila, et al.
Pubblicazione: (2026)
Identifying latent state transition in non-linear dynamical systems
di: Hızlı, Çağlar, et al.
Pubblicazione: (2024)
di: Hızlı, Çağlar, et al.
Pubblicazione: (2024)
Multi-Objective Bayesian Optimization via Adaptive \varepsilon-Constraints Decomposition
di: Yang, Yaohong, et al.
Pubblicazione: (2026)
di: Yang, Yaohong, et al.
Pubblicazione: (2026)
NavHint: Vision and Language Navigation Agent with a Hint Generator
di: Zhang, Yue, et al.
Pubblicazione: (2024)
di: Zhang, Yue, et al.
Pubblicazione: (2024)
MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents
di: Zeng, Ziyun, et al.
Pubblicazione: (2026)
di: Zeng, Ziyun, et al.
Pubblicazione: (2026)
Memento 2: Learning by Stateful Reflective Memory
di: Wang, Jun
Pubblicazione: (2025)
di: Wang, Jun
Pubblicazione: (2025)
Documenti analoghi
-
Edit Knowledge, Not Just Facts via Multi-Step Reasoning over Background Stories
di: Gao, Ya, et al.
Pubblicazione: (2026) -
Efficient Knowledge Injection in LLMs via Self-Distillation
di: Kujanpää, Kalle, et al.
Pubblicazione: (2024) -
Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search
di: Dainese, Nicola, et al.
Pubblicazione: (2024) -
Recursive Decomposition with Dependencies for Generic Divide-and-Conquer Reasoning
di: Hernández-Gutiérrez, Sergio, et al.
Pubblicazione: (2025) -
Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics
di: Alakuijala, Minttu, et al.
Pubblicazione: (2024)