Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies
Fuente:
arXiv
Salvato in:
| Autori principali: | Lou, Zhanzhi, Chen, Hui, Li, Yibo, Wang, Qian, Hooi, Bryan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
di: Li, Yibo, et al.
Pubblicazione: (2026)
di: Li, Yibo, et al.
Pubblicazione: (2026)
Enabling Self-Improving Agents to Learn at Test Time With Human-In-The-Loop Guidance
di: He, Yufei, et al.
Pubblicazione: (2025)
di: He, Yufei, et al.
Pubblicazione: (2025)
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
di: Li, Yibo, et al.
Pubblicazione: (2026)
di: Li, Yibo, et al.
Pubblicazione: (2026)
On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective
di: Zhou, Zhi, et al.
Pubblicazione: (2026)
di: Zhou, Zhi, et al.
Pubblicazione: (2026)
MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research
di: Chen, Hui, et al.
Pubblicazione: (2025)
di: Chen, Hui, et al.
Pubblicazione: (2025)
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
Federated Nested Learning: Collaborative Training of Self-Referential Memories for Test-Time Adaptation
di: Chen, Hong, et al.
Pubblicazione: (2026)
di: Chen, Hong, et al.
Pubblicazione: (2026)
A Learnability Analysis on Neuro-Symbolic Learning
di: He, Hao-Yuan, et al.
Pubblicazione: (2025)
di: He, Hao-Yuan, et al.
Pubblicazione: (2025)
DGP: A Dual-Granularity Prompting Framework for Fraud Detection with Graph-Enhanced LLMs
di: Li, Yuan, et al.
Pubblicazione: (2025)
di: Li, Yuan, et al.
Pubblicazione: (2025)
Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models
di: Sui, Yuan, et al.
Pubblicazione: (2025)
di: Sui, Yuan, et al.
Pubblicazione: (2025)
Understanding Representation Learnability of Nonlinear Self-Supervised Learning
di: Yang, Ruofeng, et al.
Pubblicazione: (2024)
di: Yang, Ruofeng, et al.
Pubblicazione: (2024)
Test Time Learning for Time Series Forecasting
di: Christou, Panayiotis, et al.
Pubblicazione: (2024)
di: Christou, Panayiotis, et al.
Pubblicazione: (2024)
Heterogeneous Multi-Agent Reinforcement Learning with Attention for Cooperative and Scalable Feature Transformation
di: Zhe, Tao, et al.
Pubblicazione: (2025)
di: Zhe, Tao, et al.
Pubblicazione: (2025)
Memory Sequence Length of Data Sampling Impacts the Adaptation of Meta-Reinforcement Learning Agents
di: Zhang, Menglong, et al.
Pubblicazione: (2024)
di: Zhang, Menglong, et al.
Pubblicazione: (2024)
Learning to Discover at Test Time
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2026)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2026)
Test-Time Learning for Large Language Models
di: Hu, Jinwu, et al.
Pubblicazione: (2025)
di: Hu, Jinwu, et al.
Pubblicazione: (2025)
Online Adaptation for Enhancing Imitation Learning Policies
di: Malato, Federico, et al.
Pubblicazione: (2024)
di: Malato, Federico, et al.
Pubblicazione: (2024)
Lyapunov-Guided Self-Alignment: Test-Time Adaptation for Offline Safe Reinforcement Learning
di: Han, Seungyub, et al.
Pubblicazione: (2026)
di: Han, Seungyub, et al.
Pubblicazione: (2026)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
di: Liu, Zongkai, et al.
Pubblicazione: (2024)
di: Liu, Zongkai, et al.
Pubblicazione: (2024)
LAC: Graph Contrastive Learning with Learnable Augmentation in Continuous Space
di: Lin, Zhenyu, et al.
Pubblicazione: (2024)
di: Lin, Zhenyu, et al.
Pubblicazione: (2024)
Sensi: Learn One Thing at a Time -- Curriculum-Based Test-Time Learning for LLM Game Agents
di: Arjmandi, Mohsen
Pubblicazione: (2026)
di: Arjmandi, Mohsen
Pubblicazione: (2026)
Policy and World Modeling Co-Training for Language Agents
di: Lu, Ning, et al.
Pubblicazione: (2026)
di: Lu, Ning, et al.
Pubblicazione: (2026)
UniTST: Effectively Modeling Inter-Series and Intra-Series Dependencies for Multivariate Time Series Forecasting
di: Liu, Juncheng, et al.
Pubblicazione: (2024)
di: Liu, Juncheng, et al.
Pubblicazione: (2024)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
di: Zhuang, Yuan, et al.
Pubblicazione: (2026)
di: Zhuang, Yuan, et al.
Pubblicazione: (2026)
Evaluating the Paperclip Maximizer: Are RL-Based Language Models More Likely to Pursue Instrumental Goals?
di: He, Yufei, et al.
Pubblicazione: (2025)
di: He, Yufei, et al.
Pubblicazione: (2025)
Active Test-Time Adaptation: Theoretical Analyses and An Algorithm
di: Gui, Shurui, et al.
Pubblicazione: (2024)
di: Gui, Shurui, et al.
Pubblicazione: (2024)
Learnable Chernoff Baselines for Inference-Time Alignment
di: Madhow, Sunil, et al.
Pubblicazione: (2026)
di: Madhow, Sunil, et al.
Pubblicazione: (2026)
Test-Time Adaptation with Binary Feedback
di: Lee, Taeckyung, et al.
Pubblicazione: (2025)
di: Lee, Taeckyung, et al.
Pubblicazione: (2025)
Monitoring Risks in Test-Time Adaptation
di: Schirmer, Mona, et al.
Pubblicazione: (2025)
di: Schirmer, Mona, et al.
Pubblicazione: (2025)
Adaptive Social Learning via Mode Policy Optimization for Language Agents
di: Wang, Minzheng, et al.
Pubblicazione: (2025)
di: Wang, Minzheng, et al.
Pubblicazione: (2025)
Augmented Contrastive Clustering with Uncertainty-Aware Prototyping for Time Series Test Time Adaptation
di: Gong, Peiliang, et al.
Pubblicazione: (2025)
di: Gong, Peiliang, et al.
Pubblicazione: (2025)
MiGrATe: Mixed-Policy GRPO for Adaptation at Test-Time
di: Phan, Peter, et al.
Pubblicazione: (2025)
di: Phan, Peter, et al.
Pubblicazione: (2025)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
di: Deng, Ailin, et al.
Pubblicazione: (2024)
di: Deng, Ailin, et al.
Pubblicazione: (2024)
Quantum-Enhanced Multi-Task Learning with Learnable Weighting for Pharmacokinetic and Toxicity Prediction
di: Zhang, Han, et al.
Pubblicazione: (2025)
di: Zhang, Han, et al.
Pubblicazione: (2025)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
di: Huang, Qian, et al.
Pubblicazione: (2023)
di: Huang, Qian, et al.
Pubblicazione: (2023)
Learning to Reason at the Frontier of Learnability
di: Foster, Thomas, et al.
Pubblicazione: (2025)
di: Foster, Thomas, et al.
Pubblicazione: (2025)
Open-World Test-Time Training: Self-Training with Contrast Learning
di: Su, Houcheng, et al.
Pubblicazione: (2024)
di: Su, Houcheng, et al.
Pubblicazione: (2024)
Rethinking Time Encoding via Learnable Transformation Functions
di: Chen, Xi, et al.
Pubblicazione: (2025)
di: Chen, Xi, et al.
Pubblicazione: (2025)
TimeCAP: Learning to Contextualize, Augment, and Predict Time Series Events with Large Language Model Agents
di: Lee, Geon, et al.
Pubblicazione: (2025)
di: Lee, Geon, et al.
Pubblicazione: (2025)
Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning
di: Zhang, Beining, et al.
Pubblicazione: (2025)
di: Zhang, Beining, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
di: Li, Yibo, et al.
Pubblicazione: (2026) -
Enabling Self-Improving Agents to Learn at Test Time With Human-In-The-Loop Guidance
di: He, Yufei, et al.
Pubblicazione: (2025) -
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
di: Li, Yibo, et al.
Pubblicazione: (2026) -
On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective
di: Zhou, Zhi, et al.
Pubblicazione: (2026) -
MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research
di: Chen, Hui, et al.
Pubblicazione: (2025)