Revisiting Dynamic Evaluation: Online Adaptation for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Rannen-Triki, Amal, Bornschein, Jorg, Pascanu, Razvan, Hutter, Marcus, György, Andras, Galashov, Alexandre, Teh, Yee Whye, Titsias, Michalis K. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Kalman Filter for Online Classification of Non-Stationary Data
di: Titsias, Michalis K., et al.
Pubblicazione: (2023)
di: Titsias, Michalis K., et al.
Pubblicazione: (2023)
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
di: Galashov, Alexandre, et al.
Pubblicazione: (2024)
di: Galashov, Alexandre, et al.
Pubblicazione: (2024)
Transformers for Supervised Online Continual Learning
di: Bornschein, Jorg, et al.
Pubblicazione: (2024)
di: Bornschein, Jorg, et al.
Pubblicazione: (2024)
Fine-Tuned In-Context Learners for Efficient Adaptation
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
di: Li, Qinyu, et al.
Pubblicazione: (2025)
di: Li, Qinyu, et al.
Pubblicazione: (2025)
Evaluating Representations with Readout Model Switching
di: Li, Yazhe, et al.
Pubblicazione: (2023)
di: Li, Yazhe, et al.
Pubblicazione: (2023)
What Can Grokking Teach Us About Learning Under Nonstationarity?
di: Lyle, Clare, et al.
Pubblicazione: (2025)
di: Lyle, Clare, et al.
Pubblicazione: (2025)
The Illusion of Stochasticity in LLMs
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
How do language models learn facts? Dynamics, curricula and hallucinations
di: Zucchet, Nicolas, et al.
Pubblicazione: (2025)
di: Zucchet, Nicolas, et al.
Pubblicazione: (2025)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
di: Schmied, Thomas, et al.
Pubblicazione: (2025)
di: Schmied, Thomas, et al.
Pubblicazione: (2025)
Online Adaptation of Language Models with a Memory of Amortized Contexts
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
Sparse Gaussian Processes: Structured Approximations and Power-EP Revisited
di: Bui, Thang D., et al.
Pubblicazione: (2025)
di: Bui, Thang D., et al.
Pubblicazione: (2025)
New Bounds for Sparse Variational Gaussian Processes
di: Titsias, Michalis K.
Pubblicazione: (2025)
di: Titsias, Michalis K.
Pubblicazione: (2025)
L3Ms -- Lagrange Large Language Models
di: Dhillon, Guneet S., et al.
Pubblicazione: (2024)
di: Dhillon, Guneet S., et al.
Pubblicazione: (2024)
Benchmarking Diversity in Image Generation via Attribute-Conditional Human Evaluation
di: Albuquerque, Isabela, et al.
Pubblicazione: (2025)
di: Albuquerque, Isabela, et al.
Pubblicazione: (2025)
Revisiting Adam for Streaming Reinforcement Learning
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
Demystifying Diffusion Objectives: Reweighted Losses are Better Variational Bounds
di: Shi, Jiaxin, et al.
Pubblicazione: (2025)
di: Shi, Jiaxin, et al.
Pubblicazione: (2025)
From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing
di: Liu, Wei, et al.
Pubblicazione: (2026)
di: Liu, Wei, et al.
Pubblicazione: (2026)
LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories
di: Kang, Liwei, et al.
Pubblicazione: (2026)
di: Kang, Liwei, et al.
Pubblicazione: (2026)
Enhancing Large Language Model Reasoning with Reward Models: An Analytical Survey
di: Liu, Qiyuan, et al.
Pubblicazione: (2025)
di: Liu, Qiyuan, et al.
Pubblicazione: (2025)
Prompting Strategies for Enabling Large Language Models to Infer Causation from Correlation
di: Sgouritsa, Eleni, et al.
Pubblicazione: (2024)
di: Sgouritsa, Eleni, et al.
Pubblicazione: (2024)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
di: Veness, Joel, et al.
Pubblicazione: (2025)
di: Veness, Joel, et al.
Pubblicazione: (2025)
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
di: Lai, Yuhang, et al.
Pubblicazione: (2026)
di: Lai, Yuhang, et al.
Pubblicazione: (2026)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
di: Sims, Anya, et al.
Pubblicazione: (2024)
di: Sims, Anya, et al.
Pubblicazione: (2024)
Incorporating Unlabelled Data into Bayesian Neural Networks
di: Sharma, Mrinank, et al.
Pubblicazione: (2023)
di: Sharma, Mrinank, et al.
Pubblicazione: (2023)
SymDiff: Equivariant Diffusion via Stochastic Symmetrisation
di: Zhang, Leo, et al.
Pubblicazione: (2024)
di: Zhang, Leo, et al.
Pubblicazione: (2024)
Sparse Orthogonal Variational Inference for Gaussian Processes
di: Shi, Jiaxin, et al.
Pubblicazione: (2019)
di: Shi, Jiaxin, et al.
Pubblicazione: (2019)
Personalized Federated Learning with Exact Stochastic Gradient Descent
di: Nikoloutsopoulos, Sotirios, et al.
Pubblicazione: (2022)
di: Nikoloutsopoulos, Sotirios, et al.
Pubblicazione: (2022)
Variance Reduction for the Independent Metropolis Sampler
di: Liu, Siran, et al.
Pubblicazione: (2024)
di: Liu, Siran, et al.
Pubblicazione: (2024)
Manifold Aware Denoising Score Matching (MAD)
di: Levy-Jurgenson, Alona, et al.
Pubblicazione: (2026)
di: Levy-Jurgenson, Alona, et al.
Pubblicazione: (2026)
Mind the Graph When Balancing Data for Fairness or Robustness
di: Schrouff, Jessica, et al.
Pubblicazione: (2024)
di: Schrouff, Jessica, et al.
Pubblicazione: (2024)
Rao-Blackwellised Reparameterisation Gradients
di: Lam, Kevin H., et al.
Pubblicazione: (2025)
di: Lam, Kevin H., et al.
Pubblicazione: (2025)
Investigating Low-Rank Training in Transformer Language Models: Efficiency and Scaling Analysis
di: Wei, Xiuying, et al.
Pubblicazione: (2024)
di: Wei, Xiuying, et al.
Pubblicazione: (2024)
Lattice: Learning to Efficiently Compress the Memory
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
Deep Grokking: Would Deep Neural Networks Generalize Better?
di: Fan, Simin, et al.
Pubblicazione: (2024)
di: Fan, Simin, et al.
Pubblicazione: (2024)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
di: Nguyen-Hien, T. Duy, et al.
Pubblicazione: (2025)
di: Nguyen-Hien, T. Duy, et al.
Pubblicazione: (2025)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
di: Zheng, Zhi, et al.
Pubblicazione: (2025)
di: Zheng, Zhi, et al.
Pubblicazione: (2025)
Selective Safety Steering via Value-Filtered Decoding
di: Einbinder, Bat-Sheva, et al.
Pubblicazione: (2026)
di: Einbinder, Bat-Sheva, et al.
Pubblicazione: (2026)
Meta-Learning Objectives for Preference Optimization
di: Alfano, Carlo, et al.
Pubblicazione: (2024)
di: Alfano, Carlo, et al.
Pubblicazione: (2024)
EvIL: Evolution Strategies for Generalisable Imitation Learning
di: Sapora, Silvia, et al.
Pubblicazione: (2024)
di: Sapora, Silvia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Kalman Filter for Online Classification of Non-Stationary Data
di: Titsias, Michalis K., et al.
Pubblicazione: (2023) -
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
di: Galashov, Alexandre, et al.
Pubblicazione: (2024) -
Transformers for Supervised Online Continual Learning
di: Bornschein, Jorg, et al.
Pubblicazione: (2024) -
Fine-Tuned In-Context Learners for Efficient Adaptation
di: Bornschein, Jorg, et al.
Pubblicazione: (2025) -
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
di: Li, Qinyu, et al.
Pubblicazione: (2025)