Fine-tuning Timeseries Predictors Using Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Cazaux, Hugo, Rudd, Ralph, Stefánsson, Hlynur, Ólafsson, Sverrir, Ásgeirsson, Eyjólfur Ingi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Does Synthetic Data Help? Empirical Evidence from Deep Learning Time Series Forecasters
by: Cazaux, Hugo, et al.
Published: (2026)
by: Cazaux, Hugo, et al.
Published: (2026)
On Distributional Reinforcement Learning in Chaotic Dynamical Systems
by: Rudd-Jones, James, et al.
Published: (2026)
by: Rudd-Jones, James, et al.
Published: (2026)
Operator Flow Matching for Timeseries Forecasting
by: Lee, Yolanne Yi Ran, et al.
Published: (2025)
by: Lee, Yolanne Yi Ran, et al.
Published: (2025)
Comparing Prior and Learned Time Representations in Transformer Models of Timeseries
by: Koliou, Natalia, et al.
Published: (2024)
by: Koliou, Natalia, et al.
Published: (2024)
Fine-tuning Behavioral Cloning Policies with Preference-Based Reinforcement Learning
by: Macuglia, Maël, et al.
Published: (2025)
by: Macuglia, Maël, et al.
Published: (2025)
Demystifying Design Choices of Reinforcement Fine-tuning: A Batched Contextual Bandit Learning Perspective
by: Xie, Hong, et al.
Published: (2026)
by: Xie, Hong, et al.
Published: (2026)
Step-wise Adaptive Integration of Supervised Fine-tuning and Reinforcement Learning for Task-Specific LLMs
by: Chen, Jack, et al.
Published: (2025)
by: Chen, Jack, et al.
Published: (2025)
Towards Revealing the Effectiveness of Small-Scale Fine-tuning in R1-style Reinforcement Learning
by: Chen, Yutong, et al.
Published: (2025)
by: Chen, Yutong, et al.
Published: (2025)
On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training
by: Niu, Xueyan, et al.
Published: (2026)
by: Niu, Xueyan, et al.
Published: (2026)
REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning
by: Gu, Zhaoyuan, et al.
Published: (2026)
by: Gu, Zhaoyuan, et al.
Published: (2026)
Generating Expressive and Customizable Evals for Timeseries Data Analysis Agents with AgentFuel
by: Maddi, Aadyaa, et al.
Published: (2026)
by: Maddi, Aadyaa, et al.
Published: (2026)
New Insights on Unfolding and Fine-tuning Quantum Federated Learning
by: Nanayakkara, Shanika Iroshi, et al.
Published: (2025)
by: Nanayakkara, Shanika Iroshi, et al.
Published: (2025)
Mechanistic Fine-tuning for In-context Learning
by: Cho, Hakaze, et al.
Published: (2025)
by: Cho, Hakaze, et al.
Published: (2025)
Uncovering Memorization in Timeseries Imputation models: LBRM Membership Inference and its link to attribute Leakage
by: Taleb, Faiz, et al.
Published: (2026)
by: Taleb, Faiz, et al.
Published: (2026)
Activated LoRA: Fine-tuned LLMs for Intrinsics
by: Greenewald, Kristjan, et al.
Published: (2025)
by: Greenewald, Kristjan, et al.
Published: (2025)
Rethinking Safety in LLM Fine-tuning: An Optimization Perspective
by: Kim, Minseon, et al.
Published: (2025)
by: Kim, Minseon, et al.
Published: (2025)
Aggregating Low Rank Adapters in Federated Fine-tuning
by: Trautmann, Evelyn, et al.
Published: (2025)
by: Trautmann, Evelyn, et al.
Published: (2025)
Speculative Coreset Selection for Task-Specific Fine-tuning
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Compiler-R1: Towards Agentic Compiler Auto-tuning with Reinforcement Learning
by: Pan, Haolin, et al.
Published: (2025)
by: Pan, Haolin, et al.
Published: (2025)
A Practical Guide for Evaluating LLMs and LLM-Reliant Systems
by: Rudd, Ethan M., et al.
Published: (2025)
by: Rudd, Ethan M., et al.
Published: (2025)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
by: Kim, Hyeonjun, et al.
Published: (2025)
by: Kim, Hyeonjun, et al.
Published: (2025)
Fine-tuning Large Language Models for Entity Matching
by: Steiner, Aaron, et al.
Published: (2024)
by: Steiner, Aaron, et al.
Published: (2024)
Fine-tuning Diffusion Policies with Backpropagation Through Diffusion Timesteps
by: Yang, Ningyuan, et al.
Published: (2025)
by: Yang, Ningyuan, et al.
Published: (2025)
Fine-tuning Large Language Model for Automated Algorithm Design
by: Liu, Fei, et al.
Published: (2025)
by: Liu, Fei, et al.
Published: (2025)
Fine-tuning Flow Matching Generative Models with Intermediate Feedback
by: Fan, Jiajun, et al.
Published: (2025)
by: Fan, Jiajun, et al.
Published: (2025)
CF-VLM:CounterFactual Vision-Language Fine-tuning
by: Zhang, Jusheng, et al.
Published: (2025)
by: Zhang, Jusheng, et al.
Published: (2025)
Faster Convergence for Transformer Fine-tuning with Line Search Methods
by: Kenneweg, Philip, et al.
Published: (2024)
by: Kenneweg, Philip, et al.
Published: (2024)
Rethinking Reinforcement fine-tuning of LLMs: A Multi-armed Bandit Learning Perspective
by: Hu, Xiao, et al.
Published: (2026)
by: Hu, Xiao, et al.
Published: (2026)
Enhancing Reinforcement Learning Fine-Tuning with an Online Refiner
by: Ma, Hao, et al.
Published: (2026)
by: Ma, Hao, et al.
Published: (2026)
Membership Inference Attacks Against Fine-tuned Diffusion Language Models
by: Chen, Yuetian, et al.
Published: (2026)
by: Chen, Yuetian, et al.
Published: (2026)
Ratio-Variance Regularized Policy Optimization for Efficient LLM Fine-tuning
by: Luo, Yu, et al.
Published: (2026)
by: Luo, Yu, et al.
Published: (2026)
Safeguarding LLM Fine-tuning via Push-Pull Distributional Alignment
by: Wang, Haozhong, et al.
Published: (2026)
by: Wang, Haozhong, et al.
Published: (2026)
The Blessing of Dimensionality in LLM Fine-tuning: A Variance-Curvature Perspective
by: Liang, Qiyao, et al.
Published: (2026)
by: Liang, Qiyao, et al.
Published: (2026)
A Minimalist Method for Fine-tuning Text-to-Image Diffusion Models
by: Miao, Yanting, et al.
Published: (2025)
by: Miao, Yanting, et al.
Published: (2025)
Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models
by: Kim, Kyuyoung, et al.
Published: (2024)
by: Kim, Kyuyoung, et al.
Published: (2024)
TuneComp: Joint Fine-tuning and Compression for Large Foundation Models
by: Chen, Xiangyu, et al.
Published: (2025)
by: Chen, Xiangyu, et al.
Published: (2025)
Adaptive Divergence Regularized Policy Optimization for Fine-tuning Generative Models
by: Fan, Jiajun, et al.
Published: (2025)
by: Fan, Jiajun, et al.
Published: (2025)
Neural ODE Transformers: Analyzing Internal Dynamics and Adaptive Fine-tuning
by: Tong, Anh, et al.
Published: (2025)
by: Tong, Anh, et al.
Published: (2025)
Learning Universal Predictors
by: Grau-Moya, Jordi, et al.
Published: (2024)
by: Grau-Moya, Jordi, et al.
Published: (2024)
Similar Items
-
Does Synthetic Data Help? Empirical Evidence from Deep Learning Time Series Forecasters
by: Cazaux, Hugo, et al.
Published: (2026) -
On Distributional Reinforcement Learning in Chaotic Dynamical Systems
by: Rudd-Jones, James, et al.
Published: (2026) -
Operator Flow Matching for Timeseries Forecasting
by: Lee, Yolanne Yi Ran, et al.
Published: (2025) -
Comparing Prior and Learned Time Representations in Transformer Models of Timeseries
by: Koliou, Natalia, et al.
Published: (2024) -
Fine-tuning Behavioral Cloning Policies with Preference-Based Reinforcement Learning
by: Macuglia, Maël, et al.
Published: (2025)