Yes, Q-learning Helps Offline In-Context RL
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tarasov, Denis, Nikulin, Alexander, Zisman, Ilya, Klepach, Albina, Polubarov, Andrei, Lyubaykin, Nikita, Derevyagin, Alexander, Kiselev, Igor, Kurenkov, Vladislav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Object-Centric Latent Action Learning
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
Vision-Language Models Unlock Task-Centric Latent Actions
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026)
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
Vintix: Action Model via In-Context Reinforcement Learning
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
Latent Action Learning Requires Supervision in the Presence of Distractors
von: Nikulin, Alexander, et al.
Veröffentlicht: (2025)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2025)
N-Gram Induction Heads for In-Context RL: Improving Stability and Reducing Data Needs
von: Zisman, Ilya, et al.
Veröffentlicht: (2024)
von: Zisman, Ilya, et al.
Veröffentlicht: (2024)
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
von: Polubarov, Andrei, et al.
Veröffentlicht: (2026)
von: Polubarov, Andrei, et al.
Veröffentlicht: (2026)
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning
von: Nikulin, Alexander, et al.
Veröffentlicht: (2024)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2024)
In-Context Reinforcement Learning for Variable Action Spaces
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
Emergence of In-Context Reinforcement Learning from Noise Distillation
von: Zisman, Ilya, et al.
Veröffentlicht: (2023)
von: Zisman, Ilya, et al.
Veröffentlicht: (2023)
Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics
von: Bobrin, Maksim, et al.
Veröffentlicht: (2025)
von: Bobrin, Maksim, et al.
Veröffentlicht: (2025)
XLand-MiniGrid: Scalable Meta-Reinforcement Learning Environments in JAX
von: Nikulin, Alexander, et al.
Veröffentlicht: (2023)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2023)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
The Role of Deep Learning Regularizations on Actors in Offline RL
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
An effective control of large systems of active particles: An application to evacuation problem
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
Is Value Functions Estimation with Classification Plug-and-play for Offline Reinforcement Learning?
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
ABRA: Agent Benchmark for Radiology Applications
von: Maksudov, Bulat, et al.
Veröffentlicht: (2026)
von: Maksudov, Bulat, et al.
Veröffentlicht: (2026)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL
von: Dayal, Sarthak, et al.
Veröffentlicht: (2026)
von: Dayal, Sarthak, et al.
Veröffentlicht: (2026)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
Electrostatics from Laplacian Eigenbasis for Neural Network Interatomic Potentials
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
von: Kim, Changyeon, et al.
Veröffentlicht: (2025)
von: Kim, Changyeon, et al.
Veröffentlicht: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
Budgeting Counterfactual for Offline RL
von: Liu, Yao, et al.
Veröffentlicht: (2023)
von: Liu, Yao, et al.
Veröffentlicht: (2023)
Selective Uncertainty Propagation in Offline RL
von: Krishnamurthy, Sanath Kumar, et al.
Veröffentlicht: (2023)
von: Krishnamurthy, Sanath Kumar, et al.
Veröffentlicht: (2023)
Decoupled Prioritized Resampling for Offline RL
von: Yue, Yang, et al.
Veröffentlicht: (2023)
von: Yue, Yang, et al.
Veröffentlicht: (2023)
Augmenting Offline RL with Unlabeled Data
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
An Empirical Study on the Effectiveness of Incorporating Offline RL As Online RL Subroutines
von: Su, Jianhai, et al.
Veröffentlicht: (2025)
von: Su, Jianhai, et al.
Veröffentlicht: (2025)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
A Tractable Inference Perspective of Offline RL
von: Liu, Xuejie, et al.
Veröffentlicht: (2023)
von: Liu, Xuejie, et al.
Veröffentlicht: (2023)
Design Considerations in Offline Preference-based RL
von: Agarwal, Alekh, et al.
Veröffentlicht: (2025)
von: Agarwal, Alekh, et al.
Veröffentlicht: (2025)
Are Expressive Models Truly Necessary for Offline RL?
von: Wang, Guan, et al.
Veröffentlicht: (2024)
von: Wang, Guan, et al.
Veröffentlicht: (2024)
OGBench: Benchmarking Offline Goal-Conditioned RL
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
Distilling LLMs' Decomposition Abilities into Compact Language Models
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
Continual Learning with Columnar Spiking Neural Networks
von: Larionov, Denis, et al.
Veröffentlicht: (2025)
von: Larionov, Denis, et al.
Veröffentlicht: (2025)
Offline Multi-task Transfer RL with Representational Penalization
von: Bose, Avinandan, et al.
Veröffentlicht: (2024)
von: Bose, Avinandan, et al.
Veröffentlicht: (2024)
Is Value Learning Really the Main Bottleneck in Offline RL?
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
Scalable Offline Model-Based RL with Action Chunks
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
Pre-training with Synthetic Data Helps Offline Reinforcement Learning
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
Language-Conditioned Offline RL for Multi-Robot Navigation
von: Morad, Steven, et al.
Veröffentlicht: (2024)
von: Morad, Steven, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Object-Centric Latent Action Learning
von: Klepach, Albina, et al.
Veröffentlicht: (2025) -
Vision-Language Models Unlock Task-Centric Latent Actions
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026) -
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
von: Tarasov, Denis, et al.
Veröffentlicht: (2025) -
Vintix: Action Model via In-Context Reinforcement Learning
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025) -
Latent Action Learning Requires Supervision in the Presence of Distractors
von: Nikulin, Alexander, et al.
Veröffentlicht: (2025)