Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Polubarov, Andrei, Nikita, Lyubaykin, Derevyagin, Alexander, Grishin, Artyom, Saprygin, Igor, Serkov, Aleksandr, Averchenko, Mark, Tikhonov, Daniil, Zhdanov, Maksim, Nikulin, Alexander, Zisman, Ilya, Klepach, Albina, Zemtsov, Alexey, Kurenkov, Vladislav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vintix: Action Model via In-Context Reinforcement Learning
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
Vision-Language Models Unlock Task-Centric Latent Actions
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026)
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
Yes, Q-learning Helps Offline In-Context RL
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
Object-Centric Latent Action Learning
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning
von: Nikulin, Alexander, et al.
Veröffentlicht: (2024)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2024)
Latent Action Learning Requires Supervision in the Presence of Distractors
von: Nikulin, Alexander, et al.
Veröffentlicht: (2025)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2025)
N-Gram Induction Heads for In-Context RL: Improving Stability and Reducing Data Needs
von: Zisman, Ilya, et al.
Veröffentlicht: (2024)
von: Zisman, Ilya, et al.
Veröffentlicht: (2024)
Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics
von: Bobrin, Maksim, et al.
Veröffentlicht: (2025)
von: Bobrin, Maksim, et al.
Veröffentlicht: (2025)
In-Context Reinforcement Learning for Variable Action Spaces
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
Emergence of In-Context Reinforcement Learning from Noise Distillation
von: Zisman, Ilya, et al.
Veröffentlicht: (2023)
von: Zisman, Ilya, et al.
Veröffentlicht: (2023)
Electrostatics from Laplacian Eigenbasis for Neural Network Interatomic Potentials
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
XLand-MiniGrid: Scalable Meta-Reinforcement Learning Environments in JAX
von: Nikulin, Alexander, et al.
Veröffentlicht: (2023)
von: Nikulin, Alexander, et al.
Veröffentlicht: (2023)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
From bedside to genetic analysis: New insights into pathophysiology of melanoma, basal cell carcinoma, and other cancers
von: Alexander Zemtsov
Veröffentlicht: (2024)
von: Alexander Zemtsov
Veröffentlicht: (2024)
Effect of group-velocity dispersion on the generation of multimode pulsed squeezed light in a synchronously pumped optical parametric oscillator
von: Averchenko, Valentin, et al.
Veröffentlicht: (2024)
von: Averchenko, Valentin, et al.
Veröffentlicht: (2024)
Limits of Perturbation Theory for Multimode Light Propagation in Dispersive Optical Cavities
von: Tikhonov, K. S., et al.
Veröffentlicht: (2025)
von: Tikhonov, K. S., et al.
Veröffentlicht: (2025)
Conditional Clifford-Steerable CNNs with Complete Kernel Basis for PDE Modeling
von: Szarvas, Bálint László, et al.
Veröffentlicht: (2025)
von: Szarvas, Bálint László, et al.
Veröffentlicht: (2025)
Weak-to-Strong 3D Object Detection with X-Ray Distillation
von: Gambashidze, Alexander, et al.
Veröffentlicht: (2024)
von: Gambashidze, Alexander, et al.
Veröffentlicht: (2024)
An effective control of large systems of active particles: An application to evacuation problem
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
von: Klepach, Albina, et al.
Veröffentlicht: (2025)
Identity Curvature Laplace Approximation for Improved Out-of-Distribution Detection
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2023)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2023)
One-Point Feedback for Composite Optimization with Applications to Distributed and Federated Learning
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2021)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2021)
Improved Probabilistic Lower Bounds for Separable Matrices
von: Goshkoder, Daniil, et al.
Veröffentlicht: (2024)
von: Goshkoder, Daniil, et al.
Veröffentlicht: (2024)
Optimal Analysis of Method with Batching for Monotone Stochastic Finite-Sum Variational Inequalities
von: Pichugin, Alexander, et al.
Veröffentlicht: (2024)
von: Pichugin, Alexander, et al.
Veröffentlicht: (2024)
Multifunctional 2d infrared photodetectors enabled by asymmetric singular metasurfaces
von: Semkin, Valentin, et al.
Veröffentlicht: (2024)
von: Semkin, Valentin, et al.
Veröffentlicht: (2024)
ECLAIR: A High-Fidelity Aerial LiDAR Dataset for Semantic Segmentation
von: Melekhov, Iaroslav, et al.
Veröffentlicht: (2024)
von: Melekhov, Iaroslav, et al.
Veröffentlicht: (2024)
From Images to Decisions: Assistive Computer Vision for Non-Metallic Content Estimation in Scrap Metal
von: Storonkin, Daniil, et al.
Veröffentlicht: (2026)
von: Storonkin, Daniil, et al.
Veröffentlicht: (2026)
Motives, cohomological invariants and Freudenthal magic square
von: Geldhauser, Nikita, et al.
Veröffentlicht: (2026)
von: Geldhauser, Nikita, et al.
Veröffentlicht: (2026)
AdS-GNN -- a Conformally Equivariant Graph Neural Network
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
Diversifying Deep Ensembles: A Saliency Map Approach for Enhanced OOD Detection, Calibration, and Accuracy
von: Dereka, Stanislav, et al.
Veröffentlicht: (2023)
von: Dereka, Stanislav, et al.
Veröffentlicht: (2023)
Method with Batching for Stochastic Finite-Sum Variational Inequalities in Non-Euclidean Setting
von: Pichugin, Alexander, et al.
Veröffentlicht: (2024)
von: Pichugin, Alexander, et al.
Veröffentlicht: (2024)
Optimal Data Splitting in Distributed Optimization for Machine Learning
von: Medyakov, Daniil, et al.
Veröffentlicht: (2024)
von: Medyakov, Daniil, et al.
Veröffentlicht: (2024)
ABRA: Agent Benchmark for Radiology Applications
von: Maksudov, Bulat, et al.
Veröffentlicht: (2026)
von: Maksudov, Bulat, et al.
Veröffentlicht: (2026)
The sl(2)-weight system at c = 3/8 for graphs
von: Fomichev, Daniil, et al.
Veröffentlicht: (2024)
von: Fomichev, Daniil, et al.
Veröffentlicht: (2024)
Erwin: A Tree-based Hierarchical Transformer for Large-scale Physical Systems
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2025)
Synthesis of Chiral Dispiro‐Indolinone‐Pyrrolidine‐Imidazolones Using ( R ) or ( S )‐1‐(2,4‐Dimethoxyphenyl)Ethyl Amine
von: Vladislav S. Polyakov, et al.
Veröffentlicht: (2025)
von: Vladislav S. Polyakov, et al.
Veröffentlicht: (2025)
AIpom at SemEval-2024 Task 8: Detecting AI-produced Outputs in M4
von: Shirnin, Alexander, et al.
Veröffentlicht: (2024)
von: Shirnin, Alexander, et al.
Veröffentlicht: (2024)
Papilusion at DAGPap24: Paper or Illusion? Detecting AI-generated Scientific Papers
von: Andreev, Nikita, et al.
Veröffentlicht: (2024)
von: Andreev, Nikita, et al.
Veröffentlicht: (2024)
Evolution of perturbations in the model of Tsallis holographic dark energy
von: Astashenok, Artyom V., et al.
Veröffentlicht: (2024)
von: Astashenok, Artyom V., et al.
Veröffentlicht: (2024)
Non-flat Universe with Tsallis holographic dark energy
von: Astashenok, Artyom V., et al.
Veröffentlicht: (2025)
von: Astashenok, Artyom V., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Vintix: Action Model via In-Context Reinforcement Learning
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025) -
Vision-Language Models Unlock Task-Centric Latent Actions
von: Nikulin, Alexander, et al.
Veröffentlicht: (2026) -
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
von: Tarasov, Denis, et al.
Veröffentlicht: (2025) -
Yes, Q-learning Helps Offline In-Context RL
von: Tarasov, Denis, et al.
Veröffentlicht: (2025) -
Object-Centric Latent Action Learning
von: Klepach, Albina, et al.
Veröffentlicht: (2025)