QT-TDM: Planning With Transformer Dynamics Model and Autoregressive Q-Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kotb, Mostafa, Weber, Cornelius, Hafez, Muhammad Burhan, Wermter, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024)
by: Zhao, Xufeng, et al.
Published: (2024)
Survey on reinforcement learning for language processing
by: Uc-Cetina, Victor, et al.
Published: (2021)
by: Uc-Cetina, Victor, et al.
Published: (2021)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
by: Ghaemi, Hafez, et al.
Published: (2025)
by: Ghaemi, Hafez, et al.
Published: (2025)
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
by: Zhao, Xufeng, et al.
Published: (2023)
by: Zhao, Xufeng, et al.
Published: (2023)
LLM+MAP: Bimanual Robot Task Planning using Large Language Models and Planning Domain Definition Language
by: Chu, Kun, et al.
Published: (2025)
by: Chu, Kun, et al.
Published: (2025)
HopCast: Calibration of Autoregressive Dynamics Models
by: Shahid, Muhammad Bilal, et al.
Published: (2025)
by: Shahid, Muhammad Bilal, et al.
Published: (2025)
Robotic Imitation of Human Actions
by: Spisak, Josua, et al.
Published: (2024)
by: Spisak, Josua, et al.
Published: (2024)
Large Language Model Data Generation for Enhanced Intent Recognition in German Speech
by: Rosin, Theresa Pekarek, et al.
Published: (2025)
by: Rosin, Theresa Pekarek, et al.
Published: (2025)
Mental Modeling of Reinforcement Learning Agents by Language Models
by: Lu, Wenhao, et al.
Published: (2024)
by: Lu, Wenhao, et al.
Published: (2024)
Read Between the Layers: Leveraging Multi-Layer Representations for Rehearsal-Free Continual Learning with Pre-Trained Models
by: Ahrens, Kyra, et al.
Published: (2023)
by: Ahrens, Kyra, et al.
Published: (2023)
Efficient Autoregressive Inference for Transformer Probabilistic Models
by: Hassan, Conor, et al.
Published: (2025)
by: Hassan, Conor, et al.
Published: (2025)
The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level
by: Herbst, Jeremy, et al.
Published: (2026)
by: Herbst, Jeremy, et al.
Published: (2026)
Global Context Enhanced Anomaly Detection of Cyber Attacks via Decoupled Graph Neural Networks
by: Hafez, Ahmad
Published: (2024)
by: Hafez, Ahmad
Published: (2024)
How do Transformers perform In-Context Autoregressive Learning?
by: Sander, Michael E., et al.
Published: (2024)
by: Sander, Michael E., et al.
Published: (2024)
LLM-based Interactive Imitation Learning for Robotic Manipulation
by: Werner, Jonas, et al.
Published: (2025)
by: Werner, Jonas, et al.
Published: (2025)
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers
by: Anagnostidis, Sotiris, et al.
Published: (2023)
by: Anagnostidis, Sotiris, et al.
Published: (2023)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
by: Wang, Siwei, et al.
Published: (2024)
by: Wang, Siwei, et al.
Published: (2024)
Snapture -- A Novel Neural Architecture for Combined Static and Dynamic Hand Gesture Recognition
by: Ali, Hassan, et al.
Published: (2022)
by: Ali, Hassan, et al.
Published: (2022)
Transformer Neural Autoregressive Flows
by: Patacchiola, Massimiliano, et al.
Published: (2024)
by: Patacchiola, Massimiliano, et al.
Published: (2024)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
by: Sharify, Sayeh, et al.
Published: (2026)
by: Sharify, Sayeh, et al.
Published: (2026)
Linear Transformers as VAR Models: Aligning Autoregressive Attention Mechanisms with Autoregressive Forecasting
by: Lu, Jiecheng, et al.
Published: (2025)
by: Lu, Jiecheng, et al.
Published: (2025)
FabuLight-ASD: Unveiling Speech Activity via Body Language
by: Carneiro, Hugo, et al.
Published: (2024)
by: Carneiro, Hugo, et al.
Published: (2024)
Q-value Regularized Transformer for Offline Reinforcement Learning
by: Hu, Shengchao, et al.
Published: (2024)
by: Hu, Shengchao, et al.
Published: (2024)
Quantization-Free Autoregressive Action Transformer
by: Sheebaelhamd, Ziyad, et al.
Published: (2025)
by: Sheebaelhamd, Ziyad, et al.
Published: (2025)
The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning
by: Hafez, Wael, et al.
Published: (2026)
by: Hafez, Wael, et al.
Published: (2026)
Bias Leaves a Gradient Trail: Label-Free Bias Identification via Gradient Probes on Concept Decompositions
by: Vitry, Thomas, et al.
Published: (2026)
by: Vitry, Thomas, et al.
Published: (2026)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
by: Reid, Cameron, et al.
Published: (2025)
by: Reid, Cameron, et al.
Published: (2025)
Masked Diffusion Models are Secretly Learned-Order Autoregressive Models
by: Garg, Prateek, et al.
Published: (2025)
by: Garg, Prateek, et al.
Published: (2025)
On-chain Validation of Tracking Data Messages (TDM) Using Distributed Deep Learning on a Proof of Stake (PoS) Blockchain
by: Latif, Yasir, et al.
Published: (2024)
by: Latif, Yasir, et al.
Published: (2024)
A Parameter-free Adaptive Resonance Theory-based Topological Clustering Algorithm Capable of Continual Learning
by: Masuyama, Naoki, et al.
Published: (2023)
by: Masuyama, Naoki, et al.
Published: (2023)
PHIDA: Persistence-Guided Node-to-Cluster Mapping for Online Clustering
by: Masuyama, Naoki, et al.
Published: (2026)
by: Masuyama, Naoki, et al.
Published: (2026)
Dynamic Estimation of Learning Rates Using a Non-Linear Autoregressive Model
by: Okhrati, Ramin
Published: (2024)
by: Okhrati, Ramin
Published: (2024)
Causal State Distillation for Explainable Reinforcement Learning
by: Lu, Wenhao, et al.
Published: (2023)
by: Lu, Wenhao, et al.
Published: (2023)
Language Model-Based Paired Variational Autoencoders for Robotic Language Learning
by: Özdemir, Ozan, et al.
Published: (2022)
by: Özdemir, Ozan, et al.
Published: (2022)
Exploring Design Choices for Autoregressive Deep Learning Climate Models
by: Gallusser, Florian, et al.
Published: (2025)
by: Gallusser, Florian, et al.
Published: (2025)
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
by: Zhang, Ruixiang, et al.
Published: (2025)
by: Zhang, Ruixiang, et al.
Published: (2025)
Replication Study: Enhancing Hydrological Modeling with Physics-Guided Machine Learning
by: Esmaeilzadeh, Mostafa, et al.
Published: (2024)
by: Esmaeilzadeh, Mostafa, et al.
Published: (2024)
SeriesGAN: Time Series Generation via Adversarial and Autoregressive Learning
by: EskandariNasab, MohammadReza, et al.
Published: (2024)
by: EskandariNasab, MohammadReza, et al.
Published: (2024)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
by: Ye, Jiacheng, et al.
Published: (2024)
by: Ye, Jiacheng, et al.
Published: (2024)
Similar Items
-
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
by: Hafez, Muhammad Burhan, et al.
Published: (2024) -
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024) -
Survey on reinforcement learning for language processing
by: Uc-Cetina, Victor, et al.
Published: (2021) -
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
by: Ghaemi, Hafez, et al.
Published: (2025) -
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
by: Zhao, Xufeng, et al.
Published: (2023)