QT-TDM: Planning With Transformer Dynamics Model and Autoregressive Q-Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Kotb, Mostafa, Weber, Cornelius, Hafez, Muhammad Burhan, Wermter, Stefan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
par: Hafez, Muhammad Burhan, et autres
Publié: (2024)
par: Hafez, Muhammad Burhan, et autres
Publié: (2024)
Agentic Skill Discovery
par: Zhao, Xufeng, et autres
Publié: (2024)
par: Zhao, Xufeng, et autres
Publié: (2024)
Survey on reinforcement learning for language processing
par: Uc-Cetina, Victor, et autres
Publié: (2021)
par: Uc-Cetina, Victor, et autres
Publié: (2021)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
par: Ghaemi, Hafez, et autres
Publié: (2025)
par: Ghaemi, Hafez, et autres
Publié: (2025)
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
par: Zhao, Xufeng, et autres
Publié: (2023)
par: Zhao, Xufeng, et autres
Publié: (2023)
LLM+MAP: Bimanual Robot Task Planning using Large Language Models and Planning Domain Definition Language
par: Chu, Kun, et autres
Publié: (2025)
par: Chu, Kun, et autres
Publié: (2025)
HopCast: Calibration of Autoregressive Dynamics Models
par: Shahid, Muhammad Bilal, et autres
Publié: (2025)
par: Shahid, Muhammad Bilal, et autres
Publié: (2025)
Robotic Imitation of Human Actions
par: Spisak, Josua, et autres
Publié: (2024)
par: Spisak, Josua, et autres
Publié: (2024)
Large Language Model Data Generation for Enhanced Intent Recognition in German Speech
par: Rosin, Theresa Pekarek, et autres
Publié: (2025)
par: Rosin, Theresa Pekarek, et autres
Publié: (2025)
Mental Modeling of Reinforcement Learning Agents by Language Models
par: Lu, Wenhao, et autres
Publié: (2024)
par: Lu, Wenhao, et autres
Publié: (2024)
Read Between the Layers: Leveraging Multi-Layer Representations for Rehearsal-Free Continual Learning with Pre-Trained Models
par: Ahrens, Kyra, et autres
Publié: (2023)
par: Ahrens, Kyra, et autres
Publié: (2023)
Efficient Autoregressive Inference for Transformer Probabilistic Models
par: Hassan, Conor, et autres
Publié: (2025)
par: Hassan, Conor, et autres
Publié: (2025)
The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level
par: Herbst, Jeremy, et autres
Publié: (2026)
par: Herbst, Jeremy, et autres
Publié: (2026)
Global Context Enhanced Anomaly Detection of Cyber Attacks via Decoupled Graph Neural Networks
par: Hafez, Ahmad
Publié: (2024)
par: Hafez, Ahmad
Publié: (2024)
How do Transformers perform In-Context Autoregressive Learning?
par: Sander, Michael E., et autres
Publié: (2024)
par: Sander, Michael E., et autres
Publié: (2024)
LLM-based Interactive Imitation Learning for Robotic Manipulation
par: Werner, Jonas, et autres
Publié: (2025)
par: Werner, Jonas, et autres
Publié: (2025)
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers
par: Anagnostidis, Sotiris, et autres
Publié: (2023)
par: Anagnostidis, Sotiris, et autres
Publié: (2023)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
par: Wang, Siwei, et autres
Publié: (2024)
par: Wang, Siwei, et autres
Publié: (2024)
Snapture -- A Novel Neural Architecture for Combined Static and Dynamic Hand Gesture Recognition
par: Ali, Hassan, et autres
Publié: (2022)
par: Ali, Hassan, et autres
Publié: (2022)
Transformer Neural Autoregressive Flows
par: Patacchiola, Massimiliano, et autres
Publié: (2024)
par: Patacchiola, Massimiliano, et autres
Publié: (2024)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
par: Sharify, Sayeh, et autres
Publié: (2026)
par: Sharify, Sayeh, et autres
Publié: (2026)
Linear Transformers as VAR Models: Aligning Autoregressive Attention Mechanisms with Autoregressive Forecasting
par: Lu, Jiecheng, et autres
Publié: (2025)
par: Lu, Jiecheng, et autres
Publié: (2025)
FabuLight-ASD: Unveiling Speech Activity via Body Language
par: Carneiro, Hugo, et autres
Publié: (2024)
par: Carneiro, Hugo, et autres
Publié: (2024)
Q-value Regularized Transformer for Offline Reinforcement Learning
par: Hu, Shengchao, et autres
Publié: (2024)
par: Hu, Shengchao, et autres
Publié: (2024)
Quantization-Free Autoregressive Action Transformer
par: Sheebaelhamd, Ziyad, et autres
Publié: (2025)
par: Sheebaelhamd, Ziyad, et autres
Publié: (2025)
The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning
par: Hafez, Wael, et autres
Publié: (2026)
par: Hafez, Wael, et autres
Publié: (2026)
Bias Leaves a Gradient Trail: Label-Free Bias Identification via Gradient Probes on Concept Decompositions
par: Vitry, Thomas, et autres
Publié: (2026)
par: Vitry, Thomas, et autres
Publié: (2026)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
par: Reid, Cameron, et autres
Publié: (2025)
par: Reid, Cameron, et autres
Publié: (2025)
Masked Diffusion Models are Secretly Learned-Order Autoregressive Models
par: Garg, Prateek, et autres
Publié: (2025)
par: Garg, Prateek, et autres
Publié: (2025)
On-chain Validation of Tracking Data Messages (TDM) Using Distributed Deep Learning on a Proof of Stake (PoS) Blockchain
par: Latif, Yasir, et autres
Publié: (2024)
par: Latif, Yasir, et autres
Publié: (2024)
A Parameter-free Adaptive Resonance Theory-based Topological Clustering Algorithm Capable of Continual Learning
par: Masuyama, Naoki, et autres
Publié: (2023)
par: Masuyama, Naoki, et autres
Publié: (2023)
PHIDA: Persistence-Guided Node-to-Cluster Mapping for Online Clustering
par: Masuyama, Naoki, et autres
Publié: (2026)
par: Masuyama, Naoki, et autres
Publié: (2026)
Dynamic Estimation of Learning Rates Using a Non-Linear Autoregressive Model
par: Okhrati, Ramin
Publié: (2024)
par: Okhrati, Ramin
Publié: (2024)
Causal State Distillation for Explainable Reinforcement Learning
par: Lu, Wenhao, et autres
Publié: (2023)
par: Lu, Wenhao, et autres
Publié: (2023)
Language Model-Based Paired Variational Autoencoders for Robotic Language Learning
par: Özdemir, Ozan, et autres
Publié: (2022)
par: Özdemir, Ozan, et autres
Publié: (2022)
Exploring Design Choices for Autoregressive Deep Learning Climate Models
par: Gallusser, Florian, et autres
Publié: (2025)
par: Gallusser, Florian, et autres
Publié: (2025)
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
par: Zhang, Ruixiang, et autres
Publié: (2025)
par: Zhang, Ruixiang, et autres
Publié: (2025)
Replication Study: Enhancing Hydrological Modeling with Physics-Guided Machine Learning
par: Esmaeilzadeh, Mostafa, et autres
Publié: (2024)
par: Esmaeilzadeh, Mostafa, et autres
Publié: (2024)
SeriesGAN: Time Series Generation via Adversarial and Autoregressive Learning
par: EskandariNasab, MohammadReza, et autres
Publié: (2024)
par: EskandariNasab, MohammadReza, et autres
Publié: (2024)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
par: Ye, Jiacheng, et autres
Publié: (2024)
par: Ye, Jiacheng, et autres
Publié: (2024)
Documents similaires
-
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
par: Hafez, Muhammad Burhan, et autres
Publié: (2024) -
Agentic Skill Discovery
par: Zhao, Xufeng, et autres
Publié: (2024) -
Survey on reinforcement learning for language processing
par: Uc-Cetina, Victor, et autres
Publié: (2021) -
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
par: Ghaemi, Hafez, et autres
Publié: (2025) -
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
par: Zhao, Xufeng, et autres
Publié: (2023)