QT-TDM: Planning With Transformer Dynamics Model and Autoregressive Q-Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kotb, Mostafa, Weber, Cornelius, Hafez, Muhammad Burhan, Wermter, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
von: Hafez, Muhammad Burhan, et al.
Veröffentlicht: (2024)
von: Hafez, Muhammad Burhan, et al.
Veröffentlicht: (2024)
Agentic Skill Discovery
von: Zhao, Xufeng, et al.
Veröffentlicht: (2024)
von: Zhao, Xufeng, et al.
Veröffentlicht: (2024)
Survey on reinforcement learning for language processing
von: Uc-Cetina, Victor, et al.
Veröffentlicht: (2021)
von: Uc-Cetina, Victor, et al.
Veröffentlicht: (2021)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
von: Zhao, Xufeng, et al.
Veröffentlicht: (2023)
von: Zhao, Xufeng, et al.
Veröffentlicht: (2023)
LLM+MAP: Bimanual Robot Task Planning using Large Language Models and Planning Domain Definition Language
von: Chu, Kun, et al.
Veröffentlicht: (2025)
von: Chu, Kun, et al.
Veröffentlicht: (2025)
HopCast: Calibration of Autoregressive Dynamics Models
von: Shahid, Muhammad Bilal, et al.
Veröffentlicht: (2025)
von: Shahid, Muhammad Bilal, et al.
Veröffentlicht: (2025)
Robotic Imitation of Human Actions
von: Spisak, Josua, et al.
Veröffentlicht: (2024)
von: Spisak, Josua, et al.
Veröffentlicht: (2024)
Large Language Model Data Generation for Enhanced Intent Recognition in German Speech
von: Rosin, Theresa Pekarek, et al.
Veröffentlicht: (2025)
von: Rosin, Theresa Pekarek, et al.
Veröffentlicht: (2025)
Mental Modeling of Reinforcement Learning Agents by Language Models
von: Lu, Wenhao, et al.
Veröffentlicht: (2024)
von: Lu, Wenhao, et al.
Veröffentlicht: (2024)
Read Between the Layers: Leveraging Multi-Layer Representations for Rehearsal-Free Continual Learning with Pre-Trained Models
von: Ahrens, Kyra, et al.
Veröffentlicht: (2023)
von: Ahrens, Kyra, et al.
Veröffentlicht: (2023)
Efficient Autoregressive Inference for Transformer Probabilistic Models
von: Hassan, Conor, et al.
Veröffentlicht: (2025)
von: Hassan, Conor, et al.
Veröffentlicht: (2025)
The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level
von: Herbst, Jeremy, et al.
Veröffentlicht: (2026)
von: Herbst, Jeremy, et al.
Veröffentlicht: (2026)
Global Context Enhanced Anomaly Detection of Cyber Attacks via Decoupled Graph Neural Networks
von: Hafez, Ahmad
Veröffentlicht: (2024)
von: Hafez, Ahmad
Veröffentlicht: (2024)
How do Transformers perform In-Context Autoregressive Learning?
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
LLM-based Interactive Imitation Learning for Robotic Manipulation
von: Werner, Jonas, et al.
Veröffentlicht: (2025)
von: Werner, Jonas, et al.
Veröffentlicht: (2025)
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers
von: Anagnostidis, Sotiris, et al.
Veröffentlicht: (2023)
von: Anagnostidis, Sotiris, et al.
Veröffentlicht: (2023)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
Snapture -- A Novel Neural Architecture for Combined Static and Dynamic Hand Gesture Recognition
von: Ali, Hassan, et al.
Veröffentlicht: (2022)
von: Ali, Hassan, et al.
Veröffentlicht: (2022)
Transformer Neural Autoregressive Flows
von: Patacchiola, Massimiliano, et al.
Veröffentlicht: (2024)
von: Patacchiola, Massimiliano, et al.
Veröffentlicht: (2024)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
von: Sharify, Sayeh, et al.
Veröffentlicht: (2026)
von: Sharify, Sayeh, et al.
Veröffentlicht: (2026)
Linear Transformers as VAR Models: Aligning Autoregressive Attention Mechanisms with Autoregressive Forecasting
von: Lu, Jiecheng, et al.
Veröffentlicht: (2025)
von: Lu, Jiecheng, et al.
Veröffentlicht: (2025)
FabuLight-ASD: Unveiling Speech Activity via Body Language
von: Carneiro, Hugo, et al.
Veröffentlicht: (2024)
von: Carneiro, Hugo, et al.
Veröffentlicht: (2024)
Q-value Regularized Transformer for Offline Reinforcement Learning
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
Quantization-Free Autoregressive Action Transformer
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2025)
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2025)
The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
Bias Leaves a Gradient Trail: Label-Free Bias Identification via Gradient Probes on Concept Decompositions
von: Vitry, Thomas, et al.
Veröffentlicht: (2026)
von: Vitry, Thomas, et al.
Veröffentlicht: (2026)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
von: Reid, Cameron, et al.
Veröffentlicht: (2025)
von: Reid, Cameron, et al.
Veröffentlicht: (2025)
Masked Diffusion Models are Secretly Learned-Order Autoregressive Models
von: Garg, Prateek, et al.
Veröffentlicht: (2025)
von: Garg, Prateek, et al.
Veröffentlicht: (2025)
On-chain Validation of Tracking Data Messages (TDM) Using Distributed Deep Learning on a Proof of Stake (PoS) Blockchain
von: Latif, Yasir, et al.
Veröffentlicht: (2024)
von: Latif, Yasir, et al.
Veröffentlicht: (2024)
A Parameter-free Adaptive Resonance Theory-based Topological Clustering Algorithm Capable of Continual Learning
von: Masuyama, Naoki, et al.
Veröffentlicht: (2023)
von: Masuyama, Naoki, et al.
Veröffentlicht: (2023)
PHIDA: Persistence-Guided Node-to-Cluster Mapping for Online Clustering
von: Masuyama, Naoki, et al.
Veröffentlicht: (2026)
von: Masuyama, Naoki, et al.
Veröffentlicht: (2026)
Dynamic Estimation of Learning Rates Using a Non-Linear Autoregressive Model
von: Okhrati, Ramin
Veröffentlicht: (2024)
von: Okhrati, Ramin
Veröffentlicht: (2024)
Causal State Distillation for Explainable Reinforcement Learning
von: Lu, Wenhao, et al.
Veröffentlicht: (2023)
von: Lu, Wenhao, et al.
Veröffentlicht: (2023)
Language Model-Based Paired Variational Autoencoders for Robotic Language Learning
von: Özdemir, Ozan, et al.
Veröffentlicht: (2022)
von: Özdemir, Ozan, et al.
Veröffentlicht: (2022)
Exploring Design Choices for Autoregressive Deep Learning Climate Models
von: Gallusser, Florian, et al.
Veröffentlicht: (2025)
von: Gallusser, Florian, et al.
Veröffentlicht: (2025)
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2025)
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2025)
Replication Study: Enhancing Hydrological Modeling with Physics-Guided Machine Learning
von: Esmaeilzadeh, Mostafa, et al.
Veröffentlicht: (2024)
von: Esmaeilzadeh, Mostafa, et al.
Veröffentlicht: (2024)
SeriesGAN: Time Series Generation via Adversarial and Autoregressive Learning
von: EskandariNasab, MohammadReza, et al.
Veröffentlicht: (2024)
von: EskandariNasab, MohammadReza, et al.
Veröffentlicht: (2024)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
von: Hafez, Muhammad Burhan, et al.
Veröffentlicht: (2024) -
Agentic Skill Discovery
von: Zhao, Xufeng, et al.
Veröffentlicht: (2024) -
Survey on reinforcement learning for language processing
von: Uc-Cetina, Victor, et al.
Veröffentlicht: (2021) -
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025) -
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
von: Zhao, Xufeng, et al.
Veröffentlicht: (2023)