Hadamax Encoding: Elevating Performance in Model-Free Atari
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kooi, Jacob E., Yang, Zhao, François-Lavet, Vincent |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangled (Un)Controllable Features
von: Kooi, Jacob E., et al.
Veröffentlicht: (2022)
von: Kooi, Jacob E., et al.
Veröffentlicht: (2022)
Novel RL approach for efficient Elevator Group Control Systems
von: Vaartjes, Nathan, et al.
Veröffentlicht: (2025)
von: Vaartjes, Nathan, et al.
Veröffentlicht: (2025)
Hadamard Representation: Scaffolding Performance Across Model-free RL
von: Kooi, Jacob E., et al.
Veröffentlicht: (2024)
von: Kooi, Jacob E., et al.
Veröffentlicht: (2024)
Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability
von: Kim, Taewoon, et al.
Veröffentlicht: (2026)
von: Kim, Taewoon, et al.
Veröffentlicht: (2026)
Temporal Knowledge-Graph Memory in a Partially Observable Environment
von: Kim, Taewoon, et al.
Veröffentlicht: (2024)
von: Kim, Taewoon, et al.
Veröffentlicht: (2024)
Diffusion for World Modeling: Visual Details Matter in Atari
von: Alonso, Eloi, et al.
Veröffentlicht: (2024)
von: Alonso, Eloi, et al.
Veröffentlicht: (2024)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
Enhancing Two-Player Performance Through Single-Player Knowledge Transfer: An Empirical Study on Atari 2600 Games
von: Saadat, Kimiya, et al.
Veröffentlicht: (2024)
von: Saadat, Kimiya, et al.
Veröffentlicht: (2024)
Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari
von: Kim, Jooyeon
Veröffentlicht: (2026)
von: Kim, Jooyeon
Veröffentlicht: (2026)
Learning To Play Atari Games Using Dueling Q-Learning and Hebbian Plasticity
von: Salehin, Md Ashfaq
Veröffentlicht: (2024)
von: Salehin, Md Ashfaq
Veröffentlicht: (2024)
Decision Transformer vs. Decision Mamba: Analysing the Complexity of Sequential Decision Making in Atari Games
von: Yan, Ke
Veröffentlicht: (2024)
von: Yan, Ke
Veröffentlicht: (2024)
MiniZero: Comparative Analysis of AlphaZero and MuZero on Go, Othello, and Atari Games
von: Wu, Ti-Rong, et al.
Veröffentlicht: (2023)
von: Wu, Ti-Rong, et al.
Veröffentlicht: (2023)
Learnable Behavior Control: Breaking Atari Human World Records via Sample-Efficient Behavior Selection
von: Fan, Jiajun, et al.
Veröffentlicht: (2023)
von: Fan, Jiajun, et al.
Veröffentlicht: (2023)
TextAtari: 100K Frames Game Playing with Language Agents
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
A quantum-classical reinforcement learning model to play Atari games
von: Freinberger, Dominik, et al.
Veröffentlicht: (2024)
von: Freinberger, Dominik, et al.
Veröffentlicht: (2024)
Read and Reap the Rewards: Learning to Play Atari with the Help of Instruction Manuals
von: Wu, Yue, et al.
Veröffentlicht: (2023)
von: Wu, Yue, et al.
Veröffentlicht: (2023)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
Rotary Position Encodings for Graphs
von: Reid, Isaac, et al.
Veröffentlicht: (2025)
von: Reid, Isaac, et al.
Veröffentlicht: (2025)
The Relationship Between Reasoning and Performance in Large Language Models -- o3 (mini) Thinks Harder, Not Longer
von: Ballon, Marthe, et al.
Veröffentlicht: (2025)
von: Ballon, Marthe, et al.
Veröffentlicht: (2025)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
On the Stability of Expressive Positional Encodings for Graphs
von: Huang, Yinan, et al.
Veröffentlicht: (2023)
von: Huang, Yinan, et al.
Veröffentlicht: (2023)
Guiding Skill Discovery with Foundation Models
von: Yang, Zhao, et al.
Veröffentlicht: (2025)
von: Yang, Zhao, et al.
Veröffentlicht: (2025)
Support Vector Boosting Machine (SVBM): Enhancing Classification Performance with AdaBoost and Residual Connections
von: Lian, Junbo Jacob
Veröffentlicht: (2024)
von: Lian, Junbo Jacob
Veröffentlicht: (2024)
Locality Sensitive Sparse Encoding for Learning World Models Online
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Physics-Encoded Inverse Modeling for Arctic Snow Depth Prediction
von: Sampath, Akila, et al.
Veröffentlicht: (2026)
von: Sampath, Akila, et al.
Veröffentlicht: (2026)
Subjective Logic Encodings
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
Algebraic Positional Encodings
von: Kogkalidis, Konstantinos, et al.
Veröffentlicht: (2023)
von: Kogkalidis, Konstantinos, et al.
Veröffentlicht: (2023)
Training Language Models to Explain Their Own Computations
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
Confidence Optimization for Probabilistic Encoding
von: Xia, Pengjiu, et al.
Veröffentlicht: (2025)
von: Xia, Pengjiu, et al.
Veröffentlicht: (2025)
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
Leveraging weights signals -- Predicting and improving generalizability in reinforcement learning
von: Moulin, Olivier, et al.
Veröffentlicht: (2025)
von: Moulin, Olivier, et al.
Veröffentlicht: (2025)
REFORMER: A ChatGPT-Driven Data Synthesis Framework Elevating Text-to-SQL Models
von: Liu, Shenyang, et al.
Veröffentlicht: (2025)
von: Liu, Shenyang, et al.
Veröffentlicht: (2025)
Mesa-Extrapolation: A Weave Position Encoding Method for Enhanced Extrapolation in LLMs
von: Ma, Xin, et al.
Veröffentlicht: (2024)
von: Ma, Xin, et al.
Veröffentlicht: (2024)
Trajectory Encoding Temporal Graph Networks
von: Xiong, Jiafeng, et al.
Veröffentlicht: (2025)
von: Xiong, Jiafeng, et al.
Veröffentlicht: (2025)
Graph Transformers without Positional Encodings
von: Garg, Ayush
Veröffentlicht: (2024)
von: Garg, Ayush
Veröffentlicht: (2024)
Axial Neural Networks for Dimension-Free Foundation Models
von: Kim, Hyunsu, et al.
Veröffentlicht: (2025)
von: Kim, Hyunsu, et al.
Veröffentlicht: (2025)
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
von: Wei, Zhao, et al.
Veröffentlicht: (2025)
von: Wei, Zhao, et al.
Veröffentlicht: (2025)
FedLED: Label-Free Equipment Fault Diagnosis with Vertical Federated Transfer Learning
von: Shen, Jie, et al.
Veröffentlicht: (2023)
von: Shen, Jie, et al.
Veröffentlicht: (2023)
Analyzing Patient Daily Movement Behavior Dynamics Using Two-Stage Encoding Model
von: Cui, Jin, et al.
Veröffentlicht: (2025)
von: Cui, Jin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Disentangled (Un)Controllable Features
von: Kooi, Jacob E., et al.
Veröffentlicht: (2022) -
Novel RL approach for efficient Elevator Group Control Systems
von: Vaartjes, Nathan, et al.
Veröffentlicht: (2025) -
Hadamard Representation: Scaffolding Performance Across Model-free RL
von: Kooi, Jacob E., et al.
Veröffentlicht: (2024) -
Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability
von: Kim, Taewoon, et al.
Veröffentlicht: (2026) -
Temporal Knowledge-Graph Memory in a Partially Observable Environment
von: Kim, Taewoon, et al.
Veröffentlicht: (2024)