Saved in:
| Main Authors: | Voelcker, Claas, Kastner, Tyler, Gilitschenski, Igor, Farahmand, Amir-massoud |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.17718 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$λ$-models: Effective Decision-Aware Reinforcement Learning with Latent Models
by: Voelcker, Claas A, et al.
Published: (2023)
by: Voelcker, Claas A, et al.
Published: (2023)
Dissecting Deep RL with High Update Ratios: Combatting Value Divergence
by: Hussing, Marcel, et al.
Published: (2024)
by: Hussing, Marcel, et al.
Published: (2024)
MAD-TD: Model-Augmented Data stabilizes High Update Ratio RL
by: Voelcker, Claas A, et al.
Published: (2024)
by: Voelcker, Claas A, et al.
Published: (2024)
Calibrated Value-Aware Model Learning with Probabilistic Environment Models
by: Voelcker, Claas, et al.
Published: (2025)
by: Voelcker, Claas, et al.
Published: (2025)
Relative Entropy Pathwise Policy Optimization
by: Voelcker, Claas, et al.
Published: (2025)
by: Voelcker, Claas, et al.
Published: (2025)
Test-Time Graph Search for Goal-Conditioned Reinforcement Learning
by: Opryshko, Evgenii, et al.
Published: (2025)
by: Opryshko, Evgenii, et al.
Published: (2025)
PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling
by: Ma, Avery, et al.
Published: (2025)
by: Ma, Avery, et al.
Published: (2025)
PID Accelerated Temporal Difference Algorithms
by: Bedaywi, Mark, et al.
Published: (2024)
by: Bedaywi, Mark, et al.
Published: (2024)
Efficient and Accurate Optimal Transport with Mirror Descent and Conjugate Gradients
by: Kemertas, Mete, et al.
Published: (2023)
by: Kemertas, Mete, et al.
Published: (2023)
Press Start to Charge: Videogaming the Online Centralized Charging Scheduling Problem
by: Ghahtarani, Alireza, et al.
Published: (2026)
by: Ghahtarani, Alireza, et al.
Published: (2026)
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024)
by: Lee, Jongmin, et al.
Published: (2024)
A Truncated Newton Method for Optimal Transport
by: Kemertas, Mete, et al.
Published: (2025)
by: Kemertas, Mete, et al.
Published: (2025)
Majority of the Bests: Improving Best-of-N via Bootstrapping
by: Rakhsha, Amin, et al.
Published: (2025)
by: Rakhsha, Amin, et al.
Published: (2025)
Improving Adversarial Transferability via Model Alignment
by: Ma, Avery, et al.
Published: (2023)
by: Ma, Avery, et al.
Published: (2023)
Behavior-Consistent Deep Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2026)
by: Hussing, Marcel, et al.
Published: (2026)
Can we hop in general? A discussion of benchmark selection and design using the Hopper environment
by: Voelcker, Claas A, et al.
Published: (2024)
by: Voelcker, Claas A, et al.
Published: (2024)
Generating Auxiliary Tasks with Reinforcement Learning
by: Goldfeder, Judah, et al.
Published: (2025)
by: Goldfeder, Judah, et al.
Published: (2025)
TESPEC: Temporally-Enhanced Self-Supervised Pretraining for Event Cameras
by: Mohammadi, Mohammad, et al.
Published: (2025)
by: Mohammadi, Mohammad, et al.
Published: (2025)
CoCoNUT: Structural Code Understanding does not fall out of a tree
by: Beger, Claas, et al.
Published: (2025)
by: Beger, Claas, et al.
Published: (2025)
Energy Efficient Task Offloading in UAV-Enabled MEC Using a Fully Decentralized Deep Reinforcement Learning Approach
by: Asadian-Rad, Hamidreza, et al.
Published: (2025)
by: Asadian-Rad, Hamidreza, et al.
Published: (2025)
Reinforcement Learning via Auxiliary Task Distillation
by: Harish, Abhinav Narayan, et al.
Published: (2024)
by: Harish, Abhinav Narayan, et al.
Published: (2024)
Enabling Asymmetric Knowledge Transfer in Multi-Task Learning with Self-Auxiliaries
by: Graffeuille, Olivier, et al.
Published: (2024)
by: Graffeuille, Olivier, et al.
Published: (2024)
Temporal-Difference Learning Using Distributed Error Signals
by: Guan, Jonas, et al.
Published: (2024)
by: Guan, Jonas, et al.
Published: (2024)
Producing and Leveraging Online Map Uncertainty in Trajectory Prediction
by: Gu, Xunjiang, et al.
Published: (2024)
by: Gu, Xunjiang, et al.
Published: (2024)
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
by: Gu, Xunjiang, et al.
Published: (2024)
by: Gu, Xunjiang, et al.
Published: (2024)
Enhancing Molecular Property Prediction with Auxiliary Learning and Task-Specific Adaptation
by: Dey, Vishal, et al.
Published: (2024)
by: Dey, Vishal, et al.
Published: (2024)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
by: Zheng, Shuhong, et al.
Published: (2025)
by: Zheng, Shuhong, et al.
Published: (2025)
Improving Reinforcement Learning Efficiency with Auxiliary Tasks in Non-Visual Environments: A Comparison
by: Lange, Moritz, et al.
Published: (2023)
by: Lange, Moritz, et al.
Published: (2023)
When Does Neuroevolution Outcompete Reinforcement Learning in Transfer Learning Tasks?
by: Nisioti, Eleni, et al.
Published: (2025)
by: Nisioti, Eleni, et al.
Published: (2025)
$α$VIL: Learning to Leverage Auxiliary Tasks for Multitask Learning
by: Kourdis, Rafael, et al.
Published: (2024)
by: Kourdis, Rafael, et al.
Published: (2024)
When does Subagging Work?
by: Revelas, Christos, et al.
Published: (2024)
by: Revelas, Christos, et al.
Published: (2024)
SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation
by: Namekata, Koichi, et al.
Published: (2024)
by: Namekata, Koichi, et al.
Published: (2024)
How does Bayesian Sampling help Membership Inference Attacks?
by: Liu, Zhenlong, et al.
Published: (2025)
by: Liu, Zhenlong, et al.
Published: (2025)
Sorrel: A simple and flexible framework for multi-agent reinforcement learning
by: Gelpí, Rebekah A., et al.
Published: (2025)
by: Gelpí, Rebekah A., et al.
Published: (2025)
When does a bridge become an aeroplane?
by: Dardeno, Tina A., et al.
Published: (2024)
by: Dardeno, Tina A., et al.
Published: (2024)
Anomaly Detection for Scalable Task Grouping in Reinforcement Learning-based RAN Optimization
by: Li, Jimmy, et al.
Published: (2023)
by: Li, Jimmy, et al.
Published: (2023)
Robust Model-Based Reinforcement Learning with an Adversarial Auxiliary Model
by: Herremans, Siemen, et al.
Published: (2024)
by: Herremans, Siemen, et al.
Published: (2024)
Learning Representation for Multitask learning through Self Supervised Auxiliary learning
by: Shin, Seokwon, et al.
Published: (2024)
by: Shin, Seokwon, et al.
Published: (2024)
Deep Learning Optimization Using Self-Adaptive Weighted Auxiliary Variables
by: Liu, Yaru, et al.
Published: (2025)
by: Liu, Yaru, et al.
Published: (2025)
Realistic Evaluation of Model Merging for Compositional Generalization
by: Tam, Derek, et al.
Published: (2024)
by: Tam, Derek, et al.
Published: (2024)
Similar Items
-
$λ$-models: Effective Decision-Aware Reinforcement Learning with Latent Models
by: Voelcker, Claas A, et al.
Published: (2023) -
Dissecting Deep RL with High Update Ratios: Combatting Value Divergence
by: Hussing, Marcel, et al.
Published: (2024) -
MAD-TD: Model-Augmented Data stabilizes High Update Ratio RL
by: Voelcker, Claas A, et al.
Published: (2024) -
Calibrated Value-Aware Model Learning with Probabilistic Environment Models
by: Voelcker, Claas, et al.
Published: (2025) -
Relative Entropy Pathwise Policy Optimization
by: Voelcker, Claas, et al.
Published: (2025)