Saved in:
| Main Authors: | Voelcker, Claas A, Hussing, Marcel, Eaton, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.08870 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dissecting Deep RL with High Update Ratios: Combatting Value Divergence
by: Hussing, Marcel, et al.
Published: (2024)
by: Hussing, Marcel, et al.
Published: (2024)
MAD-TD: Model-Augmented Data stabilizes High Update Ratio RL
by: Voelcker, Claas A, et al.
Published: (2024)
by: Voelcker, Claas A, et al.
Published: (2024)
Behavior-Consistent Deep Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2026)
by: Hussing, Marcel, et al.
Published: (2026)
Relative Entropy Pathwise Policy Optimization
by: Voelcker, Claas, et al.
Published: (2025)
by: Voelcker, Claas, et al.
Published: (2025)
Distributed Continual Learning
by: Le, Long, et al.
Published: (2024)
by: Le, Long, et al.
Published: (2024)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2023)
by: Hussing, Marcel, et al.
Published: (2023)
Replicable Reinforcement Learning with Linear Function Approximation
by: Eaton, Eric, et al.
Published: (2025)
by: Eaton, Eric, et al.
Published: (2025)
Intersectional Fairness in Reinforcement Learning with Large State and Constraint Spaces
by: Eaton, Eric, et al.
Published: (2025)
by: Eaton, Eric, et al.
Published: (2025)
Iterative Compositional Data Generation for Robot Control
by: Pham, Anh-Quan, et al.
Published: (2025)
by: Pham, Anh-Quan, et al.
Published: (2025)
Model Agreement via Anchoring
by: Eaton, Eric, et al.
Published: (2026)
by: Eaton, Eric, et al.
Published: (2026)
When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning
by: Voelcker, Claas, et al.
Published: (2024)
by: Voelcker, Claas, et al.
Published: (2024)
Test-Time Graph Search for Goal-Conditioned Reinforcement Learning
by: Opryshko, Evgenii, et al.
Published: (2025)
by: Opryshko, Evgenii, et al.
Published: (2025)
$λ$-models: Effective Decision-Aware Reinforcement Learning with Latent Models
by: Voelcker, Claas A, et al.
Published: (2023)
by: Voelcker, Claas A, et al.
Published: (2023)
Calibrated Value-Aware Model Learning with Probabilistic Environment Models
by: Voelcker, Claas, et al.
Published: (2025)
by: Voelcker, Claas, et al.
Published: (2025)
Sorrel: A simple and flexible framework for multi-agent reinforcement learning
by: Gelpí, Rebekah A., et al.
Published: (2025)
by: Gelpí, Rebekah A., et al.
Published: (2025)
Temporal-Difference Learning Using Distributed Error Signals
by: Guan, Jonas, et al.
Published: (2024)
by: Guan, Jonas, et al.
Published: (2024)
Oracle-Efficient Reinforcement Learning for Max Value Ensembles
by: Hussing, Marcel, et al.
Published: (2024)
by: Hussing, Marcel, et al.
Published: (2024)
A practical generalization metric for deep networks benchmarking
by: Huang, Mengqing, et al.
Published: (2024)
by: Huang, Mengqing, et al.
Published: (2024)
Can time series forecasting be automated? A benchmark and analysis
by: Sreedhara, Anvitha Thirthapura, et al.
Published: (2024)
by: Sreedhara, Anvitha Thirthapura, et al.
Published: (2024)
Evaluating CUDA Tile for AI Workloads on Hopper and Blackwell GPUs
by: Yadav, Divakar Kumar, et al.
Published: (2026)
by: Yadav, Divakar Kumar, et al.
Published: (2026)
Can we generate portable representations for clinical time series data using LLMs?
by: Ji, Zongliang, et al.
Published: (2026)
by: Ji, Zongliang, et al.
Published: (2026)
MotifBench: A standardized protein design benchmark for motif-scaffolding problems
by: Zheng, Zhuoqi, et al.
Published: (2025)
by: Zheng, Zhuoqi, et al.
Published: (2025)
IBCL: Zero-shot Model Generation under Stability-Plasticity Trade-offs
by: Lu, Pengyuan, et al.
Published: (2023)
by: Lu, Pengyuan, et al.
Published: (2023)
Practical FP4 Training for Large-Scale MoE Models on Hopper GPUs
by: Zhang, Wuyue, et al.
Published: (2026)
by: Zhang, Wuyue, et al.
Published: (2026)
LLM-Cave: A benchmark and light environment for large language models reasoning and decision-making system
by: Li, Huanyu, et al.
Published: (2025)
by: Li, Huanyu, et al.
Published: (2025)
Influence functions and regularity tangents for efficient active learning
by: Eaton, Frederik
Published: (2024)
by: Eaton, Frederik
Published: (2024)
FORLA: Federated Object-centric Representation Learning with Slot Attention
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Can we trust the evaluation on ChatGPT?
by: Aiyappa, Rachith, et al.
Published: (2023)
by: Aiyappa, Rachith, et al.
Published: (2023)
Is CLIP ideal? No. Can we fix it? Yes!
by: Kang, Raphi, et al.
Published: (2025)
by: Kang, Raphi, et al.
Published: (2025)
Can we Soft Prompt LLMs for Graph Learning Tasks?
by: Liu, Zheyuan, et al.
Published: (2024)
by: Liu, Zheyuan, et al.
Published: (2024)
Can we Improve Prediction of Psychotherapy Outcomes Through Pretraining With Simulated Data?
by: Jacobs, Niklas, et al.
Published: (2026)
by: Jacobs, Niklas, et al.
Published: (2026)
GreenLight-Gym: Reinforcement learning benchmark environment for control of greenhouse production systems
by: van Laatum, Bart, et al.
Published: (2024)
by: van Laatum, Bart, et al.
Published: (2024)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Extreme Weather Bench: A framework and benchmark for evaluation of high-impact weather
by: McGovern, Amy, et al.
Published: (2026)
by: McGovern, Amy, et al.
Published: (2026)
Deployment-complete benchmarking
by: Mansouri, El Mustapha, et al.
Published: (2026)
by: Mansouri, El Mustapha, et al.
Published: (2026)
BlendedNet++: A dataset and benchmark for field-resolved aerodynamics and inverse design of blended wing body aircraft
by: Sung, Nicholas, et al.
Published: (2025)
by: Sung, Nicholas, et al.
Published: (2025)
BEARCUBS: A benchmark for computer-using web agents
by: Song, Yixiao, et al.
Published: (2025)
by: Song, Yixiao, et al.
Published: (2025)
Decoding Human Preferences in Alignment: An Improved Approach to Inverse Constitutional AI
by: Henneking, Carl-Leander, et al.
Published: (2025)
by: Henneking, Carl-Leander, et al.
Published: (2025)
Citegeist: Automated Generation of Related Work Analysis on the arXiv Corpus
by: Beger, Claas, et al.
Published: (2025)
by: Beger, Claas, et al.
Published: (2025)
PDCNet: a benchmark and general deep learning framework for activity prediction of peptide-drug conjugates
by: Liu, Yun, et al.
Published: (2025)
by: Liu, Yun, et al.
Published: (2025)
Similar Items
-
Dissecting Deep RL with High Update Ratios: Combatting Value Divergence
by: Hussing, Marcel, et al.
Published: (2024) -
MAD-TD: Model-Augmented Data stabilizes High Update Ratio RL
by: Voelcker, Claas A, et al.
Published: (2024) -
Behavior-Consistent Deep Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2026) -
Relative Entropy Pathwise Policy Optimization
by: Voelcker, Claas, et al.
Published: (2025) -
Distributed Continual Learning
by: Le, Long, et al.
Published: (2024)