Saved in:
| Main Authors: | Armengol-Estapé, Jordi, Michalski, Vincent, Kumar, Ramnath, St-Charles, Pierre-Luc, Precup, Doina, Kahou, Samira Ebrahimi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.18751 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Behaviour Discovery and Attribution for Explainable Reinforcement Learning
by: Rishav, Rishav, et al.
Published: (2025)
by: Rishav, Rishav, et al.
Published: (2025)
Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment
by: Rahman, Aamer Abdul, et al.
Published: (2024)
by: Rahman, Aamer Abdul, et al.
Published: (2024)
Estimation of Head Motion in Structural MRI and its Impact on Cortical Thickness Measurements in Retrospective Data
by: Bricout, Charles, et al.
Published: (2025)
by: Bricout, Charles, et al.
Published: (2025)
Learning Multi-agent Multi-machine Tending by Mobile Robots
by: Abdalwhab, Abdalwhab, et al.
Published: (2024)
by: Abdalwhab, Abdalwhab, et al.
Published: (2024)
CAMMARL: Conformal Action Modeling in Multi Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2023)
by: Gupta, Nikunj, et al.
Published: (2023)
Zero-Shot Anomaly Detection with Dual-Branch Prompt Selection
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Adaptive Group Robust Ensemble Knowledge Distillation
by: Kenfack, Patrik, et al.
Published: (2024)
by: Kenfack, Patrik, et al.
Published: (2024)
Towards Fair In-Context Learning with Tabular Foundation Models
by: Kenfack, Patrik, et al.
Published: (2025)
by: Kenfack, Patrik, et al.
Published: (2025)
Locally Constrained Representations in Reinforcement Learning
by: Nath, Somjit, et al.
Published: (2022)
by: Nath, Somjit, et al.
Published: (2022)
Learning to Play Atari in a World of Tokens
by: Agarwal, Pranav, et al.
Published: (2024)
by: Agarwal, Pranav, et al.
Published: (2024)
Fairness Under Demographic Scarce Regime
by: Kenfack, Patrik Joslin, et al.
Published: (2023)
by: Kenfack, Patrik Joslin, et al.
Published: (2023)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
by: Carr, Jonathan Colaço, et al.
Published: (2023)
by: Carr, Jonathan Colaço, et al.
Published: (2023)
Diversity-Enriched Option-Critic
by: Kamat, Anand, et al.
Published: (2020)
by: Kamat, Anand, et al.
Published: (2020)
Functional Acceleration for Policy Mirror Descent
by: Chelu, Veronica, et al.
Published: (2024)
by: Chelu, Veronica, et al.
Published: (2024)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
by: Alver, Safa, et al.
Published: (2022)
by: Alver, Safa, et al.
Published: (2022)
Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments
by: Luo, Ziyan, et al.
Published: (2025)
by: Luo, Ziyan, et al.
Published: (2025)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
by: Yu, Xuemin, et al.
Published: (2026)
by: Yu, Xuemin, et al.
Published: (2026)
Cross-Layer Discrete Concept Discovery for Interpreting Language Models
by: Garg, Ankur, et al.
Published: (2025)
by: Garg, Ankur, et al.
Published: (2025)
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation
by: Azimi, Rambod, et al.
Published: (2024)
by: Azimi, Rambod, et al.
Published: (2024)
SLaDe: A Portable Small Language Model Decompiler for Optimized Assembly
by: Armengol-Estapé, Jordi, et al.
Published: (2023)
by: Armengol-Estapé, Jordi, et al.
Published: (2023)
Rotation-Preserving Supervised Fine-Tuning
by: Jin, Hangzhan, et al.
Published: (2026)
by: Jin, Hangzhan, et al.
Published: (2026)
Incorporating Spatial Information into Goal-Conditioned Hierarchical Reinforcement Learning via Graph Representations
by: Zhang, Shuyuan, et al.
Published: (2025)
by: Zhang, Shuyuan, et al.
Published: (2025)
Balancing Plasticity and Stability with Fast and Slow Successor Features
by: Chua, Raymond, et al.
Published: (2026)
by: Chua, Raymond, et al.
Published: (2026)
On the Privacy of Selection Mechanisms with Gaussian Noise
by: Lebensold, Jonathan, et al.
Published: (2024)
by: Lebensold, Jonathan, et al.
Published: (2024)
Comparative Analysis of Diffusion Generative Models in Computational Pathology
by: Thakkar, Denisha, et al.
Published: (2024)
by: Thakkar, Denisha, et al.
Published: (2024)
Don't Transform the Code, Code the Transforms: Towards Precise Code Rewriting using LLMs
by: Cummins, Chris, et al.
Published: (2024)
by: Cummins, Chris, et al.
Published: (2024)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
by: Jain, Arushi, et al.
Published: (2024)
by: Jain, Arushi, et al.
Published: (2024)
Fluid-Agent Reinforcement Learning
by: Sharma, Shishir, et al.
Published: (2026)
by: Sharma, Shishir, et al.
Published: (2026)
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
by: Alver, Safa, et al.
Published: (2024)
by: Alver, Safa, et al.
Published: (2024)
Handling Delay in Real-Time Reinforcement Learning
by: Anokhin, Ivan, et al.
Published: (2025)
by: Anokhin, Ivan, et al.
Published: (2025)
Survey on AI Ethics: A Socio-technical Perspective
by: Mbiazi, Dave, et al.
Published: (2023)
by: Mbiazi, Dave, et al.
Published: (2023)
Forklift: An Extensible Neural Lifter
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
LLMs versus the Halting Problem: Characterizing Program Termination Reasoning
by: Sultan, Oren, et al.
Published: (2026)
by: Sultan, Oren, et al.
Published: (2026)
Parseval Regularization for Continual Reinforcement Learning
by: Chung, Wesley, et al.
Published: (2024)
by: Chung, Wesley, et al.
Published: (2024)
Relative Trajectory Balance is equivalent to Trust-PCL
by: Deleu, Tristan, et al.
Published: (2025)
by: Deleu, Tristan, et al.
Published: (2025)
CryCeleb: A Speaker Verification Dataset Based on Infant Cry Sounds
by: Budaghyan, David, et al.
Published: (2023)
by: Budaghyan, David, et al.
Published: (2023)
Uncovering a Universal Abstract Algorithm for Modular Addition in Neural Networks
by: McCracken, Gavin, et al.
Published: (2025)
by: McCracken, Gavin, et al.
Published: (2025)
Solving Models of Economic Dynamics with Ridgeless Kernel Regressions
by: Kahou, Mahdi Ebrahimi, et al.
Published: (2024)
by: Kahou, Mahdi Ebrahimi, et al.
Published: (2024)
Prediction of Final Phosphorus Content of Steel in a Scrap-Based Electric Arc Furnace Using Artificial Neural Networks
by: Azzaz, Riadh, et al.
Published: (2024)
by: Azzaz, Riadh, et al.
Published: (2024)
Similar Items
-
Behaviour Discovery and Attribution for Explainable Reinforcement Learning
by: Rishav, Rishav, et al.
Published: (2025) -
Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment
by: Rahman, Aamer Abdul, et al.
Published: (2024) -
Estimation of Head Motion in Structural MRI and its Impact on Cortical Thickness Measurements in Retrospective Data
by: Bricout, Charles, et al.
Published: (2025) -
Learning Multi-agent Multi-machine Tending by Mobile Robots
by: Abdalwhab, Abdalwhab, et al.
Published: (2024) -
CAMMARL: Conformal Action Modeling in Multi Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2023)