Controllability in preference-conditioned multi-objective reinforcement learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Molins, Pau de las Heras, Yalcinkaya, Beyazit, Peters, Lasse, Fridovich-Keil, David, Bakirtzis, Georgios |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Approximate solutions to games of ordered preference
por: Molins, Pau de las Heras, et al.
Publicado: (2025)
por: Molins, Pau de las Heras, et al.
Publicado: (2025)
Breaking Exponential Complexity in Games of Ordered Preference: A Tractable Reformulation
por: Lee, Dong Ho, et al.
Publicado: (2026)
por: Lee, Dong Ho, et al.
Publicado: (2026)
Decomposing Control Lyapunov Functions for Efficient Reinforcement Learning
por: Lopez, Antonio, et al.
Publicado: (2024)
por: Lopez, Antonio, et al.
Publicado: (2024)
Bayesian Inverse Games with High-Dimensional Multi-Modal Observations
por: Jain, Yash, et al.
Publicado: (2026)
por: Jain, Yash, et al.
Publicado: (2026)
Stealing That Free Lunch: Exposing the Limits of Dyna-Style Reinforcement Learning
por: Barkley, Brett, et al.
Publicado: (2024)
por: Barkley, Brett, et al.
Publicado: (2024)
A Forensic Analysis of Synthetic Data in RL: Diagnosing and Solving Algorithmic Failures in Model-Based Policy Optimization
por: Barkley, Brett, et al.
Publicado: (2025)
por: Barkley, Brett, et al.
Publicado: (2025)
Auto-Encoding Bayesian Inverse Games
por: Liu, Xinjie, et al.
Publicado: (2024)
por: Liu, Xinjie, et al.
Publicado: (2024)
Learning responsibility allocations for multi-agent interactions: A differentiable optimization approach with control barrier functions
por: Remy, Isaac, et al.
Publicado: (2024)
por: Remy, Isaac, et al.
Publicado: (2024)
Scaling Pretrained Representations Enables Label-Free Out-of-Distribution Detection Without Fine-Tuning
por: Barkley, Brett, et al.
Publicado: (2026)
por: Barkley, Brett, et al.
Publicado: (2026)
SCOPED: Score-Curvature Out-of-distribution Proximity Evaluator for Diffusion
por: Barkley, Brett, et al.
Publicado: (2025)
por: Barkley, Brett, et al.
Publicado: (2025)
A Recovery Guarantee for Sparse Neural Networks
por: Fridovich-Keil, Sara, et al.
Publicado: (2025)
por: Fridovich-Keil, Sara, et al.
Publicado: (2025)
Neural Operators for Multi-Task Control and Adaptation
por: Sewell, David, et al.
Publicado: (2026)
por: Sewell, David, et al.
Publicado: (2026)
Learning to Walk from Three Minutes of Real-World Data with Semi-structured Dynamics Models
por: Levy, Jacob, et al.
Publicado: (2024)
por: Levy, Jacob, et al.
Publicado: (2024)
KLIP: localized distribution shift detection via KL-divergence with diffusion priors in Inverse Problems
por: Kheirandish, Alireza, et al.
Publicado: (2026)
por: Kheirandish, Alireza, et al.
Publicado: (2026)
You Can't Always Get What You Want: Games of Ordered Preference
por: Lee, Dong Ho, et al.
Publicado: (2024)
por: Lee, Dong Ho, et al.
Publicado: (2024)
Compositional Automata Embeddings for Goal-Conditioned Reinforcement Learning
por: Yalcinkaya, Beyazit, et al.
Publicado: (2024)
por: Yalcinkaya, Beyazit, et al.
Publicado: (2024)
Provably Correct Automata Embeddings for Optimal Automata-Conditioned Reinforcement Learning
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
Categorical semantics of compositional reinforcement learning
por: Bakirtzis, Georgios, et al.
Publicado: (2022)
por: Bakirtzis, Georgios, et al.
Publicado: (2022)
MOMA-AC: A preference-driven actor-critic framework for continuous multi-objective multi-agent reinforcement learning
por: Callaghan, Adam, et al.
Publicado: (2025)
por: Callaghan, Adam, et al.
Publicado: (2025)
Monotonic Transformation Invariant Multi-task Learning
por: Murthy, Surya, et al.
Publicado: (2025)
por: Murthy, Surya, et al.
Publicado: (2025)
Symbolic Regression on Sparse and Noisy Data with Gaussian Processes
por: Hsin, Junette, et al.
Publicado: (2023)
por: Hsin, Junette, et al.
Publicado: (2023)
HEART: A High-Efficiency Adaptive Real-Time Telemonitoring Framework for Secure Electrocardiogram Signal Transmission Using Chaotic Encryption
por: Yuksel, Beyazıt Bestami
Publicado: (2026)
por: Yuksel, Beyazıt Bestami
Publicado: (2026)
Generalized Information Gathering Under Dynamics Uncertainty
por: Palafox, Fernando, et al.
Publicado: (2026)
por: Palafox, Fernando, et al.
Publicado: (2026)
Compositional shield synthesis for safe reinforcement learning in partial observability
por: Carr, Steven, et al.
Publicado: (2025)
por: Carr, Steven, et al.
Publicado: (2025)
A Framework for Finding Local Saddle Points in Two-Player Zero-Sum Black-Box Games
por: Agarwal, Shubhankar, et al.
Publicado: (2025)
por: Agarwal, Shubhankar, et al.
Publicado: (2025)
A Flow Matching Algorithm for Many-Shot Adaptation to Unseen Distributions
por: Ingebrand, Tyler, et al.
Publicado: (2026)
por: Ingebrand, Tyler, et al.
Publicado: (2026)
A Multi-Fidelity Control Variate Approach for Policy Gradient Estimation
por: Liu, Xinjie, et al.
Publicado: (2025)
por: Liu, Xinjie, et al.
Publicado: (2025)
Data-assimilated model-informed reinforcement learning
por: Ozan, Defne E., et al.
Publicado: (2025)
por: Ozan, Defne E., et al.
Publicado: (2025)
Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
Load constrained wind farm flow control through multi-objective multi-agent reinforcement learning
por: Åstrand, Teodor, et al.
Publicado: (2026)
por: Åstrand, Teodor, et al.
Publicado: (2026)
Data-Driven Modeling and Correction of Vehicle Dynamics
por: Ly, Nguyen, et al.
Publicado: (2025)
por: Ly, Nguyen, et al.
Publicado: (2025)
Middle-mile logistics through the lens of goal-conditioned reinforcement learning
por: Eberhard, Onno, et al.
Publicado: (2026)
por: Eberhard, Onno, et al.
Publicado: (2026)
Bridging the phenotype-target gap for molecular generation via multi-objective reinforcement learning
por: Guo, Haotian, et al.
Publicado: (2025)
por: Guo, Haotian, et al.
Publicado: (2025)
Reduce, Reuse, Recycle: Categories for Compositional Reinforcement Learning
por: Bakirtzis, Georgios, et al.
Publicado: (2024)
por: Bakirtzis, Georgios, et al.
Publicado: (2024)
FoX: Formation-aware exploration in multi-agent reinforcement learning
por: Jo, Yonghyeon, et al.
Publicado: (2023)
por: Jo, Yonghyeon, et al.
Publicado: (2023)
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
por: Mirbakhsh, Shahin, et al.
Publicado: (2024)
por: Mirbakhsh, Shahin, et al.
Publicado: (2024)
Generalizing Safety Beyond Collision-Avoidance via Latent-Space Reachability Analysis
por: Nakamura, Kensuke, et al.
Publicado: (2025)
por: Nakamura, Kensuke, et al.
Publicado: (2025)
Solving Inverse Problems in Protein Space Using Diffusion-Based Priors
por: Levy, Axel, et al.
Publicado: (2024)
por: Levy, Axel, et al.
Publicado: (2024)
Accurate, provable and fast polychromatic tomographic reconstruction: A variational inequality approach
por: Lou, Mengqi, et al.
Publicado: (2025)
por: Lou, Mengqi, et al.
Publicado: (2025)
Hardware-Aware Federated Learning for Speech Emotion Recognition
por: Yuksel, Beyazit Bestami, et al.
Publicado: (2026)
por: Yuksel, Beyazit Bestami, et al.
Publicado: (2026)
Ejemplares similares
-
Approximate solutions to games of ordered preference
por: Molins, Pau de las Heras, et al.
Publicado: (2025) -
Breaking Exponential Complexity in Games of Ordered Preference: A Tractable Reformulation
por: Lee, Dong Ho, et al.
Publicado: (2026) -
Decomposing Control Lyapunov Functions for Efficient Reinforcement Learning
por: Lopez, Antonio, et al.
Publicado: (2024) -
Bayesian Inverse Games with High-Dimensional Multi-Modal Observations
por: Jain, Yash, et al.
Publicado: (2026) -
Stealing That Free Lunch: Exposing the Limits of Dyna-Style Reinforcement Learning
por: Barkley, Brett, et al.
Publicado: (2024)