Inverse Reinforcement Learning via Convex Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhu, Hao, Zhang, Yuan, Boedecker, Joschka |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Solving Inverse Problem for Multi-armed Bandits via Convex Optimization
por: Zhu, Hao, et al.
Publicado: (2025)
por: Zhu, Hao, et al.
Publicado: (2025)
Fitting Reinforcement Learning Model to Behavioral Data under Bandits
por: Zhu, Hao, et al.
Publicado: (2025)
por: Zhu, Hao, et al.
Publicado: (2025)
Probabilistic Recurrent Intention Switching Model
por: Sheng, Wenyuan, et al.
Publicado: (2026)
por: Sheng, Wenyuan, et al.
Publicado: (2026)
A Disentangled Low-Rank RNN Framework for Uncovering Neural Connectivity and Dynamics
por: Li, Chengrui, et al.
Publicado: (2025)
por: Li, Chengrui, et al.
Publicado: (2025)
Classification of Raw MEG/EEG Data with Detach-Rocket Ensemble: An Improved ROCKET Algorithm for Multivariate Time Series Analysis
por: Solana, Adrià, et al.
Publicado: (2024)
por: Solana, Adrià, et al.
Publicado: (2024)
BrainMass: Advancing Brain Network Analysis for Diagnosis with Large-scale Self-Supervised Learning
por: Yang, Yanwu, et al.
Publicado: (2024)
por: Yang, Yanwu, et al.
Publicado: (2024)
Deep Learning Models for Atypical Serotonergic Cells Recognition
por: Corradetti, Daniele, et al.
Publicado: (2024)
por: Corradetti, Daniele, et al.
Publicado: (2024)
Multi-intention Inverse Q-learning for Interpretable Behavior Representation
por: Zhu, Hao, et al.
Publicado: (2023)
por: Zhu, Hao, et al.
Publicado: (2023)
Disciplined Biconvex Programming
por: Zhu, Hao, et al.
Publicado: (2025)
por: Zhu, Hao, et al.
Publicado: (2025)
Exploring EEG Responses during Observation of Actions Performed by Human Actor and Humanoid Robot
por: Nguyen, Anh T., et al.
Publicado: (2025)
por: Nguyen, Anh T., et al.
Publicado: (2025)
Accurate identification of communication between multiple interacting neural populations
por: Liu, Belle, et al.
Publicado: (2025)
por: Liu, Belle, et al.
Publicado: (2025)
Synthetic Data for Discriminating Serotonergic Neurons using Convolutional Neural Networks
por: Corradetti, Daniele, et al.
Publicado: (2024)
por: Corradetti, Daniele, et al.
Publicado: (2024)
Identification of Epileptic Spasms (ESES) Phases Using EEG Signals: A Vision Transformer Approach
por: Gong, Wei, et al.
Publicado: (2024)
por: Gong, Wei, et al.
Publicado: (2024)
Similarity Matching Networks: Hebbian Learning and Convergence Over Multiple Time Scales
por: Centorrino, Veronica, et al.
Publicado: (2025)
por: Centorrino, Veronica, et al.
Publicado: (2025)
Brain-Like Replay Naturally Emerges in Reinforcement Learning Agents
por: Wang, Jiyi, et al.
Publicado: (2024)
por: Wang, Jiyi, et al.
Publicado: (2024)
Personalized Transcranial Electrical Stimulation: A Review of Computational Modeling and Optimization
por: Wang, Mo, et al.
Publicado: (2025)
por: Wang, Mo, et al.
Publicado: (2025)
Multi-convex Programming for Discrete Latent Factor Models Prototyping
por: Zhu, Hao, et al.
Publicado: (2025)
por: Zhu, Hao, et al.
Publicado: (2025)
The Principle of Maximum Heterogeneity Optimises Productivity in Distributed Production Systems Across Biology, Economics, and Computing
por: Artis, Guillhem, et al.
Publicado: (2026)
por: Artis, Guillhem, et al.
Publicado: (2026)
Plastic Arbor: a modern simulation framework for synaptic plasticity -- from single synapses to networks of morphological neurons
por: Luboeinski, Jannik, et al.
Publicado: (2024)
por: Luboeinski, Jannik, et al.
Publicado: (2024)
State-Space NTK Collapse Near Bifurcations
por: Hazelden, James, et al.
Publicado: (2026)
por: Hazelden, James, et al.
Publicado: (2026)
Role of Data-driven Regional Growth Model in Shaping Brain Folding Patterns
por: Hou, Jixin, et al.
Publicado: (2024)
por: Hou, Jixin, et al.
Publicado: (2024)
From Brain Models to Executable Digital Twins: Execution Semantics and Neuro-Neuromorphic Systems
por: Muzy, Alexandre
Publicado: (2026)
por: Muzy, Alexandre
Publicado: (2026)
Accelerating Fleet Upgrade Decisions with Machine-Learning Enhanced Optimization
por: Chai, Kenrick Howin, et al.
Publicado: (2025)
por: Chai, Kenrick Howin, et al.
Publicado: (2025)
Multi-fidelity Bayesian Optimization: A Review
por: Do, Bach, et al.
Publicado: (2023)
por: Do, Bach, et al.
Publicado: (2023)
Adjoint-Based Aerodynamic Shape Optimization with a Manifold Constraint Learned by Diffusion Models
por: Chen, Long, et al.
Publicado: (2025)
por: Chen, Long, et al.
Publicado: (2025)
Bayesian Optimization under Uncertainty for Training a Scale Parameter in Stochastic Models
por: Yadav, Akash, et al.
Publicado: (2025)
por: Yadav, Akash, et al.
Publicado: (2025)
Laser Scan Path Design for Controlled Microstructure in Additive Manufacturing with Integrated Reduced-Order Phase-Field Modeling and Deep Reinforcement Learning
por: Twumasi, Augustine, et al.
Publicado: (2025)
por: Twumasi, Augustine, et al.
Publicado: (2025)
An Online Machine Learning Multi-resolution Optimization Framework for Energy System Design Limit of Performance Analysis
por: Amusat, Oluwamayowa O., et al.
Publicado: (2026)
por: Amusat, Oluwamayowa O., et al.
Publicado: (2026)
Multi-Source Neural Activity Indices for EEG/MEG Localization: A Two-Stage Spatial Filtering Framework and Extension to MNE-Python
por: Jurkowska, Julia, et al.
Publicado: (2025)
por: Jurkowska, Julia, et al.
Publicado: (2025)
Definition of Cybernetical Neuroscience
por: Fradkov, Alexander
Publicado: (2024)
por: Fradkov, Alexander
Publicado: (2024)
A Differentiable Approach to Multi-scale Brain Modeling
por: Wang, Chaoming, et al.
Publicado: (2024)
por: Wang, Chaoming, et al.
Publicado: (2024)
Minimizing Structural Vibrations via Guided Flow Matching Design Optimization
por: van Delden, Jan, et al.
Publicado: (2025)
por: van Delden, Jan, et al.
Publicado: (2025)
Controlgym: Large-Scale Control Environments for Benchmarking Reinforcement Learning Algorithms
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease
por: Ravivarapu, Harsh, et al.
Publicado: (2025)
por: Ravivarapu, Harsh, et al.
Publicado: (2025)
Positive Competitive Networks for Sparse Reconstruction
por: Centorrino, Veronica, et al.
Publicado: (2023)
por: Centorrino, Veronica, et al.
Publicado: (2023)
Towards Autonomous Experimentation: Bayesian Optimization over Problem Formulation Space for Accelerated Alloy Development
por: Khatamsaz, Danial, et al.
Publicado: (2025)
por: Khatamsaz, Danial, et al.
Publicado: (2025)
Predicting Neuromodulation Outcome for Parkinson's Disease with Generative Virtual Brain Model
por: Du, Siyuan, et al.
Publicado: (2026)
por: Du, Siyuan, et al.
Publicado: (2026)
Learning Mixtures of Linear Dynamical Systems via Hybrid Tensor-EM Method
por: Gong, Lulu, et al.
Publicado: (2025)
por: Gong, Lulu, et al.
Publicado: (2025)
AI Driven Laser Parameter Search: Inverse Design of Photonic Surfaces using Greedy Surrogate-based Optimization
por: Grbcic, Luka, et al.
Publicado: (2024)
por: Grbcic, Luka, et al.
Publicado: (2024)
Estimating the Unobservable Components of Electricity Demand Response with Inverse Optimization
por: Esteban-Perez, Adrian, et al.
Publicado: (2024)
por: Esteban-Perez, Adrian, et al.
Publicado: (2024)
Ejemplares similares
-
Solving Inverse Problem for Multi-armed Bandits via Convex Optimization
por: Zhu, Hao, et al.
Publicado: (2025) -
Fitting Reinforcement Learning Model to Behavioral Data under Bandits
por: Zhu, Hao, et al.
Publicado: (2025) -
Probabilistic Recurrent Intention Switching Model
por: Sheng, Wenyuan, et al.
Publicado: (2026) -
A Disentangled Low-Rank RNN Framework for Uncovering Neural Connectivity and Dynamics
por: Li, Chengrui, et al.
Publicado: (2025) -
Classification of Raw MEG/EEG Data with Detach-Rocket Ensemble: An Improved ROCKET Algorithm for Multivariate Time Series Analysis
por: Solana, Adrià, et al.
Publicado: (2024)