Model-Free Inference of Investor Preferences: A Relative Entropy IRL Approach
Fuente:
arXiv
Salvato in:
| Autore principale: | Xu, Chen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Maximum Causal Entropy IRL in Mean-Field Games and GNEP Framework for Forward RL
di: Anahtarci, Berkay, et al.
Pubblicazione: (2024)
di: Anahtarci, Berkay, et al.
Pubblicazione: (2024)
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning
di: Wan, Zhenglin, et al.
Pubblicazione: (2025)
di: Wan, Zhenglin, et al.
Pubblicazione: (2025)
Scouting By Reward: VLM-TO-IRL-Driven Player Selection For Esports
di: Yan, Qing, et al.
Pubblicazione: (2026)
di: Yan, Qing, et al.
Pubblicazione: (2026)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
di: Jain, Gauri, et al.
Pubblicazione: (2024)
di: Jain, Gauri, et al.
Pubblicazione: (2024)
CoMI-IRL: Contrastive Multi-Intention Inverse Reinforcement Learning
di: Mone, Antonio, et al.
Pubblicazione: (2026)
di: Mone, Antonio, et al.
Pubblicazione: (2026)
Relative Entropy Estimation in Function Space: Theory and Applications to Trajectory Inference
di: Wang, Chao, et al.
Pubblicazione: (2026)
di: Wang, Chao, et al.
Pubblicazione: (2026)
TreeIRL: Safe Urban Driving with Tree Search and Inverse Reinforcement Learning
di: Tomov, Momchil S., et al.
Pubblicazione: (2025)
di: Tomov, Momchil S., et al.
Pubblicazione: (2025)
SurgIRL: Towards Life-Long Learning for Surgical Automation by Incremental Reinforcement Learning
di: Ho, Yun-Jie, et al.
Pubblicazione: (2024)
di: Ho, Yun-Jie, et al.
Pubblicazione: (2024)
CoIRL-AD: Collaborative-Competitive Imitation-Reinforcement Learning in Latent World Models for Autonomous Driving
di: Zheng, Xiaoji, et al.
Pubblicazione: (2025)
di: Zheng, Xiaoji, et al.
Pubblicazione: (2025)
A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models
di: Kang, Enoch Hyunwook
Pubblicazione: (2026)
di: Kang, Enoch Hyunwook
Pubblicazione: (2026)
FP-IRL: Fokker--Planck Inverse Reinforcement Learning -- A Physics-Constrained Approach to Markov Decision Processes
di: Huang, Chengyang, et al.
Pubblicazione: (2023)
di: Huang, Chengyang, et al.
Pubblicazione: (2023)
Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time
di: Lan, Yifan, et al.
Pubblicazione: (2025)
di: Lan, Yifan, et al.
Pubblicazione: (2025)
State-Free Inference of State-Space Models: The Transfer Function Approach
di: Parnichkun, Rom N., et al.
Pubblicazione: (2024)
di: Parnichkun, Rom N., et al.
Pubblicazione: (2024)
From Reward-Free Representations to Preferences: Rethinking Offline Preference-Based Reinforcement Learning
di: Yang, Jun-Jie, et al.
Pubblicazione: (2026)
di: Yang, Jun-Jie, et al.
Pubblicazione: (2026)
AMPS: Adaptive Modality Preference Steering via Functional Entropy
di: Huang, Zihan, et al.
Pubblicazione: (2026)
di: Huang, Zihan, et al.
Pubblicazione: (2026)
GEM: Generative Entropy-Guided Preference Modeling for Few-shot Alignment of LLMs
di: Zhao, Yiyang, et al.
Pubblicazione: (2025)
di: Zhao, Yiyang, et al.
Pubblicazione: (2025)
Relative Entropy Pathwise Policy Optimization
di: Voelcker, Claas, et al.
Pubblicazione: (2025)
di: Voelcker, Claas, et al.
Pubblicazione: (2025)
Fisher Random Walk: Automatic Debiasing Contextual Preference Inference for Large Language Model Evaluation
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
Entropy Controllable Direct Preference Optimization
di: Omura, Motoki, et al.
Pubblicazione: (2024)
di: Omura, Motoki, et al.
Pubblicazione: (2024)
Graph State-Space Models and Latent Relational Inference
di: Zambon, Daniele, et al.
Pubblicazione: (2023)
di: Zambon, Daniele, et al.
Pubblicazione: (2023)
Inference of Utilities and Time Preference in Sequential Decision-Making
di: Cao, Haoyang, et al.
Pubblicazione: (2024)
di: Cao, Haoyang, et al.
Pubblicazione: (2024)
Mix- and MoE-DPO: A Variational Inference Approach to Direct Preference Optimization
di: Bohne, Jason, et al.
Pubblicazione: (2025)
di: Bohne, Jason, et al.
Pubblicazione: (2025)
Entropy-regularized Gradient Estimators for Approximate Bayesian Inference
di: Kaur, Jasmeet
Pubblicazione: (2025)
di: Kaur, Jasmeet
Pubblicazione: (2025)
Variational Inference, Entropy, and Orthogonality: A Unified Theory of Mixture-of-Experts
di: Su, Ye, et al.
Pubblicazione: (2026)
di: Su, Ye, et al.
Pubblicazione: (2026)
Benchmark of Likelihood-Free Inference Methods based on Neural and Optimal Transport Approaches
di: Aka, Samira, et al.
Pubblicazione: (2026)
di: Aka, Samira, et al.
Pubblicazione: (2026)
No Free Lunch: Non-Asymptotic Analysis of Prediction-Powered Inference
di: Mani, Pranav, et al.
Pubblicazione: (2025)
di: Mani, Pranav, et al.
Pubblicazione: (2025)
IRPM: Intergroup Relative Preference Modeling for Pointwise Generative Reward Models
di: Song, Haonan, et al.
Pubblicazione: (2026)
di: Song, Haonan, et al.
Pubblicazione: (2026)
Theoretical Insights in Model Inversion Robustness and Conditional Entropy Maximization for Collaborative Inference Systems
di: Xia, Song, et al.
Pubblicazione: (2025)
di: Xia, Song, et al.
Pubblicazione: (2025)
Interpretable Relational Inference with LLM-Guided Symbolic Dynamics Modeling
di: Liang, Xiaoxiao, et al.
Pubblicazione: (2026)
di: Liang, Xiaoxiao, et al.
Pubblicazione: (2026)
Label-Free Reinforcement Learning via Cross-Model Entropy
di: Gorbett, Matt, et al.
Pubblicazione: (2026)
di: Gorbett, Matt, et al.
Pubblicazione: (2026)
A Free Probabilistic Framework for Denoising Diffusion Models: Entropy, Transport, and Reverse Processes
di: Das, Swagatam
Pubblicazione: (2025)
di: Das, Swagatam
Pubblicazione: (2025)
Entropy Regularizing Activation: Boosting Continuous Control, Large Language Models, and Image Classification with Activation as Entropy Constraints
di: Kang, Zilin, et al.
Pubblicazione: (2025)
di: Kang, Zilin, et al.
Pubblicazione: (2025)
Float8@2bits: Entropy Coding Enables Data-Free Model Compression
di: Putzky, Patrick, et al.
Pubblicazione: (2026)
di: Putzky, Patrick, et al.
Pubblicazione: (2026)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
di: Carr, Jonathan Colaço, et al.
Pubblicazione: (2023)
di: Carr, Jonathan Colaço, et al.
Pubblicazione: (2023)
Neural Stochastic Flows: Solver-Free Modelling and Inference for SDE Solutions
di: Kiyohara, Naoki, et al.
Pubblicazione: (2025)
di: Kiyohara, Naoki, et al.
Pubblicazione: (2025)
Aligning Language Models with Investor and Market Behavior for Financial Recommendations
di: Spadea, Fernando, et al.
Pubblicazione: (2025)
di: Spadea, Fernando, et al.
Pubblicazione: (2025)
From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs
di: Gu, Xihe, et al.
Pubblicazione: (2026)
di: Gu, Xihe, et al.
Pubblicazione: (2026)
Quantum Maximum Entropy Inference and Hamiltonian Learning
di: Gao, Minbo, et al.
Pubblicazione: (2024)
di: Gao, Minbo, et al.
Pubblicazione: (2024)
Optimizing Language Models for Human Preferences is a Causal Inference Problem
di: Lin, Victoria, et al.
Pubblicazione: (2024)
di: Lin, Victoria, et al.
Pubblicazione: (2024)
Inference-Time Personalized Alignment with a Few User Preference Queries
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Maximum Causal Entropy IRL in Mean-Field Games and GNEP Framework for Forward RL
di: Anahtarci, Berkay, et al.
Pubblicazione: (2024) -
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning
di: Wan, Zhenglin, et al.
Pubblicazione: (2025) -
Scouting By Reward: VLM-TO-IRL-Driven Player Selection For Esports
di: Yan, Qing, et al.
Pubblicazione: (2026) -
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
di: Jain, Gauri, et al.
Pubblicazione: (2024) -
CoMI-IRL: Contrastive Multi-Intention Inverse Reinforcement Learning
di: Mone, Antonio, et al.
Pubblicazione: (2026)