Inferring Transition Dynamics from Value Functions
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Adamczyk, Jacob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploration Behavior of Untrained Policies
von: Adamczyk, Jacob
Veröffentlicht: (2025)
von: Adamczyk, Jacob
Veröffentlicht: (2025)
Maximum Entropy Exploration Without the Rollouts
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026)
Thermodynamics of Reinforcement Learning Curricula
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026)
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
EVAL: EigenVector-based Average-reward Learning
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Average-Reward Soft Actor-Critic
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Boosting Soft Q-Learning by Bounding
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2024)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2024)
Active Inference with Reusable State-Dependent Value Profiles
von: Poschl, Jacob
Veröffentlicht: (2025)
von: Poschl, Jacob
Veröffentlicht: (2025)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
von: Li, Zongyue, et al.
Veröffentlicht: (2025)
von: Li, Zongyue, et al.
Veröffentlicht: (2025)
Evaluating machine learning models for predicting pesticide toxicity to honey bees
von: Adamczyk, Jakub, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jakub, et al.
Veröffentlicht: (2025)
Benchmarking Pretrained Molecular Embedding Models For Molecular Representation Learning
von: Praski, Mateusz, et al.
Veröffentlicht: (2025)
von: Praski, Mateusz, et al.
Veröffentlicht: (2025)
Inferring Reward Machines and Transition Machines from Partially Observable Markov Decision Processes
von: Wu, Yuly, et al.
Veröffentlicht: (2025)
von: Wu, Yuly, et al.
Veröffentlicht: (2025)
Transitive RL: Value Learning via Divide and Conquer
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Universal Value-Function Uncertainties
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
Positive-Unlabeled Constraint Learning for Inferring Nonlinear Continuous Constraints Functions from Expert Demonstrations
von: Peng, Baiyu, et al.
Veröffentlicht: (2024)
von: Peng, Baiyu, et al.
Veröffentlicht: (2024)
Quasimetric Value Functions with Dense Rewards
von: Valieva, Khadichabonu, et al.
Veröffentlicht: (2024)
von: Valieva, Khadichabonu, et al.
Veröffentlicht: (2024)
Inferring response times of perceptual decisions with Poisson variational autoencoders
von: Johnson, Hayden R., et al.
Veröffentlicht: (2025)
von: Johnson, Hayden R., et al.
Veröffentlicht: (2025)
Executable Functional Abstractions: Inferring Generative Programs for Advanced Math Problems
von: Khan, Zaid, et al.
Veröffentlicht: (2025)
von: Khan, Zaid, et al.
Veröffentlicht: (2025)
Massively Scaling Explicit Policy-conditioned Value Functions
von: Bohlinger, Nico, et al.
Veröffentlicht: (2025)
von: Bohlinger, Nico, et al.
Veröffentlicht: (2025)
Learning Exposure Mapping Functions for Inferring Heterogeneous Peer Effects
von: Adhikari, Shishir, et al.
Veröffentlicht: (2025)
von: Adhikari, Shishir, et al.
Veröffentlicht: (2025)
Catapult Dynamics and Phase Transitions in Quadratic Nets
von: Meltzer, David, et al.
Veröffentlicht: (2023)
von: Meltzer, David, et al.
Veröffentlicht: (2023)
VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
Stable Offline Value Function Learning with Bisimulation-based Representations
von: Pavse, Brahma S., et al.
Veröffentlicht: (2024)
von: Pavse, Brahma S., et al.
Veröffentlicht: (2024)
Tensor Low-rank Approximation of Finite-horizon Value Functions
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
LOCAL: Learning with Orientation Matrix to Infer Causal Structure from Time Series Data
von: Zhang, Jiajun, et al.
Veröffentlicht: (2024)
von: Zhang, Jiajun, et al.
Veröffentlicht: (2024)
A Poisson-Gamma Dynamic Factor Model with Time-Varying Transition Dynamics
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
Bayesian Optimization for Function-Valued Responses under Min-Max Criteria
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2022)
von: Rozada, Sergio, et al.
Veröffentlicht: (2022)
On the Limited Representational Power of Value Functions and its Links to Statistical (In)Efficiency
von: Cheikhi, David, et al.
Veröffentlicht: (2024)
von: Cheikhi, David, et al.
Veröffentlicht: (2024)
Is Value Functions Estimation with Classification Plug-and-play for Offline Reinforcement Learning?
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
ViVa: Video-Trained Value Functions for Guiding Online RL from Diverse Data
von: Dashora, Nitish, et al.
Veröffentlicht: (2025)
von: Dashora, Nitish, et al.
Veröffentlicht: (2025)
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
von: Nair, Jishnu Sethumadhavan, et al.
Veröffentlicht: (2026)
von: Nair, Jishnu Sethumadhavan, et al.
Veröffentlicht: (2026)
Stop Regressing: Training Value Functions via Classification for Scalable Deep RL
von: Farebrother, Jesse, et al.
Veröffentlicht: (2024)
von: Farebrother, Jesse, et al.
Veröffentlicht: (2024)
On the Curses of Future and History in Future-dependent Value Functions for Off-policy Evaluation
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Sample and Oracle Efficient Reinforcement Learning for MDPs with Linearly-Realizable Value Functions
von: Mhammedi, Zakaria
Veröffentlicht: (2024)
von: Mhammedi, Zakaria
Veröffentlicht: (2024)
Posterior-First Neural PDE Simulation: Inferring Hidden Problem State from a Single Field
von: Wang, Wenshuo, et al.
Veröffentlicht: (2026)
von: Wang, Wenshuo, et al.
Veröffentlicht: (2026)
AirRadar: Inferring Nationwide Air Quality in China with Deep Neural Networks
von: Wang, Qiongyan, et al.
Veröffentlicht: (2025)
von: Wang, Qiongyan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploration Behavior of Untrained Policies
von: Adamczyk, Jacob
Veröffentlicht: (2025) -
Maximum Entropy Exploration Without the Rollouts
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026) -
Thermodynamics of Reinforcement Learning Curricula
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2026) -
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025) -
EVAL: EigenVector-based Average-reward Learning
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)