Gespeichert in:
| Hauptverfasser: | Bicer, Osman, Kara, Ali D., Yuksel, Serdar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.04355 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Near Optimal Approximations and Finite Memory Policies for POMPDs with Continuous Spaces
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024)
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024)
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024)
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024)
Robustness to Model Approximation, Model Learning From Data, and Sample Complexity in Wasserstein Regular MDPs
von: Zhou, Yichen, et al.
Veröffentlicht: (2024)
von: Zhou, Yichen, et al.
Veröffentlicht: (2024)
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
von: Kara, Ali Devran, et al.
Veröffentlicht: (2023)
von: Kara, Ali Devran, et al.
Veröffentlicht: (2023)
Kernel Mean Embedding Topology: Weak and Strong Forms for Stochastic Kernels and Implications for Model Learning
von: Saldi, Naci, et al.
Veröffentlicht: (2025)
von: Saldi, Naci, et al.
Veröffentlicht: (2025)
Refined Bounds on Near Optimality Finite Window Policies in POMDPs and Their Reinforcement Learning
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2024)
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2024)
Data-Driven Non-Parametric Model Learning and Adaptive Control of MDPs with Borel spaces: Identifiability and Near Optimal Design
von: Mrani-Zentar, Omar, et al.
Veröffentlicht: (2025)
von: Mrani-Zentar, Omar, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Function Approximation for Non-Markov Processes
von: Kara, Ali Devran
Veröffentlicht: (2026)
von: Kara, Ali Devran
Veröffentlicht: (2026)
An Optimal Control Approach To Transformer Training
von: Akman, Kağan, et al.
Veröffentlicht: (2026)
von: Akman, Kağan, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Discounted and Ergodic Control of Diffusion Processes
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
Satisficing Paths and Independent Multi-Agent Reinforcement Learning in Stochastic Games
von: Yongacoglu, Bora, et al.
Veröffentlicht: (2021)
von: Yongacoglu, Bora, et al.
Veröffentlicht: (2021)
Layer-wise Quantization for Quantized Optimistic Dual Averaging
von: Nguyen, Anh Duc, et al.
Veröffentlicht: (2025)
von: Nguyen, Anh Duc, et al.
Veröffentlicht: (2025)
Online Distributed Learning with Quantized Finite-Time Coordination
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
Another Look at Partially Observed Optimal Stochastic Control: Existence, Ergodicity, and Approximations without Belief-Reduction
von: Yüksel, Serdar
Veröffentlicht: (2023)
von: Yüksel, Serdar
Veröffentlicht: (2023)
Q-Learning under Finite Model Uncertainty
von: Sester, Julian, et al.
Veröffentlicht: (2024)
von: Sester, Julian, et al.
Veröffentlicht: (2024)
WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points
von: Li, Dongyue, et al.
Veröffentlicht: (2026)
von: Li, Dongyue, et al.
Veröffentlicht: (2026)
Learning POMDPs with Linear Function Approximation and Finite Memory
von: Kara, Ali Devran
Veröffentlicht: (2025)
von: Kara, Ali Devran
Veröffentlicht: (2025)
Planning and Learning in Average Risk-aware MDPs
von: Wang, Weikai, et al.
Veröffentlicht: (2025)
von: Wang, Weikai, et al.
Veröffentlicht: (2025)
Landscape of Policy Optimization for Finite Horizon MDPs with General State and Action
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Offline-Online Reinforcement Learning for Linear Mixture MDPs
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
Joint MDPs and Reinforcement Learning in Coupled-Dynamics Environments
von: Kaya, Ege C., et al.
Veröffentlicht: (2026)
von: Kaya, Ege C., et al.
Veröffentlicht: (2026)
Sensitivity of Filter Kernels and Robustness Bounds to Transition and Measurement Kernel Perturbations in Partially Observable Stochastic Control
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2025)
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2025)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
Model approximation in MDPs with unbounded per-step cost
von: Bozkurt, Berk, et al.
Veröffentlicht: (2024)
von: Bozkurt, Berk, et al.
Veröffentlicht: (2024)
Efficient Model-Free Exploration in Low-Rank MDPs
von: Mhammedi, Zakaria, et al.
Veröffentlicht: (2023)
von: Mhammedi, Zakaria, et al.
Veröffentlicht: (2023)
Approximations and Learning for Decentralized Stochastic Control and Near Optimal Finite Window Policies
von: Mrani-Zentar, Omar, et al.
Veröffentlicht: (2026)
von: Mrani-Zentar, Omar, et al.
Veröffentlicht: (2026)
Sliding Window Codes: Near-Optimality and Q-Learning for Zero-Delay Coding
von: Cregg, Liam, et al.
Veröffentlicht: (2023)
von: Cregg, Liam, et al.
Veröffentlicht: (2023)
PARQ: Piecewise-Affine Regularized Quantization
von: Jin, Lisa, et al.
Veröffentlicht: (2025)
von: Jin, Lisa, et al.
Veröffentlicht: (2025)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
von: Mitra, Aritra
Veröffentlicht: (2024)
von: Mitra, Aritra
Veröffentlicht: (2024)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
Incremental Learning of Sparse Attention Patterns in Transformers
von: Yüksel, Oğuz Kaan, et al.
Veröffentlicht: (2026)
von: Yüksel, Oğuz Kaan, et al.
Veröffentlicht: (2026)
Quantization Avoids Saddle Points in Distributed Optimization
von: Bo, Yanan, et al.
Veröffentlicht: (2024)
von: Bo, Yanan, et al.
Veröffentlicht: (2024)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
von: Zhang, Runyu, et al.
Veröffentlicht: (2023)
von: Zhang, Runyu, et al.
Veröffentlicht: (2023)
Representative Action Selection for Large Action Space: From Bandits to MDPs
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2025)
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2025)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
Decentralized Optimization on Compact Submanifolds by Quantized Riemannian Gradient Tracking
von: Chen, Jun, et al.
Veröffentlicht: (2025)
von: Chen, Jun, et al.
Veröffentlicht: (2025)
On Borkar and Young Relaxed Control Topologies and Continuous Dependence of Invariant Measures on Control Policy
von: Yüksel, Serdar
Veröffentlicht: (2023)
von: Yüksel, Serdar
Veröffentlicht: (2023)
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Ähnliche Einträge
-
Near Optimal Approximations and Finite Memory Policies for POMPDs with Continuous Spaces
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024) -
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
von: Kara, Ali Devran, et al.
Veröffentlicht: (2024) -
Robustness to Model Approximation, Model Learning From Data, and Sample Complexity in Wasserstein Regular MDPs
von: Zhou, Yichen, et al.
Veröffentlicht: (2024) -
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
von: Kara, Ali Devran, et al.
Veröffentlicht: (2023) -
Kernel Mean Embedding Topology: Weak and Strong Forms for Stochastic Kernels and Implications for Model Learning
von: Saldi, Naci, et al.
Veröffentlicht: (2025)