Guardado en:
| Autores principales: | Abolhassani, Bahman, Tadrous, John, Eryilmaz, Atilla, Yüksel, Serdar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2401.03613 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SwiftCache: Model-Based Learning for Dynamic Content Caching in CDNs
por: Abolhassani, Bahman, et al.
Publicado: (2024)
por: Abolhassani, Bahman, et al.
Publicado: (2024)
Another Look at Partially Observed Optimal Stochastic Control: Existence, Ergodicity, and Approximations without Belief-Reduction
por: Yüksel, Serdar
Publicado: (2023)
por: Yüksel, Serdar
Publicado: (2023)
Convergence of Gradient Descent for Recurrent Neural Networks: A Nonasymptotic Analysis
por: Cayci, Semih, et al.
Publicado: (2024)
por: Cayci, Semih, et al.
Publicado: (2024)
Recurrent Natural Policy Gradient for POMDPs
por: Cayci, Semih, et al.
Publicado: (2024)
por: Cayci, Semih, et al.
Publicado: (2024)
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
por: Kara, Ali Devran, et al.
Publicado: (2024)
por: Kara, Ali Devran, et al.
Publicado: (2024)
Decentralized Learning for Optimality in Stochastic Dynamic Teams and Games with Local Control and Global State Information
por: Yongacoglu, Bora, et al.
Publicado: (2019)
por: Yongacoglu, Bora, et al.
Publicado: (2019)
Decentralized Exchangeable Stochastic Dynamic Teams in Continuous-time, their Mean-Field Limits and Optimality of Symmetric Policies
por: Sanjari, Sina, et al.
Publicado: (2024)
por: Sanjari, Sina, et al.
Publicado: (2024)
Quantum Markov Decision Processes: Dynamic and Semi-Definite Programs for Optimal Solutions
por: Saldi, Naci, et al.
Publicado: (2024)
por: Saldi, Naci, et al.
Publicado: (2024)
On Borkar and Young Relaxed Control Topologies and Continuous Dependence of Invariant Measures on Control Policy
por: Yüksel, Serdar
Publicado: (2023)
por: Yüksel, Serdar
Publicado: (2023)
Finite-Time Analysis of Gradient Descent for Shallow Transformers
por: Arda, Enes, et al.
Publicado: (2026)
por: Arda, Enes, et al.
Publicado: (2026)
Near Optimal Approximations and Finite Memory Policies for POMPDs with Continuous Spaces
por: Kara, Ali Devran, et al.
Publicado: (2024)
por: Kara, Ali Devran, et al.
Publicado: (2024)
Data-Driven Non-Parametric Model Learning and Adaptive Control of MDPs with Borel spaces: Identifiability and Near Optimal Design
por: Mrani-Zentar, Omar, et al.
Publicado: (2025)
por: Mrani-Zentar, Omar, et al.
Publicado: (2025)
Reinforcement Learning for Jointly Optimal Coding and Control Policies for a Controlled Markovian System over a Communication Channel
por: Hubbard, Evelyn, et al.
Publicado: (2024)
por: Hubbard, Evelyn, et al.
Publicado: (2024)
An Optimal Control Approach To Transformer Training
por: Akman, Kağan, et al.
Publicado: (2026)
por: Akman, Kağan, et al.
Publicado: (2026)
SLAM as a Stochastic Control Problem with Partial Information: Optimal Solutions and Rigorous Approximations
por: Gusija, Ilir, et al.
Publicado: (2026)
por: Gusija, Ilir, et al.
Publicado: (2026)
Optimality of Symmetric Independent Policies under Decentralized Mean-Field Information Sharing for Stochastic Teams and Equivalence with McKean-Vlasov Control of a Representative Agent
por: Sanjari, Sina, et al.
Publicado: (2024)
por: Sanjari, Sina, et al.
Publicado: (2024)
Subjective Equilibria under Beliefs of Exogenous Uncertainty: Linear Quadratic Case
por: Arslan, Gürdal, et al.
Publicado: (2024)
por: Arslan, Gürdal, et al.
Publicado: (2024)
Decentralized Detection with Many Sensors: Optimality of Exchangeable and Identical Encoding Policies
por: Sanjari, Sina, et al.
Publicado: (2025)
por: Sanjari, Sina, et al.
Publicado: (2025)
Reinforcement Learning for Near-Optimal Design of Zero-Delay Codes for Markov Sources
por: Cregg, Liam, et al.
Publicado: (2023)
por: Cregg, Liam, et al.
Publicado: (2023)
Sliding Window Codes: Near-Optimality and Q-Learning for Zero-Delay Coding
por: Cregg, Liam, et al.
Publicado: (2023)
por: Cregg, Liam, et al.
Publicado: (2023)
Refined Bounds on Near Optimality Finite Window Policies in POMDPs and Their Reinforcement Learning
por: Demirci, Yunus Emre, et al.
Publicado: (2024)
por: Demirci, Yunus Emre, et al.
Publicado: (2024)
Robust Decentralized Control of Coupled Systems via Risk Sensitive Control of Decoupled or Simple Models with Measure Change
por: Selk, Zachary, et al.
Publicado: (2024)
por: Selk, Zachary, et al.
Publicado: (2024)
Stochastic Push-Pull for Decentralized Nonconvex Optimization
por: You, Runze, et al.
Publicado: (2025)
por: You, Runze, et al.
Publicado: (2025)
Robustness to Model Approximation, Model Learning From Data, and Sample Complexity in Wasserstein Regular MDPs
por: Zhou, Yichen, et al.
Publicado: (2024)
por: Zhou, Yichen, et al.
Publicado: (2024)
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
por: Kara, Ali Devran, et al.
Publicado: (2023)
por: Kara, Ali Devran, et al.
Publicado: (2023)
Kernel Mean Embedding Topology: Weak and Strong Forms for Stochastic Kernels and Implications for Model Learning
por: Saldi, Naci, et al.
Publicado: (2025)
por: Saldi, Naci, et al.
Publicado: (2025)
Quantum Markov Decision Processes: General Theory, Approximations, and Classes of Policies
por: Saldi, Naci, et al.
Publicado: (2024)
por: Saldi, Naci, et al.
Publicado: (2024)
Near Optimality of Lipschitz and Smooth Policies in Controlled Diffusions
por: Pradhan, Somnath, et al.
Publicado: (2024)
por: Pradhan, Somnath, et al.
Publicado: (2024)
Non-Sequential Decentralized Stochastic Control Revisited: Causality and Static Reducibility
por: Mrani-Zentar, Omar, et al.
Publicado: (2022)
por: Mrani-Zentar, Omar, et al.
Publicado: (2022)
A Robust Compressed Push-Pull Method for Decentralized Nonconvex Optimization
por: Liao, Yiwei, et al.
Publicado: (2024)
por: Liao, Yiwei, et al.
Publicado: (2024)
On the Linear Speedup of the Push-Pull Method for Decentralized Optimization over Digraphs
por: Liang, Liyuan, et al.
Publicado: (2025)
por: Liang, Liyuan, et al.
Publicado: (2025)
Mean-Field Systems with Heterogeneous Subteams: Optimality of Cluster-Symmetric Independent Policies and Equivalence with Decentralized McKean-Vlasov Control of Cluster-Representative Agents
por: Braun, Connor S., et al.
Publicado: (2026)
por: Braun, Connor S., et al.
Publicado: (2026)
Sensitivity of Filter Kernels and Robustness Bounds to Transition and Measurement Kernel Perturbations in Partially Observable Stochastic Control
por: Demirci, Yunus Emre, et al.
Publicado: (2025)
por: Demirci, Yunus Emre, et al.
Publicado: (2025)
Stochastic Momentum Tracking Push-Pull for Decentralized Optimization over Directed Graphs
por: Fan, Wenqi, et al.
Publicado: (2026)
por: Fan, Wenqi, et al.
Publicado: (2026)
Approximations and Learning for Decentralized Stochastic Control and Near Optimal Finite Window Policies
por: Mrani-Zentar, Omar, et al.
Publicado: (2026)
por: Mrani-Zentar, Omar, et al.
Publicado: (2026)
Discrete-Time Approximations of Controlled Diffusions with Infinite Horizon Discounted and Average Cost
por: Pradhan, Somnath, et al.
Publicado: (2025)
por: Pradhan, Somnath, et al.
Publicado: (2025)
Near Optimality of Discrete-Time Approximations for Controlled McKean-Vlasov Diffusions and Interacting Particle Systems
por: Pradhan, Somnath, et al.
Publicado: (2025)
por: Pradhan, Somnath, et al.
Publicado: (2025)
Reinforcement Learning for Discounted and Ergodic Control of Diffusion Processes
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Quantizer Design for Finite Model Approximations, Model Learning, and Quantized Q-Learning for MDPs with Unbounded Spaces
por: Bicer, Osman, et al.
Publicado: (2025)
por: Bicer, Osman, et al.
Publicado: (2025)
B-ary Tree Push-Pull Method is Provably Efficient for Distributed Learning on Heterogeneous Data
por: You, Runze, et al.
Publicado: (2024)
por: You, Runze, et al.
Publicado: (2024)
Ejemplares similares
-
SwiftCache: Model-Based Learning for Dynamic Content Caching in CDNs
por: Abolhassani, Bahman, et al.
Publicado: (2024) -
Another Look at Partially Observed Optimal Stochastic Control: Existence, Ergodicity, and Approximations without Belief-Reduction
por: Yüksel, Serdar
Publicado: (2023) -
Convergence of Gradient Descent for Recurrent Neural Networks: A Nonasymptotic Analysis
por: Cayci, Semih, et al.
Publicado: (2024) -
Recurrent Natural Policy Gradient for POMDPs
por: Cayci, Semih, et al.
Publicado: (2024) -
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
por: Kara, Ali Devran, et al.
Publicado: (2024)