Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Ruining, Han, Haoran, Lv, Maolong, Yang, Qisong, Cheng, Jian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On-Meter Graph Machine Learning: A Case Study of PV Power Forecasting for Grid Edge Intelligence
por: Huang, Jian, et al.
Publicado: (2026)
por: Huang, Jian, et al.
Publicado: (2026)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
por: Ibrahim, Sinan, et al.
Publicado: (2026)
por: Ibrahim, Sinan, et al.
Publicado: (2026)
Specification Generation for Neural Networks in Systems
por: Chaudhary, Isha, et al.
Publicado: (2024)
por: Chaudhary, Isha, et al.
Publicado: (2024)
Integrating LLMs for Explainable Fault Diagnosis in Complex Systems
por: Dave, Akshay J., et al.
Publicado: (2024)
por: Dave, Akshay J., et al.
Publicado: (2024)
Physically Plausible Multi-System Trajectory Generation and Symmetry Discovery
por: Liu, Jiayin, et al.
Publicado: (2025)
por: Liu, Jiayin, et al.
Publicado: (2025)
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial
por: Li, Haoyu, et al.
Publicado: (2026)
por: Li, Haoyu, et al.
Publicado: (2026)
DCoPilot: Generative AI-Empowered Policy Adaptation for Dynamic Data Center Operations
por: Li, Minghao, et al.
Publicado: (2026)
por: Li, Minghao, et al.
Publicado: (2026)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
Spatiotemporal Decision Transformer for Traffic Coordination
por: Su, Haoran, et al.
Publicado: (2026)
por: Su, Haoran, et al.
Publicado: (2026)
Continual Reinforcement Learning for HVAC Systems Control: Integrating Hypernetworks and Transfer Learning
por: Bekal, Gautham Udayakumar, et al.
Publicado: (2025)
por: Bekal, Gautham Udayakumar, et al.
Publicado: (2025)
Distributionally Robust Safety Verification of Neural Networks via Worst-Case CVaR
por: Kishida, Masako
Publicado: (2025)
por: Kishida, Masako
Publicado: (2025)
A Neural Network-Based Real-time Casing Collar Recognition System for Downhole Instruments
por: Xiao, Si-Yu, et al.
Publicado: (2025)
por: Xiao, Si-Yu, et al.
Publicado: (2025)
Towards Developing Safety Assurance Cases for Learning-Enabled Medical Cyber-Physical Systems
por: Bagheri, Maryam, et al.
Publicado: (2022)
por: Bagheri, Maryam, et al.
Publicado: (2022)
Generative AI and Process Systems Engineering: The Next Frontier
por: Decardi-Nelson, Benjamin, et al.
Publicado: (2024)
por: Decardi-Nelson, Benjamin, et al.
Publicado: (2024)
Improving Variational Autoencoder using Random Fourier Transformation: An Aviation Safety Anomaly Detection Case-Study
por: Asanjan, Ata Akbari, et al.
Publicado: (2026)
por: Asanjan, Ata Akbari, et al.
Publicado: (2026)
Decision Transformer as a Foundation Model for Partially Observable Continuous Control
por: Zhang, Xiangyuan, et al.
Publicado: (2024)
por: Zhang, Xiangyuan, et al.
Publicado: (2024)
Generative Modeling and Data Augmentation for Power System Production Simulation
por: Xu, Linna, et al.
Publicado: (2024)
por: Xu, Linna, et al.
Publicado: (2024)
Scaling Learning based Policy Optimization for Temporal Logic Tasks by Controller Network Dropout
por: Hashemi, Navid, et al.
Publicado: (2024)
por: Hashemi, Navid, et al.
Publicado: (2024)
Deep Policy Gradient Methods Without Batch Updates, Target Networks, or Replay Buffers
por: Vasan, Gautham, et al.
Publicado: (2024)
por: Vasan, Gautham, et al.
Publicado: (2024)
Scalable Data-Driven Reachability Analysis and Control via Koopman Operators with Conformal Coverage Guarantees
por: Nath, Devesh, et al.
Publicado: (2026)
por: Nath, Devesh, et al.
Publicado: (2026)
PowerGrow: Feasible Co-Growth of Structures and Dynamics for Power Grid Synthesis
por: He, Xinyu, et al.
Publicado: (2025)
por: He, Xinyu, et al.
Publicado: (2025)
Graph Embedding Dynamic Feature-based Supervised Contrastive Learning of Transient Stability for Changing Power Grid Topologies
por: Lv, Zijian, et al.
Publicado: (2023)
por: Lv, Zijian, et al.
Publicado: (2023)
Learning a Generalized Model for Substation Level Voltage Estimation in Distribution Networks
por: Za'ter, Muhy Eddin, et al.
Publicado: (2025)
por: Za'ter, Muhy Eddin, et al.
Publicado: (2025)
An Optimal Policy for Learning Controllable Dynamics by Exploration
por: Loxley, Peter N.
Publicado: (2025)
por: Loxley, Peter N.
Publicado: (2025)
Policy Optimization Algorithms in a Unified Framework
por: Wu, Shuang
Publicado: (2025)
por: Wu, Shuang
Publicado: (2025)
Certifiably Robust Policies for Uncertain Parametric Environments
por: Schnitzer, Yannik, et al.
Publicado: (2024)
por: Schnitzer, Yannik, et al.
Publicado: (2024)
Stable-by-Design Neural Network-Based LPV State-Space Models for System Identification
por: Sertbaş, Ahmet Eren, et al.
Publicado: (2025)
por: Sertbaş, Ahmet Eren, et al.
Publicado: (2025)
Vegetable Peeling: A Case Study in Constrained Dexterous Manipulation
por: Chen, Tao, et al.
Publicado: (2024)
por: Chen, Tao, et al.
Publicado: (2024)
FaultXformer: A Transformer-Encoder Based Fault Classification and Location Identification model in PMU-Integrated Active Electrical Distribution System
por: Thakur, Kriti, et al.
Publicado: (2026)
por: Thakur, Kriti, et al.
Publicado: (2026)
Feature Engineering Approach to Building Load Prediction: A Case Study for Commercial Building Chiller Plant Optimization in Tropical Weather
por: Wang, Zhan, et al.
Publicado: (2025)
por: Wang, Zhan, et al.
Publicado: (2025)
GLinSAT: The General Linear Satisfiability Neural Network Layer By Accelerated Gradient Descent
por: Zeng, Hongtai, et al.
Publicado: (2024)
por: Zeng, Hongtai, et al.
Publicado: (2024)
On the Generalization of Data-Assisted Control in port-Hamiltonian Systems (DAC-pH)
por: Eslami, Mostafa, et al.
Publicado: (2025)
por: Eslami, Mostafa, et al.
Publicado: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
por: Foffano, Daniele, et al.
Publicado: (2023)
por: Foffano, Daniele, et al.
Publicado: (2023)
Stabilizing Policy Gradient Methods via Reward Profiling
por: Ahmed, Shihab, et al.
Publicado: (2025)
por: Ahmed, Shihab, et al.
Publicado: (2025)
Time and Frequency Domain-based Anomaly Detection in Smart Meter Data for Distribution Network Studies
por: Labura, Petar, et al.
Publicado: (2025)
por: Labura, Petar, et al.
Publicado: (2025)
Trustworthy and Explainable Deep Reinforcement Learning for Safe and Energy-Efficient Process Control: A Use Case in Industrial Compressed Air Systems
por: Bezold, Vincent, et al.
Publicado: (2025)
por: Bezold, Vincent, et al.
Publicado: (2025)
Decentralized Dynamic Event-triggered Output-feedback Control of Stochastic Non-triangular Interconnected Systems with Unknown Time-varying Sensor Sensitivity
por: Sun, Libei, et al.
Publicado: (2024)
por: Sun, Libei, et al.
Publicado: (2024)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
por: Anand, Akhil S, et al.
Publicado: (2025)
por: Anand, Akhil S, et al.
Publicado: (2025)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
por: Fard, Amir, et al.
Publicado: (2025)
por: Fard, Amir, et al.
Publicado: (2025)
An LLM-Based Digital Twin for Optimizing Human-in-the Loop Systems
por: Yang, Hanqing, et al.
Publicado: (2024)
por: Yang, Hanqing, et al.
Publicado: (2024)
Ejemplares similares
-
On-Meter Graph Machine Learning: A Case Study of PV Power Forecasting for Grid Edge Intelligence
por: Huang, Jian, et al.
Publicado: (2026) -
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
por: Ibrahim, Sinan, et al.
Publicado: (2026) -
Specification Generation for Neural Networks in Systems
por: Chaudhary, Isha, et al.
Publicado: (2024) -
Integrating LLMs for Explainable Fault Diagnosis in Complex Systems
por: Dave, Akshay J., et al.
Publicado: (2024) -
Physically Plausible Multi-System Trajectory Generation and Symmetry Discovery
por: Liu, Jiayin, et al.
Publicado: (2025)