How to ensure a safe control strategy? Towards a SRL for urban transit autonomous operation
Fuente:
arXiv
Saved in:
| Main Author: | Zhao, Zicong |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DAHRS: Divergence-Aware Hallucination-Remediated SRL Projection
by: Youm, Sangpil, et al.
Published: (2024)
by: Youm, Sangpil, et al.
Published: (2024)
Survey on safe robot control via learning
by: Mabsout, Bassel El
Published: (2024)
by: Mabsout, Bassel El
Published: (2024)
How to safely discard features based on aggregate SHAP values
by: Bhattacharjee, Robi, et al.
Published: (2025)
by: Bhattacharjee, Robi, et al.
Published: (2025)
SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores
by: Mei, Zhiyu, et al.
Published: (2023)
by: Mei, Zhiyu, et al.
Published: (2023)
xSRL: Safety-Aware Explainable Reinforcement Learning -- Safety as a Product of Explainability
by: Shefin, Risal Shahriar, et al.
Published: (2024)
by: Shefin, Risal Shahriar, et al.
Published: (2024)
Pedestrian motion prediction evaluation for urban autonomous driving
by: Zabolotnii, Dmytro, et al.
Published: (2024)
by: Zabolotnii, Dmytro, et al.
Published: (2024)
Critic as Lyapunov function (CALF): a model-free, stability-ensuring agent
by: Osinenko, Pavel, et al.
Published: (2024)
by: Osinenko, Pavel, et al.
Published: (2024)
Can AI autonomously build, operate, and use the entire data stack?
by: Agarwal, Arvind, et al.
Published: (2025)
by: Agarwal, Arvind, et al.
Published: (2025)
Hierarchical learning control for autonomous robots inspired by central nervous system
by: Zhang, Pei, et al.
Published: (2024)
by: Zhang, Pei, et al.
Published: (2024)
What Shapes a Creative Machine Mind? Comprehensively Benchmarking Creativity in Foundation Models
by: He, Zicong, et al.
Published: (2025)
by: He, Zicong, et al.
Published: (2025)
End-to-end autonomous scientific discovery on a real optical platform
by: Yang, Shuxing, et al.
Published: (2026)
by: Yang, Shuxing, et al.
Published: (2026)
A multi-algorithm approach for operational human resources workload balancing in a last mile urban delivery system
by: Moreno-Saavedra, Luis M., et al.
Published: (2025)
by: Moreno-Saavedra, Luis M., et al.
Published: (2025)
Re2: A Consistency-ensured Dataset for Full-stage Peer Review and Multi-turn Rebuttal Discussions
by: Zhang, Daoze, et al.
Published: (2025)
by: Zhang, Daoze, et al.
Published: (2025)
Context is all you need: Towards autonomous model-based process design using agentic AI in flowsheet simulations
by: Schäfer, Pascal, et al.
Published: (2026)
by: Schäfer, Pascal, et al.
Published: (2026)
An AI-native experimental laboratory for autonomous biomolecular engineering
by: Wu, Mingyu, et al.
Published: (2025)
by: Wu, Mingyu, et al.
Published: (2025)
D$^{2}$MoE: Dual Routing and Dynamic Scheduling for Efficient On-Device MoE-based LLM Serving
by: Wang, Haodong, et al.
Published: (2025)
by: Wang, Haodong, et al.
Published: (2025)
Supertrust foundational alignment: mutual trust must replace permanent control for safe superintelligence
by: Mazzu, James M.
Published: (2024)
by: Mazzu, James M.
Published: (2024)
TimeSRL: Generalizable Time-Series Behavioral Modeling via Semantic RL-Tuned LLMs -- A Case Study in Mental Health
by: Fan, Yuang, et al.
Published: (2026)
by: Fan, Yuang, et al.
Published: (2026)
WMPO: World Model-based Policy Optimization for Vision-Language-Action Models
by: Zhu, Fangqi, et al.
Published: (2025)
by: Zhu, Fangqi, et al.
Published: (2025)
A hierarchical control framework for autonomous decision-making systems: Integrating HMDP and MPC
by: Wang, Xue-Fang, et al.
Published: (2024)
by: Wang, Xue-Fang, et al.
Published: (2024)
Navigating the safe harbor paradox in human-machine systems
by: Zanardelli, Riccardo
Published: (2025)
by: Zanardelli, Riccardo
Published: (2025)
An autonomous agent for auditing and improving the reliability of clinical AI models
by: Kuhn, Lukas, et al.
Published: (2025)
by: Kuhn, Lukas, et al.
Published: (2025)
Towards autonomous quantum physics research using LLM agents with access to intelligent tools
by: Arlt, Sören, et al.
Published: (2025)
by: Arlt, Sören, et al.
Published: (2025)
Krul: Efficient State Restoration for Multi-turn Conversations with Dynamic Cross-layer KV Sharing
by: Wen, Junyi, et al.
Published: (2025)
by: Wen, Junyi, et al.
Published: (2025)
AhaKV: Adaptive Holistic Attention-Driven KV Cache Eviction for Efficient Inference of Large Language Models
by: Gu, Yifeng, et al.
Published: (2025)
by: Gu, Yifeng, et al.
Published: (2025)
First, do NOHARM: towards clinically safe large language models
by: Wu, David, et al.
Published: (2025)
by: Wu, David, et al.
Published: (2025)
Automated urban waterlogging assessment and early warning through a mixture of foundation models
by: Zhang, Chenxu, et al.
Published: (2025)
by: Zhang, Chenxu, et al.
Published: (2025)
A methodological framework for Resilience as a Service (RaaS) in multimodal urban transportation networks
by: Jaber, Sara, et al.
Published: (2024)
by: Jaber, Sara, et al.
Published: (2024)
Monte Carlo Tree Search with Velocity Obstacles for safe and efficient motion planning in dynamic environments
by: Bonanni, Lorenzo, et al.
Published: (2025)
by: Bonanni, Lorenzo, et al.
Published: (2025)
Contingency-constrained economic dispatch with safe reinforcement learning
by: Eichelbeck, Michael, et al.
Published: (2022)
by: Eichelbeck, Michael, et al.
Published: (2022)
Aligned, Orthogonal or In-conflict: When can we safely optimize Chain-of-Thought?
by: Kaufmann, Max, et al.
Published: (2026)
by: Kaufmann, Max, et al.
Published: (2026)
Data selection method for assessment of autonomous vehicles
by: Trinh, Linh, et al.
Published: (2024)
by: Trinh, Linh, et al.
Published: (2024)
Fundamentals of legislation for autonomous artificial intelligence systems
by: Romanova, Anna
Published: (2024)
by: Romanova, Anna
Published: (2024)
Deep reinforcement learning-based longitudinal control strategy for automated vehicles at signalised intersections
by: Kumar, Pankaj, et al.
Published: (2025)
by: Kumar, Pankaj, et al.
Published: (2025)
Training and Serving System of Foundation Models: A Comprehensive Survey
by: Zhou, Jiahang, et al.
Published: (2024)
by: Zhou, Jiahang, et al.
Published: (2024)
Towards a constructive framework for control theory
by: Osinenko, Pavel
Published: (2025)
by: Osinenko, Pavel
Published: (2025)
Towards grounded autonomous research: an end-to-end LLM mini research loop on published computational physics
by: Huang, Haonan
Published: (2026)
by: Huang, Haonan
Published: (2026)
A collaborative agent with two lightweight synergistic models for autonomous crystal materials research
by: Shi, Tongyu, et al.
Published: (2026)
by: Shi, Tongyu, et al.
Published: (2026)
Prospective multi-pathogen disease forecasting using autonomous LLM-guided tree search
by: Martinson, Sarah, et al.
Published: (2026)
by: Martinson, Sarah, et al.
Published: (2026)
From nuclear safety to LLM security: Applying non-probabilistic risk management strategies to build safe and secure LLM-powered systems
by: Gutfraind, Alexander, et al.
Published: (2025)
by: Gutfraind, Alexander, et al.
Published: (2025)
Similar Items
-
DAHRS: Divergence-Aware Hallucination-Remediated SRL Projection
by: Youm, Sangpil, et al.
Published: (2024) -
Survey on safe robot control via learning
by: Mabsout, Bassel El
Published: (2024) -
How to safely discard features based on aggregate SHAP values
by: Bhattacharjee, Robi, et al.
Published: (2025) -
SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores
by: Mei, Zhiyu, et al.
Published: (2023) -
xSRL: Safety-Aware Explainable Reinforcement Learning -- Safety as a Product of Explainability
by: Shefin, Risal Shahriar, et al.
Published: (2024)