GENESIS-RL: GEnerating Natural Edge-cases with Systematic Integration of Safety considerations and Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Hsin-Jung, Beck, Joe, Hasan, Md Zahid, Beyazit, Ekin, Chakraborty, Subhadeep, Wongpiromsarn, Tichakorn, Sarkar, Soumik |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Incorporating System-level Safety Requirements in Perception Models via Reinforcement Learning
di: Fan, Weisi, et al.
Pubblicazione: (2024)
di: Fan, Weisi, et al.
Pubblicazione: (2024)
Risk-Aware Rulebooks for Multi-Objective Trajectory Evaluation under Uncertainty
di: Wongpiromsarn, Tichakorn
Pubblicazione: (2026)
di: Wongpiromsarn, Tichakorn
Pubblicazione: (2026)
ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abstract Scenarios
di: Chang, Kevin Kai-Chun, et al.
Pubblicazione: (2026)
di: Chang, Kevin Kai-Chun, et al.
Pubblicazione: (2026)
Lexicographic Multi-Objective Stochastic Shortest Path with Mixed Max-Sum Costs
di: Zhang, Zhiquan, et al.
Pubblicazione: (2025)
di: Zhang, Zhiquan, et al.
Pubblicazione: (2025)
Diagnosing and Predicting Autonomous Vehicle Operational Safety Using Multiple Simulation Modalities and a Virtual Environment
di: Beck, Joe, et al.
Pubblicazione: (2024)
di: Beck, Joe, et al.
Pubblicazione: (2024)
Fully Embedded Time-Series Generative Adversarial Networks
di: Beck, Joe, et al.
Pubblicazione: (2023)
di: Beck, Joe, et al.
Pubblicazione: (2023)
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception
di: Waite, Joshua R., et al.
Pubblicazione: (2025)
di: Waite, Joshua R., et al.
Pubblicazione: (2025)
Find the Fruit: Zero-Shot Sim2Real RL for Occlusion-Aware Plant Manipulation
di: Subedi, Nitesh, et al.
Pubblicazione: (2025)
di: Subedi, Nitesh, et al.
Pubblicazione: (2025)
LexiSafe: Offline Safe Reinforcement Learning with Lexicographic Safety-Reward Hierarchy
di: Yang, Hsin-Jung, et al.
Pubblicazione: (2026)
di: Yang, Hsin-Jung, et al.
Pubblicazione: (2026)
Balancing Utility and Privacy: Dynamically Private SGD with Random Projection
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
Zero-shot Sim-to-Real Transfer for Reinforcement Learning-based Visual Servoing of Soft Continuum Arms
di: Yang, Hsin-Jung, et al.
Pubblicazione: (2025)
di: Yang, Hsin-Jung, et al.
Pubblicazione: (2025)
FUSE: First-Order and Second-Order Unified SynthEsis in Stochastic Optimization
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
Data-driven Kinematic Modeling in Soft Robots: System Identification and Uncertainty Quantification
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
di: Jiang, Zhanhong, et al.
Pubblicazione: (2025)
Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning
di: Koirala, Prajwal, et al.
Pubblicazione: (2024)
di: Koirala, Prajwal, et al.
Pubblicazione: (2024)
CAGE: Controllable Articulation GEneration
di: Liu, Jiayi, et al.
Pubblicazione: (2023)
di: Liu, Jiayi, et al.
Pubblicazione: (2023)
FAWAC: Feasibility Informed Advantage Weighted Regression for Persistent Safety in Offline Reinforcement Learning
di: Koirala, Prajwal, et al.
Pubblicazione: (2024)
di: Koirala, Prajwal, et al.
Pubblicazione: (2024)
BitRL: Reinforcement Learning with 1-bit Quantized Language Models for Resource-Constrained Edge Deployment
di: Sajid, Md. Ashiq Ul Islam, et al.
Pubblicazione: (2026)
di: Sajid, Md. Ashiq Ul Islam, et al.
Pubblicazione: (2026)
Driving as a Diagnostic Tool: Scenario-based Cognitive Assessment in Older Drivers from Driving Video
di: Hasan, Md Zahid, et al.
Pubblicazione: (2025)
di: Hasan, Md Zahid, et al.
Pubblicazione: (2025)
Reinforcement Learning for Autonomous Point-to-Point UAV Navigation
di: Oyinlola, Salim, et al.
Pubblicazione: (2025)
di: Oyinlola, Salim, et al.
Pubblicazione: (2025)
Real-Time Multi-Modal Embedded Vision Framework for Object Detection Facial Emotion Recognition and Biometric Identification on Low-Power Edge Platforms
di: Zahid, S. M. Khalid Bin, et al.
Pubblicazione: (2026)
di: Zahid, S. M. Khalid Bin, et al.
Pubblicazione: (2026)
commensurability: a Python package for classifying astronomical orbits based on their toroid volume
di: Sarkar, Subhadeep, et al.
Pubblicazione: (2025)
di: Sarkar, Subhadeep, et al.
Pubblicazione: (2025)
commensurability: a Python package for classifying astronomical orbits based on their toroid volume
di: Sarkar, Subhadeep, et al.
Pubblicazione: (2025)
di: Sarkar, Subhadeep, et al.
Pubblicazione: (2025)
FROST: Filtering Reasoning Outliers with Attention for Efficient Reasoning
di: Luo, Haozheng, et al.
Pubblicazione: (2026)
di: Luo, Haozheng, et al.
Pubblicazione: (2026)
Vision-Language Models can Identify Distracted Driver Behavior from Naturalistic Videos
di: Hasan, Md Zahid, et al.
Pubblicazione: (2023)
di: Hasan, Md Zahid, et al.
Pubblicazione: (2023)
HEDGE: Heterogeneous Ensemble for Detection of AI-GEnerated Images in the Wild
di: Wu, Fei, et al.
Pubblicazione: (2026)
di: Wu, Fei, et al.
Pubblicazione: (2026)
Improving sensitivity of trilinear RPV SUSY searches using machine learning at the LHC
di: Choudhury, Arghya, et al.
Pubblicazione: (2023)
di: Choudhury, Arghya, et al.
Pubblicazione: (2023)
Slepton searches in the trilinear RPV SUSY scenarios at the HL-LHC and HE-LHC
di: Choudhury, Arghya, et al.
Pubblicazione: (2023)
di: Choudhury, Arghya, et al.
Pubblicazione: (2023)
Enhancing PPO with Trajectory-Aware Hybrid Policies
di: Liu, Qisai, et al.
Pubblicazione: (2025)
di: Liu, Qisai, et al.
Pubblicazione: (2025)
ORANGE: An Online Reflection ANd GEneration framework with Domain Knowledge for Text-to-SQL
di: Jiao, Yiwen, et al.
Pubblicazione: (2025)
di: Jiao, Yiwen, et al.
Pubblicazione: (2025)
Monitoring Real-Time ECG Signals on Mobile Systems
di: Yuksel, Beyazit Bestami
Pubblicazione: (2025)
di: Yuksel, Beyazit Bestami
Pubblicazione: (2025)
HEART: A High-Efficiency Adaptive Real-Time Telemonitoring Framework for Secure Electrocardiogram Signal Transmission Using Chaotic Encryption
di: Yuksel, Beyazıt Bestami
Pubblicazione: (2026)
di: Yuksel, Beyazıt Bestami
Pubblicazione: (2026)
GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
Are LLMs Ready to Replace Bangla Annotators?
di: Hasan, Md. Najib, et al.
Pubblicazione: (2026)
di: Hasan, Md. Najib, et al.
Pubblicazione: (2026)
Quantum Control of Heat Current
di: Chakraborty, Gobinda, et al.
Pubblicazione: (2023)
di: Chakraborty, Gobinda, et al.
Pubblicazione: (2023)
Searches for the BSM scenarios at the LHC using decision tree based machine learning algorithms: A comparative study and review of Random Forest, Adaboost, XGboost and LightGBM frameworks
di: Choudhury, Arghya, et al.
Pubblicazione: (2024)
di: Choudhury, Arghya, et al.
Pubblicazione: (2024)
Exploring Photophysical Characterization and Bioimaging Applications of Benzomorpholino, Benzopiperazinyl, and Quinoxalino‐Based Amino Terephthalonitriles
di: Ankita Sinha, et al.
Pubblicazione: (2025)
di: Ankita Sinha, et al.
Pubblicazione: (2025)
Robust Majorana bound state in pseudo-spin domain wall of 2-D topological insulator
di: Chakraborty, Subhadeep, et al.
Pubblicazione: (2024)
di: Chakraborty, Subhadeep, et al.
Pubblicazione: (2024)
Assuring the Safety of Reinforcement Learning Components: AMLAS-RL
di: Imrie, Calum Corrie, et al.
Pubblicazione: (2025)
di: Imrie, Calum Corrie, et al.
Pubblicazione: (2025)
Association of opportunistic bacterial pathogens with female infertility: A case–control study
di: Zahid Hasan, et al.
Pubblicazione: (2025)
di: Zahid Hasan, et al.
Pubblicazione: (2025)
Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines
di: Soumik, Sadman Kabir
Pubblicazione: (2026)
di: Soumik, Sadman Kabir
Pubblicazione: (2026)
Documenti analoghi
-
Incorporating System-level Safety Requirements in Perception Models via Reinforcement Learning
di: Fan, Weisi, et al.
Pubblicazione: (2024) -
Risk-Aware Rulebooks for Multi-Objective Trajectory Evaluation under Uncertainty
di: Wongpiromsarn, Tichakorn
Pubblicazione: (2026) -
ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abstract Scenarios
di: Chang, Kevin Kai-Chun, et al.
Pubblicazione: (2026) -
Lexicographic Multi-Objective Stochastic Shortest Path with Mixed Max-Sum Costs
di: Zhang, Zhiquan, et al.
Pubblicazione: (2025) -
Diagnosing and Predicting Autonomous Vehicle Operational Safety Using Multiple Simulation Modalities and a Virtual Environment
di: Beck, Joe, et al.
Pubblicazione: (2024)