Optimal Transport-Guided Safety in Temporal Difference Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Shahrooei, Zahra, Baheri, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CENTS: Generating synthetic electricity consumption time series for rare and unseen scenarios
by: Fuest, Michael, et al.
Published: (2025)
by: Fuest, Michael, et al.
Published: (2025)
Fine-Tuning Integrity for Modern Neural Networks: Structured Drift Proofs via Norm, Rank, and Sparsity Certificates
by: Shang, Zhenhang, et al.
Published: (2026)
by: Shang, Zhenhang, et al.
Published: (2026)
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
Over-Squashing in Graph Neural Networks: A Comprehensive survey
by: Akansha, Singh
Published: (2023)
by: Akansha, Singh
Published: (2023)
Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels
by: Zhao, Yaqi, et al.
Published: (2026)
by: Zhao, Yaqi, et al.
Published: (2026)
Accelerating Monte-Carlo Tree Search with Optimized Posterior Policies
by: Frankston, Keith, et al.
Published: (2026)
by: Frankston, Keith, et al.
Published: (2026)
Practical Code RAG at Scale: Task-Aware Retrieval Design Choices under Compute Budgets
by: Galimzyanov, Timur, et al.
Published: (2025)
by: Galimzyanov, Timur, et al.
Published: (2025)
The Exploration of Neural Collapse under Imbalanced Data
by: Liu, Haixia
Published: (2024)
by: Liu, Haixia
Published: (2024)
Semantic Content Determines Algorithmic Performance
by: Ríos-García, Martiño, et al.
Published: (2026)
by: Ríos-García, Martiño, et al.
Published: (2026)
GraphCompNet: A Position-Aware Model for Predicting and Compensating Shape Deviations in 3D Printing
by: Lee, Juheon, et al.
Published: (2025)
by: Lee, Juheon, et al.
Published: (2025)
Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models
by: Leyva-Vázquez, Maikel Yelandi, et al.
Published: (2026)
by: Leyva-Vázquez, Maikel Yelandi, et al.
Published: (2026)
HADA: Human-AI Agent Decision Alignment Architecture
by: Pitkäranta, Tapio, et al.
Published: (2025)
by: Pitkäranta, Tapio, et al.
Published: (2025)
Optimizing Hospital Capacity During Pandemics: A Dual-Component Framework for Strategic Patient Relocation
by: Tabatabaee, Sadaf, et al.
Published: (2026)
by: Tabatabaee, Sadaf, et al.
Published: (2026)
FactoryBench: Evaluating Industrial Machine Understanding
by: Merzouki, Yanis, et al.
Published: (2026)
by: Merzouki, Yanis, et al.
Published: (2026)
Benchmarking machine learning for bowel sound pattern classification from tabular features to pretrained models
by: Mansour, Zahra, et al.
Published: (2025)
by: Mansour, Zahra, et al.
Published: (2025)
Towards Objective Gastrointestinal Auscultation: Automated Segmentation and Annotation of Bowel Sound Patterns
by: Mansour, Zahra, et al.
Published: (2026)
by: Mansour, Zahra, et al.
Published: (2026)
Cycling Race Time Prediction: A Personalized Machine Learning Approach Using Route Topology and Training Load
by: Moreno, Francisco Aguilera
Published: (2026)
by: Moreno, Francisco Aguilera
Published: (2026)
Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network
by: Mitra, Sayan, et al.
Published: (2026)
by: Mitra, Sayan, et al.
Published: (2026)
Differentially Private Adaptation of Diffusion Models via Noisy Aggregated Embeddings
by: Peetathawatchai, Pura, et al.
Published: (2024)
by: Peetathawatchai, Pura, et al.
Published: (2024)
OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness
by: Barnard, Amanda S
Published: (2026)
by: Barnard, Amanda S
Published: (2026)
STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts
by: Opria, Joshua
Published: (2026)
by: Opria, Joshua
Published: (2026)
An approach to encode divergence-free stress fields in neural approximations based on stress potentials
by: Khorrami, Mohammad S., et al.
Published: (2026)
by: Khorrami, Mohammad S., et al.
Published: (2026)
AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning
by: Zhao, Haotian, et al.
Published: (2026)
by: Zhao, Haotian, et al.
Published: (2026)
Conditional Shift-Robust Conformal Prediction for Graph Neural Network
by: Akansha, S.
Published: (2024)
by: Akansha, S.
Published: (2024)
Semantic Web and Software Agents -- A Forgotten Wave of Artificial Intelligence?
by: Pitkäranta, Tapio, et al.
Published: (2025)
by: Pitkäranta, Tapio, et al.
Published: (2025)
Where did you get that? Towards Summarization Attribution for Analysts
by: B, Violet, et al.
Published: (2025)
by: B, Violet, et al.
Published: (2025)
Optimal Transport-Assisted Risk-Sensitive Q-Learning
by: Shahrooei, Zahra, et al.
Published: (2024)
by: Shahrooei, Zahra, et al.
Published: (2024)
CapTune: Adapting Non-Speech Captions With Anchored Generative Models
by: Huang, Jeremy Zhengqi, et al.
Published: (2025)
by: Huang, Jeremy Zhengqi, et al.
Published: (2025)
From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents
by: Kroehl, Lars Kersten
Published: (2026)
by: Kroehl, Lars Kersten
Published: (2026)
Security Friction Quotient for Zero Trust Identity Policy with Empirical Validation
by: Youssef, Michel
Published: (2025)
by: Youssef, Michel
Published: (2025)
Uncover and Unlearn Nuisances: Agnostic Fully Test-Time Adaptation
by: Srey, Ponhvoan, et al.
Published: (2025)
by: Srey, Ponhvoan, et al.
Published: (2025)
Into the Gray Zone: Domain Contexts Can Blur LLM Safety Boundaries
by: Hung, Ki Sen, et al.
Published: (2026)
by: Hung, Ki Sen, et al.
Published: (2026)
PG-Triggers: Triggers for Property Graphs
by: Ceri, Stefano, et al.
Published: (2023)
by: Ceri, Stefano, et al.
Published: (2023)
LLMs: A Game-Changer for Software Engineers?
by: Haque, Md Asraful
Published: (2024)
by: Haque, Md Asraful
Published: (2024)
HateClipSeg: A Segment-Level Annotated Dataset for Fine-Grained Hate Video Detection
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
The NordDRG AI Benchmark for Large Language Models
by: Pitkäranta, Tapio
Published: (2025)
by: Pitkäranta, Tapio
Published: (2025)
LLMTrace: A Corpus for Classification and Fine-Grained Localization of AI-Written Text
by: Tolstykh, Irina, et al.
Published: (2025)
by: Tolstykh, Irina, et al.
Published: (2025)
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
by: Liu, Junxiao, et al.
Published: (2026)
by: Liu, Junxiao, et al.
Published: (2026)
Addressing Sustainability-IN Software Challenges
by: Calero, Coral, et al.
Published: (2024)
by: Calero, Coral, et al.
Published: (2024)
WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games
by: Zhang, Wenyu, et al.
Published: (2026)
by: Zhang, Wenyu, et al.
Published: (2026)
Similar Items
-
CENTS: Generating synthetic electricity consumption time series for rare and unseen scenarios
by: Fuest, Michael, et al.
Published: (2025) -
Fine-Tuning Integrity for Modern Neural Networks: Structured Drift Proofs via Norm, Rank, and Sparsity Certificates
by: Shang, Zhenhang, et al.
Published: (2026) -
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025) -
Over-Squashing in Graph Neural Networks: A Comprehensive survey
by: Akansha, Singh
Published: (2023) -
Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels
by: Zhao, Yaqi, et al.
Published: (2026)