Off-Switching Not Guaranteed
Fuente:
arXiv
Saved in:
| Main Author: | Neth, Sven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Noise Contrastive Estimation-based Matching Framework for Low-Resource Security Attack Pattern Recognition
by: Nguyen, Tu, et al.
Published: (2024)
by: Nguyen, Tu, et al.
Published: (2024)
The Partially Observable Off-Switch Game
by: Garber, Andrew, et al.
Published: (2024)
by: Garber, Andrew, et al.
Published: (2024)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
Dispatch-Aware Deep Neural Network for Optimal Transmission Switching: Toward Real-Time and Feasibility Guaranteed Operation
by: Kim, Minsoo, et al.
Published: (2025)
by: Kim, Minsoo, et al.
Published: (2025)
UnIT: Scalable Unstructured Inference-Time Pruning for MAC-efficient Neural Inference on MCUs
by: Neth, Ashe, et al.
Published: (2025)
by: Neth, Ashe, et al.
Published: (2025)
Action Model Learning with Guarantees
by: Aineto, Diego, et al.
Published: (2024)
by: Aineto, Diego, et al.
Published: (2024)
Simplification of Risk Averse POMDPs with Performance Guarantees
by: Pariente, Yaacov, et al.
Published: (2024)
by: Pariente, Yaacov, et al.
Published: (2024)
Composing Reinforcement Learning Policies, with Formal Guarantees
by: Delgrange, Florent, et al.
Published: (2024)
by: Delgrange, Florent, et al.
Published: (2024)
Hybrid Convolutional Neural Networks with Reliability Guarantee
by: Doran, Hans Dermot, et al.
Published: (2024)
by: Doran, Hans Dermot, et al.
Published: (2024)
SwitchCIT: Switching for Continual Instruction Tuning
by: Wu, Xinbo, et al.
Published: (2024)
by: Wu, Xinbo, et al.
Published: (2024)
Decomposability-Guaranteed Cooperative Coevolution for Large-Scale Itinerary Planning
by: Zhang, Ziyu, et al.
Published: (2025)
by: Zhang, Ziyu, et al.
Published: (2025)
Guaranteeing consistency in evidence fusion: A novel perspective on credibility
by: Ma, Chaoxiong, et al.
Published: (2025)
by: Ma, Chaoxiong, et al.
Published: (2025)
Measurement Simplification in ρ-POMDP with Performance Guarantees
by: Yotam, Tom, et al.
Published: (2023)
by: Yotam, Tom, et al.
Published: (2023)
Reward Bound for Behavioral Guarantee of Model-based Planning Agents
by: An, Zhiyu, et al.
Published: (2024)
by: An, Zhiyu, et al.
Published: (2024)
SeeUPO: Sequence-Level Agentic-RL with Convergence Guarantees
by: Hu, Tianyi, et al.
Published: (2026)
by: Hu, Tianyi, et al.
Published: (2026)
Selective Off-Policy Reference Tuning with Plan Guidance
by: Le, Duc Anh, et al.
Published: (2026)
by: Le, Duc Anh, et al.
Published: (2026)
Probabilistic Stability Guarantees for Feature Attributions
by: Jin, Helen, et al.
Published: (2025)
by: Jin, Helen, et al.
Published: (2025)
Off-OAB: Off-Policy Policy Gradient Method with Optimal Action-Dependent Baseline
by: Meng, Wenjia, et al.
Published: (2024)
by: Meng, Wenjia, et al.
Published: (2024)
Safety Guarantees in Zero-Shot Reinforcement Learning for Cascade Dynamical Systems
by: Rabiei, Shima, et al.
Published: (2026)
by: Rabiei, Shima, et al.
Published: (2026)
Online POMDP Planning with Anytime Deterministic Optimality Guarantees
by: Barenboim, Moran, et al.
Published: (2023)
by: Barenboim, Moran, et al.
Published: (2023)
Windowed MAPF with Completeness Guarantees
by: Veerapaneni, Rishi, et al.
Published: (2024)
by: Veerapaneni, Rishi, et al.
Published: (2024)
Off-policy Reinforcement Learning with Model-based Exploration Augmentation
by: Wang, Likun, et al.
Published: (2025)
by: Wang, Likun, et al.
Published: (2025)
Off-Trajectory Reasoning: Can LLMs Collaborate on Reasoning Trajectory?
by: Li, Aochong Oliver, et al.
Published: (2025)
by: Li, Aochong Oliver, et al.
Published: (2025)
Calibrated Prediction Set in Fault Detection with Risk Guarantees via Significance Tests
by: Mei, Mingchen, et al.
Published: (2025)
by: Mei, Mingchen, et al.
Published: (2025)
Why Self-Rewarding Works: Theoretical Guarantees for Iterative Alignment of Language Models
by: Fu, Shi, et al.
Published: (2026)
by: Fu, Shi, et al.
Published: (2026)
Guaranteed satisficing and finite regret: Analysis of a cognitive satisficing value function
by: Tamatsukuri, Akihiro, et al.
Published: (2018)
by: Tamatsukuri, Akihiro, et al.
Published: (2018)
Towards Conscious Service Robots
by: Behnke, Sven
Published: (2025)
by: Behnke, Sven
Published: (2025)
Guaranteed prediction sets for functional surrogate models
by: Gray, Ander, et al.
Published: (2025)
by: Gray, Ander, et al.
Published: (2025)
Debiasing Reward Models by Representation Learning with Guarantees
by: Ng, Ignavier, et al.
Published: (2025)
by: Ng, Ignavier, et al.
Published: (2025)
Do Biological Structural Guarantees Earn Their Complexity?
by: Banu, Bogdan
Published: (2026)
by: Banu, Bogdan
Published: (2026)
Rigorous Probabilistic Guarantees for Robust Counterfactual Explanations
by: Marzari, Luca, et al.
Published: (2024)
by: Marzari, Luca, et al.
Published: (2024)
Conditional Performance Guarantee for Large Reasoning Models
by: Huang, Jianguo, et al.
Published: (2026)
by: Huang, Jianguo, et al.
Published: (2026)
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
by: Zeighami, Sepanta, et al.
Published: (2025)
by: Zeighami, Sepanta, et al.
Published: (2025)
Learning Variable Impedance Skills from Demonstrations with Passivity Guarantee
by: Zhang, Yu, et al.
Published: (2023)
by: Zhang, Yu, et al.
Published: (2023)
AdaSwitch: Balancing Exploration and Guidance in Knowledge Distillation via Adaptive Switching
by: Peng, Jingyu, et al.
Published: (2025)
by: Peng, Jingyu, et al.
Published: (2025)
Behavior-Aware Auxiliary Corrections for Off-Policy Temporal-Difference Prediction
by: Chen, Xingguo, et al.
Published: (2026)
by: Chen, Xingguo, et al.
Published: (2026)
Clustering Context in Off-Policy Evaluation
by: Guzman-Olivares, Daniel, et al.
Published: (2025)
by: Guzman-Olivares, Daniel, et al.
Published: (2025)
Pluralistic Off-policy Evaluation and Alignment
by: Huang, Chengkai, et al.
Published: (2025)
by: Huang, Chengkai, et al.
Published: (2025)
Zero-Shot Off-Policy Learning
by: Asadulaev, Arip, et al.
Published: (2026)
by: Asadulaev, Arip, et al.
Published: (2026)
Concept-driven Off Policy Evaluation
by: Majumdar, Ritam, et al.
Published: (2024)
by: Majumdar, Ritam, et al.
Published: (2024)
Similar Items
-
Noise Contrastive Estimation-based Matching Framework for Low-Resource Security Attack Pattern Recognition
by: Nguyen, Tu, et al.
Published: (2024) -
The Partially Observable Off-Switch Game
by: Garber, Andrew, et al.
Published: (2024) -
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024) -
Dispatch-Aware Deep Neural Network for Optimal Transmission Switching: Toward Real-Time and Feasibility Guaranteed Operation
by: Kim, Minsoo, et al.
Published: (2025) -
UnIT: Scalable Unstructured Inference-Time Pruning for MAC-efficient Neural Inference on MCUs
by: Neth, Ashe, et al.
Published: (2025)