Efficient Action-Constrained Reinforcement Learning via Acceptance-Rejection Method and Augmented MDPs
Fuente:
arXiv
Guardado en:
| Autores principales: | Hung, Wei, Sun, Shao-Hua, Hsieh, Ping-Chun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Predicting Future Actions of Reinforcement Learning Agents
por: Chung, Stephen, et al.
Publicado: (2024)
por: Chung, Stephen, et al.
Publicado: (2024)
Structured Contrastive Learning for Interpretable Latent Representations
por: Shen, Zhengyang, et al.
Publicado: (2025)
por: Shen, Zhengyang, et al.
Publicado: (2025)
Nonparametric Partial Disentanglement via Mechanism Sparsity: Sparse Actions, Interventions and Sparse Temporal Dependencies
por: Lachapelle, Sébastien, et al.
Publicado: (2024)
por: Lachapelle, Sébastien, et al.
Publicado: (2024)
Augmenting Replay in World Models for Continual Reinforcement Learning
por: Yang, Luke, et al.
Publicado: (2024)
por: Yang, Luke, et al.
Publicado: (2024)
Combining Euclidean Alignment and Data Augmentation for BCI decoding
por: Rodrigues, Gustavo H., et al.
Publicado: (2024)
por: Rodrigues, Gustavo H., et al.
Publicado: (2024)
The Spotlight Resonance Method: Resolving the Alignment of Embedded Activations
por: Bird, George
Publicado: (2025)
por: Bird, George
Publicado: (2025)
Annot-Mix: Learning with Noisy Class Labels from Multiple Annotators via a Mixup Extension
por: Herde, Marek, et al.
Publicado: (2024)
por: Herde, Marek, et al.
Publicado: (2024)
Feature emergence via margin maximization: case studies in algebraic tasks
por: Morwani, Depen, et al.
Publicado: (2023)
por: Morwani, Depen, et al.
Publicado: (2023)
ZClassifier: Temperature Tuning and Manifold Approximation via KL Divergence on Logit Space
por: Yong, Shim Soon
Publicado: (2025)
por: Yong, Shim Soon
Publicado: (2025)
Learning to Land Anywhere: Transferable Generative Models for Aircraft Trajectories
por: Larsen, Olav Finne Praesteng, et al.
Publicado: (2025)
por: Larsen, Olav Finne Praesteng, et al.
Publicado: (2025)
SATORIS-N: Spectral Analysis based Traffic Observation Recovery via Informed Subspaces and Nuclear-norm minimization
por: Mohanty, Sampad, et al.
Publicado: (2026)
por: Mohanty, Sampad, et al.
Publicado: (2026)
AGOP-IxG: A Gradient Covariance Filter for Local Feature Attribution on Tabular Data, with a Controlled Benchmark
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
por: Dutta, Utsav, et al.
Publicado: (2026)
por: Dutta, Utsav, et al.
Publicado: (2026)
Mamba base PKD for efficient knowledge compression
por: Medina, José, et al.
Publicado: (2025)
por: Medina, José, et al.
Publicado: (2025)
Subspace Geometry Governs Catastrophic Forgetting in Low-Rank Adaptation
por: Steele, Brady
Publicado: (2026)
por: Steele, Brady
Publicado: (2026)
Remaining Useful Life Estimation for Turbofan Engines: A Comparative Study of Classical, CNN, and LSTM Approaches
por: Goel, Astitva, et al.
Publicado: (2026)
por: Goel, Astitva, et al.
Publicado: (2026)
Contrastive Representation Modeling for Anomaly Detection
por: Lunardi, Willian T., et al.
Publicado: (2025)
por: Lunardi, Willian T., et al.
Publicado: (2025)
Improving Consistency in Large Language Models through Chain of Guidance
por: Raj, Harsh, et al.
Publicado: (2025)
por: Raj, Harsh, et al.
Publicado: (2025)
AdaCap: An Adaptive Contrastive Approach for Small-Data Neural Networks
por: Belucci, Bruno, et al.
Publicado: (2025)
por: Belucci, Bruno, et al.
Publicado: (2025)
Entropy Aware Message Passing in Graph Neural Networks
por: Nazari, Philipp, et al.
Publicado: (2024)
por: Nazari, Philipp, et al.
Publicado: (2024)
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
por: Priezzhev, I. I., et al.
Publicado: (2025)
por: Priezzhev, I. I., et al.
Publicado: (2025)
Out of Distribution Detection for Efficient Continual Learning in Quality Prediction for Arc Welding
por: Hahn, Yannik, et al.
Publicado: (2025)
por: Hahn, Yannik, et al.
Publicado: (2025)
In-situ Autoguidance: Eliciting Self-Correction in Diffusion Models
por: Gu, Enhao, et al.
Publicado: (2025)
por: Gu, Enhao, et al.
Publicado: (2025)
CRLLK: Constrained Reinforcement Learning for Lane Keeping in Autonomous Driving
por: Gao, Xinwei, et al.
Publicado: (2025)
por: Gao, Xinwei, et al.
Publicado: (2025)
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
por: Ho, Siu Hang, et al.
Publicado: (2025)
por: Ho, Siu Hang, et al.
Publicado: (2025)
Neural Reasoning Networks: Efficient Interpretable Neural Networks With Automatic Textual Explanations
por: Carrow, Stephen, et al.
Publicado: (2024)
por: Carrow, Stephen, et al.
Publicado: (2024)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
por: Sun, Yuhui, et al.
Publicado: (2025)
por: Sun, Yuhui, et al.
Publicado: (2025)
Enhanced Protein Intrinsic Disorder Prediction Through Dual-View Multiscale Features and Multi-objective Evolutionary Algorithm
por: Wang, Shaokuan, et al.
Publicado: (2026)
por: Wang, Shaokuan, et al.
Publicado: (2026)
QuIDE: Mastering the Quantized Intelligence Trade-off via Active Optimization
por: Jiang, Xiantao
Publicado: (2026)
por: Jiang, Xiantao
Publicado: (2026)
Massive Redundancy in Gradient Transport Enables Sparse Online Learning
por: Merin, Aur Shalev
Publicado: (2026)
por: Merin, Aur Shalev
Publicado: (2026)
GS-KAN: Parameter-Efficient Kolmogorov-Arnold Networks via Sprecher-Type Shared Basis Functions
por: Eliasson, Oscar
Publicado: (2025)
por: Eliasson, Oscar
Publicado: (2025)
SeizureFormer: A Transformer Model for IEA-Based Seizure Risk Forecasting
por: Feng, Tianning, et al.
Publicado: (2025)
por: Feng, Tianning, et al.
Publicado: (2025)
Step-E: A Differentiable Data Cleaning Framework for Robust Learning with Noisy Labels
por: Du, Wenzhang
Publicado: (2025)
por: Du, Wenzhang
Publicado: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
por: Semenov, Andrei, et al.
Publicado: (2024)
por: Semenov, Andrei, et al.
Publicado: (2024)
Synergizing Deep Learning and Biological Heuristics for Extreme Long-Tail White Blood Cell Classification
por: Nguyen, Duc T., et al.
Publicado: (2026)
por: Nguyen, Duc T., et al.
Publicado: (2026)
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
por: Rottenwalter, Georg, et al.
Publicado: (2025)
por: Rottenwalter, Georg, et al.
Publicado: (2025)
Advancements in synthetic data extraction for industrial injection molding
por: Rottenwalter, Georg, et al.
Publicado: (2025)
por: Rottenwalter, Georg, et al.
Publicado: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
por: Reddy, Sandeep, et al.
Publicado: (2025)
por: Reddy, Sandeep, et al.
Publicado: (2025)
Leveraging Causal Reasoning Method for Explaining Medical Image Segmentation Models
por: Jiang, Limai, et al.
Publicado: (2026)
por: Jiang, Limai, et al.
Publicado: (2026)
Ejemplares similares
-
Predicting Future Actions of Reinforcement Learning Agents
por: Chung, Stephen, et al.
Publicado: (2024) -
Structured Contrastive Learning for Interpretable Latent Representations
por: Shen, Zhengyang, et al.
Publicado: (2025) -
Nonparametric Partial Disentanglement via Mechanism Sparsity: Sparse Actions, Interventions and Sparse Temporal Dependencies
por: Lachapelle, Sébastien, et al.
Publicado: (2024) -
Augmenting Replay in World Models for Continual Reinforcement Learning
por: Yang, Luke, et al.
Publicado: (2024) -
Combining Euclidean Alignment and Data Augmentation for BCI decoding
por: Rodrigues, Gustavo H., et al.
Publicado: (2024)