Enregistré dans:
| Auteurs principaux: | Hafez, Wael, Reid, Cameron, Nazeri, Amit |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2603.01283 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Mutual Information Tracks Policy Coherence in Reinforcement Learning
par: Reid, Cameron, et autres
Publié: (2025)
par: Reid, Cameron, et autres
Publié: (2025)
A Mathematical Theory of Agency and Intelligence
par: Hafez, Wael, et autres
Publié: (2026)
par: Hafez, Wael, et autres
Publié: (2026)
Token Statistics Reveal Conversational Drift in Multi-turn LLM Interaction
par: Hafez, Wael, et autres
Publié: (2026)
par: Hafez, Wael, et autres
Publié: (2026)
Information-Theoretic Framework for Self-Adapting Model Predictive Controllers
par: Hafez, Wael, et autres
Publié: (2026)
par: Hafez, Wael, et autres
Publié: (2026)
Entropy-Based Non-Invasive Reliability Monitoring of Convolutional Neural Networks
par: Nazeri, Amirhossein, et autres
Publié: (2025)
par: Nazeri, Amirhossein, et autres
Publié: (2025)
To Measure or Not: A Cost-Sensitive, Selective Measuring Environment for Agricultural Management Decisions with Reinforcement Learning
par: Baja, Hilmy, et autres
Publié: (2025)
par: Baja, Hilmy, et autres
Publié: (2025)
V-Max: A Reinforcement Learning Framework for Autonomous Driving
par: Charraut, Valentin, et autres
Publié: (2025)
par: Charraut, Valentin, et autres
Publié: (2025)
Momentum-Based Federated Reinforcement Learning with Interaction and Communication Efficiency
par: Yue, Sheng, et autres
Publié: (2024)
par: Yue, Sheng, et autres
Publié: (2024)
Learning Markov State Abstractions for Deep Reinforcement Learning
par: Allen, Cameron, et autres
Publié: (2021)
par: Allen, Cameron, et autres
Publié: (2021)
From Pixels to Factors: Learning Independently Controllable State Variables for Reinforcement Learning
par: Rodriguez-Sanchez, Rafael, et autres
Publié: (2025)
par: Rodriguez-Sanchez, Rafael, et autres
Publié: (2025)
A Fast Anti-Jamming Cognitive Radar Deployment Algorithm Based on Reinforcement Learning
par: Cai, Wencheng, et autres
Publié: (2025)
par: Cai, Wencheng, et autres
Publié: (2025)
Imitating Cost-Constrained Behaviors in Reinforcement Learning
par: Shao, Qian, et autres
Publié: (2024)
par: Shao, Qian, et autres
Publié: (2024)
To Train or Not to Train: Balancing Efficiency and Training Cost in Deep Reinforcement Learning for Mobile Edge Computing
par: Boscaro, Maddalena, et autres
Publié: (2024)
par: Boscaro, Maddalena, et autres
Publié: (2024)
On The Sample Complexity Bounds In Bilevel Reinforcement Learning
par: Gaur, Mudit, et autres
Publié: (2025)
par: Gaur, Mudit, et autres
Publié: (2025)
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
par: Vamplew, Peter, et autres
Publié: (2024)
par: Vamplew, Peter, et autres
Publié: (2024)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
par: Vakili, Sattar, et autres
Publié: (2023)
par: Vakili, Sattar, et autres
Publié: (2023)
From Efficiency to Equity: Measuring Fairness in Preference Learning
par: Gowaikar, Shreeyash, et autres
Publié: (2024)
par: Gowaikar, Shreeyash, et autres
Publié: (2024)
Interactive Ontology Matching with Cost-Efficient Learning
par: Cheng, Bin, et autres
Publié: (2024)
par: Cheng, Bin, et autres
Publié: (2024)
Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains
par: Tao, Ruo Yu, et autres
Publié: (2025)
par: Tao, Ruo Yu, et autres
Publié: (2025)
Improved Bayesian Regret Bounds for Thompson Sampling in Reinforcement Learning
par: Moradipari, Ahmadreza, et autres
Publié: (2023)
par: Moradipari, Ahmadreza, et autres
Publié: (2023)
Regret Bounds and Reinforcement Learning Exploration of EXP-based Algorithms
par: Xu, Mengfan, et autres
Publié: (2020)
par: Xu, Mengfan, et autres
Publié: (2020)
Performance Control in Early Exiting to Deploy Large Models at the Same Cost of Smaller Ones
par: Mofakhami, Mehrnaz, et autres
Publié: (2024)
par: Mofakhami, Mehrnaz, et autres
Publié: (2024)
ML Compass: Navigating Capability, Cost, and Compliance Trade-offs in AI Model Deployment
par: Digalakis Jr, Vassilis, et autres
Publié: (2025)
par: Digalakis Jr, Vassilis, et autres
Publié: (2025)
Subgoal-based Reward Shaping to Improve Efficiency in Reinforcement Learning
par: Okudo, Takato, et autres
Publié: (2021)
par: Okudo, Takato, et autres
Publié: (2021)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
par: Ghaemi, Hafez, et autres
Publié: (2025)
par: Ghaemi, Hafez, et autres
Publié: (2025)
Bounded Ratio Reinforcement Learning
par: Ao, Yunke, et autres
Publié: (2026)
par: Ao, Yunke, et autres
Publié: (2026)
Honesty to Subterfuge: In-Context Reinforcement Learning Can Make Honest Models Reward Hack
par: McKee-Reid, Leo, et autres
Publié: (2024)
par: McKee-Reid, Leo, et autres
Publié: (2024)
Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Policy Adaptation for Sim-to-Real Deployment
par: Han, Gengyue, et autres
Publié: (2026)
par: Han, Gengyue, et autres
Publié: (2026)
Paged Attention Meets FlexAttention: Unlocking Long-Context Efficiency in Deployed Inference
par: Joshi, Thomas, et autres
Publié: (2025)
par: Joshi, Thomas, et autres
Publié: (2025)
Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
par: Talvitie, Erin J., et autres
Publié: (2024)
par: Talvitie, Erin J., et autres
Publié: (2024)
Bridging Distributional and Risk-sensitive Reinforcement Learning with Provable Regret Bounds
par: Liang, Hao, et autres
Publié: (2022)
par: Liang, Hao, et autres
Publié: (2022)
Theoretical Analysis of Meta Reinforcement Learning: Generalization Bounds and Convergence Guarantees
par: Wang, Cangqing, et autres
Publié: (2024)
par: Wang, Cangqing, et autres
Publié: (2024)
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
par: Liu, Zhishuai, et autres
Publié: (2024)
par: Liu, Zhishuai, et autres
Publié: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
par: Nofshin, Eura, et autres
Publié: (2024)
par: Nofshin, Eura, et autres
Publié: (2024)
Echoes of Socratic Doubt: Embracing Uncertainty in Calibrated Evidential Reinforcement Learning
par: Stutts, Alex Christopher, et autres
Publié: (2024)
par: Stutts, Alex Christopher, et autres
Publié: (2024)
A Cost-Benefit Analysis of On-Premise Large Language Model Deployment: Breaking Even with Commercial LLM Services
par: Pan, Guanzhong, et autres
Publié: (2025)
par: Pan, Guanzhong, et autres
Publié: (2025)
LEASE: Offline Preference-based Reinforcement Learning with High Sample Efficiency
par: Liu, Xiao-Yin, et autres
Publié: (2024)
par: Liu, Xiao-Yin, et autres
Publié: (2024)
On the Sample Efficiency of Abstractions and Potential-Based Reward Shaping in Reinforcement Learning
par: Canonaco, Giuseppe, et autres
Publié: (2024)
par: Canonaco, Giuseppe, et autres
Publié: (2024)
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
par: Huang, Jiawei, et autres
Publié: (2023)
par: Huang, Jiawei, et autres
Publié: (2023)
Interactive Multi-Objective Probabilistic Preference Learning with Soft and Hard Bounds
par: Chen, Edward, et autres
Publié: (2025)
par: Chen, Edward, et autres
Publié: (2025)
Documents similaires
-
Mutual Information Tracks Policy Coherence in Reinforcement Learning
par: Reid, Cameron, et autres
Publié: (2025) -
A Mathematical Theory of Agency and Intelligence
par: Hafez, Wael, et autres
Publié: (2026) -
Token Statistics Reveal Conversational Drift in Multi-turn LLM Interaction
par: Hafez, Wael, et autres
Publié: (2026) -
Information-Theoretic Framework for Self-Adapting Model Predictive Controllers
par: Hafez, Wael, et autres
Publié: (2026) -
Entropy-Based Non-Invasive Reliability Monitoring of Convolutional Neural Networks
par: Nazeri, Amirhossein, et autres
Publié: (2025)