A Comparative Theoretical Analysis of Entropy Control Methods in Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lei, Ming, Baehr, Christophe |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility
di: Sun, Kangkang, et al.
Pubblicazione: (2025)
di: Sun, Kangkang, et al.
Pubblicazione: (2025)
Prediction and Forecast of Short-Term Drought Impacts Using Machine Learning to Support Mitigation and Adaptation Efforts
di: Geli, Hatim M. E., et al.
Pubblicazione: (2025)
di: Geli, Hatim M. E., et al.
Pubblicazione: (2025)
Non-linear Phillips Curve for India: Evidence from Explainable Machine Learning
di: Sengupta, Shovon, et al.
Pubblicazione: (2025)
di: Sengupta, Shovon, et al.
Pubblicazione: (2025)
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
di: Wigmore, Jerrod, et al.
Pubblicazione: (2024)
di: Wigmore, Jerrod, et al.
Pubblicazione: (2024)
From Static to Adaptive Defense: Federated Multi-Agent Deep Reinforcement Learning-Driven Moving Target Defense Against DoS Attacks in UAV Swarm Networks
di: Zhou, Yuyang, et al.
Pubblicazione: (2025)
di: Zhou, Yuyang, et al.
Pubblicazione: (2025)
Learning Symbolic Task Decompositions for Multi-Agent Teams
di: Shah, Ameesh, et al.
Pubblicazione: (2025)
di: Shah, Ameesh, et al.
Pubblicazione: (2025)
Continual Learning, Not Training: Online Adaptation For Agents
di: Jaglan, Aman, et al.
Pubblicazione: (2025)
di: Jaglan, Aman, et al.
Pubblicazione: (2025)
Geometric Meta-Learning via Coupled Ricci Flow: Unifying Knowledge Representation and Quantum Entanglement
di: Lei, Ming, et al.
Pubblicazione: (2025)
di: Lei, Ming, et al.
Pubblicazione: (2025)
Uncovering Bias Paths with LLM-guided Causal Discovery: An Active Learning and Dynamic Scoring Approach
di: Zanna, Khadija, et al.
Pubblicazione: (2025)
di: Zanna, Khadija, et al.
Pubblicazione: (2025)
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
Defense-as-a-Service: Black-box Shielding against Backdoored Graph Models
di: Yang, Xiao, et al.
Pubblicazione: (2024)
di: Yang, Xiao, et al.
Pubblicazione: (2024)
EnergyMamba: An Uncertainty-Aware Graph-Enhanced Selective State Space Model for Energy Consumption Prediction
di: Yu, Dahai, et al.
Pubblicazione: (2026)
di: Yu, Dahai, et al.
Pubblicazione: (2026)
Evaluating Model Robustness Using Adaptive Sparse L0 Regularization
di: Liu, Weiyou, et al.
Pubblicazione: (2024)
di: Liu, Weiyou, et al.
Pubblicazione: (2024)
Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
di: Li, Houyi, et al.
Pubblicazione: (2025)
di: Li, Houyi, et al.
Pubblicazione: (2025)
Distribution Consistency based Self-Training for Graph Neural Networks with Sparse Labels
di: Wang, Fali, et al.
Pubblicazione: (2024)
di: Wang, Fali, et al.
Pubblicazione: (2024)
Towards Verifiable AI with Lightweight Cryptographic Proofs of Inference
di: Anchuri, Pranay, et al.
Pubblicazione: (2026)
di: Anchuri, Pranay, et al.
Pubblicazione: (2026)
Deep Learning for Human Locomotion Analysis in Lower-Limb Exoskeletons: A Comparative Study
di: Coser, Omar, et al.
Pubblicazione: (2025)
di: Coser, Omar, et al.
Pubblicazione: (2025)
LAWS: Learning from Actual Workloads Symbolically -- A Self-Certifying Parametrized Cache Architecture for Neural Inference, Robotics, and Edge Deployment
di: Magarshak, Gregory
Pubblicazione: (2026)
di: Magarshak, Gregory
Pubblicazione: (2026)
Lightweight Quantum-Enhanced ResNet for Coronary Angiography Classification: A Hybrid Quantum-Classical Feature Enhancement Framework
di: Xia, Jingsong
Pubblicazione: (2026)
di: Xia, Jingsong
Pubblicazione: (2026)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
di: Catterall, Victoria, et al.
Pubblicazione: (2026)
di: Catterall, Victoria, et al.
Pubblicazione: (2026)
Software Model Evolution with Large Language Models: Experiments on Simulated, Public, and Industrial Datasets
di: Tinnes, Christof, et al.
Pubblicazione: (2024)
di: Tinnes, Christof, et al.
Pubblicazione: (2024)
LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models
di: Sikdar, Prateek Kumar
Pubblicazione: (2026)
di: Sikdar, Prateek Kumar
Pubblicazione: (2026)
DeepContext: Stateful Real-Time Detection of Multi-Turn Adversarial Intent Drift in LLMs
di: Albrethsen, Justin, et al.
Pubblicazione: (2026)
di: Albrethsen, Justin, et al.
Pubblicazione: (2026)
Transcribing Bengali Text with Regional Dialects to IPA using District Guided Tokens
di: Islam, S M Jishanul, et al.
Pubblicazione: (2024)
di: Islam, S M Jishanul, et al.
Pubblicazione: (2024)
When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Evaluating the Limitations of Local LLMs in Solving Complex Programming Challenges
di: Matotek, Kadin, et al.
Pubblicazione: (2025)
di: Matotek, Kadin, et al.
Pubblicazione: (2025)
Intelligence Inertia: Physical Isomorphism and Applications
di: Han, Jipeng
Pubblicazione: (2026)
di: Han, Jipeng
Pubblicazione: (2026)
Submodular Benchmark Selection
di: Smola, Alexander
Pubblicazione: (2026)
di: Smola, Alexander
Pubblicazione: (2026)
A Method for Evaluating the Interpretability of Machine Learning Models in Predicting Bond Default Risk Based on LIME and SHAP
di: Zhang, Yan, et al.
Pubblicazione: (2025)
di: Zhang, Yan, et al.
Pubblicazione: (2025)
JoFormer (Journey-based Transformer): Theory and Empirical Analysis on the Tiny Shakespeare Dataset
di: Godavarti, Mahesh
Pubblicazione: (2025)
di: Godavarti, Mahesh
Pubblicazione: (2025)
A Channel Attention-Driven Hybrid CNN Framework for Paddy Leaf Disease Detection
di: V, Pandiyaraju, et al.
Pubblicazione: (2024)
di: V, Pandiyaraju, et al.
Pubblicazione: (2024)
A Theoretical Computer Science Perspective on Free Will
di: Blum, Manuel, et al.
Pubblicazione: (2022)
di: Blum, Manuel, et al.
Pubblicazione: (2022)
MACI: Multi-Agent Collaborative Intelligence for Adaptive Reasoning and Temporal Planning
di: Chang, Edward Y.
Pubblicazione: (2025)
di: Chang, Edward Y.
Pubblicazione: (2025)
StarCraft+: Benchmarking Multi-agent Algorithms in Adversary Paradigm
di: Li, Yadong, et al.
Pubblicazione: (2025)
di: Li, Yadong, et al.
Pubblicazione: (2025)
Random Heterogeneous Neurochaos Learning Architecture for Data Classification
di: S, Remya Ajai A, et al.
Pubblicazione: (2024)
di: S, Remya Ajai A, et al.
Pubblicazione: (2024)
The Empowerment of Science of Science by Large Language Models: New Tools and Methods
di: Liang, Guoqiang, et al.
Pubblicazione: (2025)
di: Liang, Guoqiang, et al.
Pubblicazione: (2025)
Graph Coloring for Multi-Task Learning
di: Patapati, Santosh
Pubblicazione: (2025)
di: Patapati, Santosh
Pubblicazione: (2025)
Blockchain As a Platform For Artificial Intelligence (AI) Transparency
di: Akther, Afroja, et al.
Pubblicazione: (2025)
di: Akther, Afroja, et al.
Pubblicazione: (2025)
Counter-Inferential Behavior in Natural and Artificial Cognitive Systems
di: Dolgikh, Serge
Pubblicazione: (2025)
di: Dolgikh, Serge
Pubblicazione: (2025)
XAI and Few-shot-based Hybrid Classification Model for Plant Leaf Disease Prognosis
di: Joseph, Diana Susan, et al.
Pubblicazione: (2026)
di: Joseph, Diana Susan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility
di: Sun, Kangkang, et al.
Pubblicazione: (2025) -
Prediction and Forecast of Short-Term Drought Impacts Using Machine Learning to Support Mitigation and Adaptation Efforts
di: Geli, Hatim M. E., et al.
Pubblicazione: (2025) -
Non-linear Phillips Curve for India: Evidence from Explainable Machine Learning
di: Sengupta, Shovon, et al.
Pubblicazione: (2025) -
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
di: Wigmore, Jerrod, et al.
Pubblicazione: (2024) -
From Static to Adaptive Defense: Federated Multi-Agent Deep Reinforcement Learning-Driven Moving Target Defense Against DoS Attacks in UAV Swarm Networks
di: Zhou, Yuyang, et al.
Pubblicazione: (2025)