PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management
Fuente:
arXiv
Salvato in:
| Autori principali: | Zaregarizi, Shadmehr, Yavari, Khashayar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Uncertainty-Aware Transfer Learning for Cross-Building Energy Forecasting: Toward Robust and Scalable District-Level Energy Management
di: Zaregarizi, Shadmehr, et al.
Pubblicazione: (2026)
di: Zaregarizi, Shadmehr, et al.
Pubblicazione: (2026)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
di: Yousaf, Iqra
Pubblicazione: (2024)
di: Yousaf, Iqra
Pubblicazione: (2024)
Constrained Auto-Bidding via Generative Response Modeling
di: Yang, Eunseok, et al.
Pubblicazione: (2026)
di: Yang, Eunseok, et al.
Pubblicazione: (2026)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
di: Aichmüller, Michael, et al.
Pubblicazione: (2024)
di: Aichmüller, Michael, et al.
Pubblicazione: (2024)
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
di: George, Robert Joseph, et al.
Pubblicazione: (2025)
di: George, Robert Joseph, et al.
Pubblicazione: (2025)
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
NeSIG: A Neuro-Symbolic Method for Learning to Generate Planning Problems
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2023)
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2023)
Learning to Select Goals in Automated Planning with Deep-Q Learning
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2024)
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2024)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
di: Belcamino, Valerio, et al.
Pubblicazione: (2026)
di: Belcamino, Valerio, et al.
Pubblicazione: (2026)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
di: Wu, Yalun, et al.
Pubblicazione: (2026)
di: Wu, Yalun, et al.
Pubblicazione: (2026)
When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State
di: Zhu, Peiying, et al.
Pubblicazione: (2026)
di: Zhu, Peiying, et al.
Pubblicazione: (2026)
Multi-Task Genetic Algorithm with Multi-Granularity Encoding for Protein-Nucleotide Binding Site Prediction
di: Gao, Yiming, et al.
Pubblicazione: (2026)
di: Gao, Yiming, et al.
Pubblicazione: (2026)
From Next Token Prediction to (STRIPS) World Models
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2025)
di: Núñez-Molina, Carlos, et al.
Pubblicazione: (2025)
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
di: Priezzhev, I. I., et al.
Pubblicazione: (2025)
di: Priezzhev, I. I., et al.
Pubblicazione: (2025)
EcoNet: Multiagent Planning and Control Of Household Energy Resources Using Active Inference
di: Boik, John C., et al.
Pubblicazione: (2025)
di: Boik, John C., et al.
Pubblicazione: (2025)
Foundational Requirements for Artificial General Intelligence: A Falsifiable Framework Based on Signal Prediction
di: Šprogar, Matej
Pubblicazione: (2025)
di: Šprogar, Matej
Pubblicazione: (2025)
Adaptable Hindsight Experience Replay for Search-Based Learning
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
di: Zhou, Yue, et al.
Pubblicazione: (2025)
di: Zhou, Yue, et al.
Pubblicazione: (2025)
Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning
di: Wang, Zizhao, et al.
Pubblicazione: (2024)
di: Wang, Zizhao, et al.
Pubblicazione: (2024)
Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy
di: He, Langzhou, et al.
Pubblicazione: (2026)
di: He, Langzhou, et al.
Pubblicazione: (2026)
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
di: Rottenwalter, Georg, et al.
Pubblicazione: (2025)
di: Rottenwalter, Georg, et al.
Pubblicazione: (2025)
Advancements in synthetic data extraction for industrial injection molding
di: Rottenwalter, Georg, et al.
Pubblicazione: (2025)
di: Rottenwalter, Georg, et al.
Pubblicazione: (2025)
Predicting Future Actions of Reinforcement Learning Agents
di: Chung, Stephen, et al.
Pubblicazione: (2024)
di: Chung, Stephen, et al.
Pubblicazione: (2024)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
di: Schipper, Olivier, et al.
Pubblicazione: (2025)
di: Schipper, Olivier, et al.
Pubblicazione: (2025)
How Metacognitive Architectures Remember Their Own Thoughts: A Systematic Review
di: Nolte, Robin, et al.
Pubblicazione: (2025)
di: Nolte, Robin, et al.
Pubblicazione: (2025)
Differentiable Symbolic Planning: A Neural Architecture for Constraint Reasoning with Learned Feasibility
di: Oruganti, Venkatakrishna Reddy
Pubblicazione: (2026)
di: Oruganti, Venkatakrishna Reddy
Pubblicazione: (2026)
Safe Reinforcement Learning with Preference-based Constraint Inference
di: Li, Chenglin, et al.
Pubblicazione: (2026)
di: Li, Chenglin, et al.
Pubblicazione: (2026)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
Regret-Aware Policy Optimization: Environment-Level Memory for Replay Suppression under Delayed Harm
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
Not All Transitions Matter: Evidence from PPO
di: Basnet, Ajhesh
Pubblicazione: (2026)
di: Basnet, Ajhesh
Pubblicazione: (2026)
AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites
di: Zhang, Qinshi, et al.
Pubblicazione: (2026)
di: Zhang, Qinshi, et al.
Pubblicazione: (2026)
What Do World Models Learn in RL? Probing Latent Representations in Learned Environment Simulators
di: Zhang, Xinyu
Pubblicazione: (2026)
di: Zhang, Xinyu
Pubblicazione: (2026)
Embedded Safety-Aligned Intelligence via Differentiable Internal Alignment Embeddings
di: Rathva, Harsh, et al.
Pubblicazione: (2025)
di: Rathva, Harsh, et al.
Pubblicazione: (2025)
Working Paper: Active Causal Structure Learning with Latent Variables: Towards Learning to Detour in Autonomous Robots
di: Riscos, Pablo de los, et al.
Pubblicazione: (2024)
di: Riscos, Pablo de los, et al.
Pubblicazione: (2024)
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
di: Zawalski, Michał, et al.
Pubblicazione: (2022)
di: Zawalski, Michał, et al.
Pubblicazione: (2022)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
di: Nuzhin, Egor E., et al.
Pubblicazione: (2024)
di: Nuzhin, Egor E., et al.
Pubblicazione: (2024)
Inverting Cryptographic Hash Functions via Cube-and-Conquer
di: Zaikin, Oleg
Pubblicazione: (2022)
di: Zaikin, Oleg
Pubblicazione: (2022)
Bridging the Reasoning Gap: Small LLMs Can Plan with Generalised Strategies
di: Borro, Andrey, et al.
Pubblicazione: (2025)
di: Borro, Andrey, et al.
Pubblicazione: (2025)
Incentives for Responsiveness, Instrumental Control and Impact
di: Carey, Ryan, et al.
Pubblicazione: (2020)
di: Carey, Ryan, et al.
Pubblicazione: (2020)
Score-informed Neural Operator for Enhancing Ordering-based Causal Discovery
di: Kang, Jiyeon, et al.
Pubblicazione: (2025)
di: Kang, Jiyeon, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Uncertainty-Aware Transfer Learning for Cross-Building Energy Forecasting: Toward Robust and Scalable District-Level Energy Management
di: Zaregarizi, Shadmehr, et al.
Pubblicazione: (2026) -
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
di: Yousaf, Iqra
Pubblicazione: (2024) -
Constrained Auto-Bidding via Generative Response Modeling
di: Yang, Eunseok, et al.
Pubblicazione: (2026) -
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
di: Aichmüller, Michael, et al.
Pubblicazione: (2024) -
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
di: George, Robert Joseph, et al.
Pubblicazione: (2025)