Meta-Reinforcement Learning with Discrete World Models for Adaptive Load Balancing
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Redovian, Cameron |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Next Token Prediction to (STRIPS) World Models
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2025)
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2025)
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
Learning How to Cube
von: Erata, Ferhat, et al.
Veröffentlicht: (2026)
von: Erata, Ferhat, et al.
Veröffentlicht: (2026)
Predicting Future Actions of Reinforcement Learning Agents
von: Chung, Stephen, et al.
Veröffentlicht: (2024)
von: Chung, Stephen, et al.
Veröffentlicht: (2024)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
von: Aichmüller, Michael, et al.
Veröffentlicht: (2024)
von: Aichmüller, Michael, et al.
Veröffentlicht: (2024)
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
Foundational Requirements for Artificial General Intelligence: A Falsifiable Framework Based on Signal Prediction
von: Šprogar, Matej
Veröffentlicht: (2025)
von: Šprogar, Matej
Veröffentlicht: (2025)
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
von: Silue, Bram, et al.
Veröffentlicht: (2025)
von: Silue, Bram, et al.
Veröffentlicht: (2025)
Social Interpretable Reinforcement Learning
von: Custode, Leonardo Lucio, et al.
Veröffentlicht: (2024)
von: Custode, Leonardo Lucio, et al.
Veröffentlicht: (2024)
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
Learning to Select Goals in Automated Planning with Deep-Q Learning
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2024)
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2024)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
von: Yousaf, Iqra
Veröffentlicht: (2024)
von: Yousaf, Iqra
Veröffentlicht: (2024)
Controlled Territory and Conflict Tracking (CONTACT): (Geo-)Mapping Occupied Territory from Open Source Intelligence
von: Mandal, Paul K., et al.
Veröffentlicht: (2025)
von: Mandal, Paul K., et al.
Veröffentlicht: (2025)
NeSIG: A Neuro-Symbolic Method for Learning to Generate Planning Problems
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
von: de Mol, Barbera, et al.
Veröffentlicht: (2025)
von: de Mol, Barbera, et al.
Veröffentlicht: (2025)
Constrained Auto-Bidding via Generative Response Modeling
von: Yang, Eunseok, et al.
Veröffentlicht: (2026)
von: Yang, Eunseok, et al.
Veröffentlicht: (2026)
A Review of Symbolic, Subsymbolic and Hybrid Methods for Sequential Decision Making
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
What Do World Models Learn in RL? Probing Latent Representations in Learned Environment Simulators
von: Zhang, Xinyu
Veröffentlicht: (2026)
von: Zhang, Xinyu
Veröffentlicht: (2026)
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
von: Rottenwalter, Georg, et al.
Veröffentlicht: (2025)
von: Rottenwalter, Georg, et al.
Veröffentlicht: (2025)
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
von: Priezzhev, I. I., et al.
Veröffentlicht: (2025)
von: Priezzhev, I. I., et al.
Veröffentlicht: (2025)
Advancements in synthetic data extraction for industrial injection molding
von: Rottenwalter, Georg, et al.
Veröffentlicht: (2025)
von: Rottenwalter, Georg, et al.
Veröffentlicht: (2025)
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
von: Wang, Caroline, et al.
Veröffentlicht: (2025)
von: Wang, Caroline, et al.
Veröffentlicht: (2025)
LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models
von: Raman, Vishal, et al.
Veröffentlicht: (2025)
von: Raman, Vishal, et al.
Veröffentlicht: (2025)
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
von: George, Robert Joseph, et al.
Veröffentlicht: (2025)
von: George, Robert Joseph, et al.
Veröffentlicht: (2025)
Uncovering Bugs in Formal Explainers: A Case Study with PyXAI
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2025)
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning with Preference-based Constraint Inference
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
von: Pinchuk, Mykola
Veröffentlicht: (2026)
von: Pinchuk, Mykola
Veröffentlicht: (2026)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites
von: Zhang, Qinshi, et al.
Veröffentlicht: (2026)
von: Zhang, Qinshi, et al.
Veröffentlicht: (2026)
IntSat: Integer Linear Programming by Conflict-Driven Constraint-Learning
von: Nieuwenhuis, Robert, et al.
Veröffentlicht: (2024)
von: Nieuwenhuis, Robert, et al.
Veröffentlicht: (2024)
Exact Synthetic Populations for Scalable Societal and Market Modeling
von: Petit, Thierry, et al.
Veröffentlicht: (2025)
von: Petit, Thierry, et al.
Veröffentlicht: (2025)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
von: Schipper, Olivier, et al.
Veröffentlicht: (2025)
von: Schipper, Olivier, et al.
Veröffentlicht: (2025)
A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification
von: Odmark, Joshua, et al.
Veröffentlicht: (2026)
von: Odmark, Joshua, et al.
Veröffentlicht: (2026)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
N-Agent Ad Hoc Teamwork
von: Wang, Caroline, et al.
Veröffentlicht: (2024)
von: Wang, Caroline, et al.
Veröffentlicht: (2024)
Sim-to-reality adaptation for Deep Reinforcement Learning applied to an underwater docking application
von: Chaarani, Alaaeddine, et al.
Veröffentlicht: (2026)
von: Chaarani, Alaaeddine, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From Next Token Prediction to (STRIPS) World Models
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2025) -
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025) -
Learning How to Cube
von: Erata, Ferhat, et al.
Veröffentlicht: (2026) -
Predicting Future Actions of Reinforcement Learning Agents
von: Chung, Stephen, et al.
Veröffentlicht: (2024) -
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
von: Aichmüller, Michael, et al.
Veröffentlicht: (2024)