Epistemic Exploration for Generalizable Planning and Learning in Non-Stationary Settings
Fuente:
arXiv
Guardado en:
| Autores principales: | Karia, Rushang, Verma, Pulkit, Speranzon, Alberto, Srivastava, Siddharth |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
$\forall$uto$\exists$val: Autonomous Assessment of LLMs in Formal Synthesis and Interpretation Tasks
por: Karia, Rushang, et al.
Publicado: (2024)
por: Karia, Rushang, et al.
Publicado: (2024)
Discovering and Learning Probabilistic Models of Black-Box AI Capabilities
por: Bramblett, Daniel, et al.
Publicado: (2025)
por: Bramblett, Daniel, et al.
Publicado: (2025)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
por: Karia, Rushang, et al.
Publicado: (2024)
por: Karia, Rushang, et al.
Publicado: (2024)
Using Explainable AI and Hierarchical Planning for Outreach with Robots
por: Karia, Rushang, et al.
Publicado: (2024)
por: Karia, Rushang, et al.
Publicado: (2024)
AI Planning: A Primer and Survey (Preliminary Report)
por: Chen, Dillon Z., et al.
Publicado: (2024)
por: Chen, Dillon Z., et al.
Publicado: (2024)
From Real World to Logic and Back: Learning Generalizable Relational Concepts For Long Horizon Robot Planning
por: Shah, Naman, et al.
Publicado: (2024)
por: Shah, Naman, et al.
Publicado: (2024)
Autonomous Option Invention for Continual Hierarchical Reinforcement Learning and Planning
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2024)
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2024)
Exploration and Adaptation in Non-Stationary Tasks with Diffusion Policies
por: Baveja, Gunbir Singh
Publicado: (2025)
por: Baveja, Gunbir Singh
Publicado: (2025)
Incentivized Exploration of Non-Stationary Stochastic Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2024)
por: Chakraborty, Sourav, et al.
Publicado: (2024)
Teaching LLMs to Plan: Logical Chain-of-Thought Instruction Tuning for Symbolic Planning
por: Verma, Pulkit, et al.
Publicado: (2025)
por: Verma, Pulkit, et al.
Publicado: (2025)
Indoor and Outdoor 3D Scene Graph Generation via Language-Enabled Spatial Ontologies
por: Strader, Jared, et al.
Publicado: (2023)
por: Strader, Jared, et al.
Publicado: (2023)
Depth-Bounded Epistemic Planning
por: Bolander, Thomas, et al.
Publicado: (2024)
por: Bolander, Thomas, et al.
Publicado: (2024)
Belief-State Query Policies for User-Aligned POMDPs
por: Bramblett, Daniel, et al.
Publicado: (2024)
por: Bramblett, Daniel, et al.
Publicado: (2024)
OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning
por: Srivastava, Siddharth, et al.
Publicado: (2025)
por: Srivastava, Siddharth, et al.
Publicado: (2025)
FAST-Q: Fast-track Exploration with Adversarially Balanced State Representations for Counterfactual Action Estimation in Offline Reinforcement Learning
por: Agrawal, Pulkit, et al.
Publicado: (2025)
por: Agrawal, Pulkit, et al.
Publicado: (2025)
In-Context Learning for Non-Stationary MIMO Equalization
por: Jiang, Jiachen, et al.
Publicado: (2025)
por: Jiang, Jiachen, et al.
Publicado: (2025)
Improving Intrinsic Exploration by Creating Stationary Objectives
por: Castanyer, Roger Creus, et al.
Publicado: (2023)
por: Castanyer, Roger Creus, et al.
Publicado: (2023)
SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition
por: Dey, Jayanta, et al.
Publicado: (2026)
por: Dey, Jayanta, et al.
Publicado: (2026)
Active Epistemic Control for Query-Efficient Verified Planning
por: Qu, Shuhui
Publicado: (2026)
por: Qu, Shuhui
Publicado: (2026)
The Epistemic Planning Domain Definition Language: Official Guideline
por: Burigana, Alessandro, et al.
Publicado: (2026)
por: Burigana, Alessandro, et al.
Publicado: (2026)
Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2025)
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2025)
Inverse Reinforcement Learning from Non-Stationary Learning Agents
por: Sivakumar, Kavinayan P., et al.
Publicado: (2024)
por: Sivakumar, Kavinayan P., et al.
Publicado: (2024)
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
por: Rustagi, Pulkit, et al.
Publicado: (2025)
por: Rustagi, Pulkit, et al.
Publicado: (2025)
Modeling Epistemic Uncertainty in Social Perception via Rashomon Set Agents
por: Yang, Jinming, et al.
Publicado: (2026)
por: Yang, Jinming, et al.
Publicado: (2026)
Multi-Label Transfer Learning in Non-Stationary Data Streams
por: Du, Honghui, et al.
Publicado: (2025)
por: Du, Honghui, et al.
Publicado: (2025)
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
por: Hamadanian, Pouya, et al.
Publicado: (2023)
por: Hamadanian, Pouya, et al.
Publicado: (2023)
Beyond Static Assumptions: the Predictive Justified Perspective Model for Epistemic Planning
por: Hu, Guang, et al.
Publicado: (2024)
por: Hu, Guang, et al.
Publicado: (2024)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
por: Gerogiannis, Argyrios, et al.
Publicado: (2024)
por: Gerogiannis, Argyrios, et al.
Publicado: (2024)
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
por: Galashov, Alexandre, et al.
Publicado: (2024)
por: Galashov, Alexandre, et al.
Publicado: (2024)
MARLINE: Multi-Source Mapping Transfer Learning for Non-Stationary Environments
por: Du, Honghui, et al.
Publicado: (2025)
por: Du, Honghui, et al.
Publicado: (2025)
Grounded World Model for Semantically Generalizable Planning
por: Li, Quanyi, et al.
Publicado: (2026)
por: Li, Quanyi, et al.
Publicado: (2026)
Non-Stationary Latent Auto-Regressive Bandits
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning
por: Wang, Tongxi, et al.
Publicado: (2026)
por: Wang, Tongxi, et al.
Publicado: (2026)
On Sample-Efficient Generalized Planning via Learned Transition Models
por: Gupta, Nitin, et al.
Publicado: (2026)
por: Gupta, Nitin, et al.
Publicado: (2026)
From Protoscience to Epistemic Monoculture: How Benchmarking Set the Stage for the Deep Learning Revolution
por: Koch, Bernard J., et al.
Publicado: (2024)
por: Koch, Bernard J., et al.
Publicado: (2024)
Learning Rate-Free Reinforcement Learning: A Case for Model Selection with Non-Stationary Objectives
por: Afshar, Aida, et al.
Publicado: (2024)
por: Afshar, Aida, et al.
Publicado: (2024)
Plan-MCTS: Plan Exploration for Action Exploitation in Web Navigation
por: Zhang, Weiming, et al.
Publicado: (2026)
por: Zhang, Weiming, et al.
Publicado: (2026)
An Epistemic and Aleatoric Decomposition of Arbitrariness to Constrain the Set of Good Models
por: Khan, Falaah Arif, et al.
Publicado: (2023)
por: Khan, Falaah Arif, et al.
Publicado: (2023)
Value Augmented Sampling for Language Model Alignment and Personalization
por: Han, Seungwook, et al.
Publicado: (2024)
por: Han, Seungwook, et al.
Publicado: (2024)
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
por: Luo, Baiting, et al.
Publicado: (2024)
por: Luo, Baiting, et al.
Publicado: (2024)
Ejemplares similares
-
$\forall$uto$\exists$val: Autonomous Assessment of LLMs in Formal Synthesis and Interpretation Tasks
por: Karia, Rushang, et al.
Publicado: (2024) -
Discovering and Learning Probabilistic Models of Black-Box AI Capabilities
por: Bramblett, Daniel, et al.
Publicado: (2025) -
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
por: Karia, Rushang, et al.
Publicado: (2024) -
Using Explainable AI and Hierarchical Planning for Outreach with Robots
por: Karia, Rushang, et al.
Publicado: (2024) -
AI Planning: A Primer and Survey (Preliminary Report)
por: Chen, Dillon Z., et al.
Publicado: (2024)