Online Feedback Efficient Active Target Discovery in Partially Observable Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Sarkar, Anindya, Ji, Binglin, Vorobeychik, Yevgeniy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Active Target Discovery under Uninformative Prior: The Power of Permanent and Transient Memory
by: Sarkar, Anindya, et al.
Published: (2025)
by: Sarkar, Anindya, et al.
Published: (2025)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
by: Sarkar, Anindya, et al.
Published: (2026)
by: Sarkar, Anindya, et al.
Published: (2026)
Active Geospatial Search for Efficient Tenant Eviction Outreach
by: Sarkar, Anindya, et al.
Published: (2024)
by: Sarkar, Anindya, et al.
Published: (2024)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
by: Lanier, Michael, et al.
Published: (2024)
by: Lanier, Michael, et al.
Published: (2024)
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
by: Ge, Luise, et al.
Published: (2025)
by: Ge, Luise, et al.
Published: (2025)
Verified Safe Reinforcement Learning for Neural Network Dynamic Models
by: Wu, Junlin, et al.
Published: (2024)
by: Wu, Junlin, et al.
Published: (2024)
CoFineLLM: Conformal Finetuning of LLMs for Language-Instructed Robot Planning
by: Wang, Jun, et al.
Published: (2025)
by: Wang, Jun, et al.
Published: (2025)
Learning Linear Utility Functions From Pairwise Comparison Queries
by: Ge, Luise, et al.
Published: (2024)
by: Ge, Luise, et al.
Published: (2024)
Multi-Agent Reinforcement Learning for Assessing False-Data Injection Attacks on Transportation Networks
by: Eghtesad, Taha, et al.
Published: (2023)
by: Eghtesad, Taha, et al.
Published: (2023)
Attacks on Node Attributes in Graph Neural Networks
by: Xu, Ying, et al.
Published: (2024)
by: Xu, Ying, et al.
Published: (2024)
GOMAA-Geo: GOal Modality Agnostic Active Geo-localization
by: Sarkar, Anindya, et al.
Published: (2024)
by: Sarkar, Anindya, et al.
Published: (2024)
Near-Optimal Partially Observable Reinforcement Learning with Partial Online State Information
by: Shi, Ming, et al.
Published: (2023)
by: Shi, Ming, et al.
Published: (2023)
Axioms for AI Alignment from Human Feedback
by: Ge, Luise, et al.
Published: (2024)
by: Ge, Luise, et al.
Published: (2024)
An Empirical Study on the Power of Future Prediction in Partially Observable Environments
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
Temporal Knowledge-Graph Memory in a Partially Observable Environment
by: Kim, Taewoon, et al.
Published: (2024)
by: Kim, Taewoon, et al.
Published: (2024)
Ontology-Enhanced Decision-Making for Autonomous Agents in Dynamic and Partially Observable Environments
by: Ghanadbashi, Saeedeh, et al.
Published: (2024)
by: Ghanadbashi, Saeedeh, et al.
Published: (2024)
Provable Representation with Efficient Planning for Partial Observable Reinforcement Learning
by: Zhang, Hongming, et al.
Published: (2023)
by: Zhang, Hongming, et al.
Published: (2023)
Transformer-Based Reinforcement Learning for Autonomous Orbital Collision Avoidance in Partially Observable Environments
by: Georges, Thomas, et al.
Published: (2026)
by: Georges, Thomas, et al.
Published: (2026)
Preference Poisoning Attacks on Reward Model Learning
by: Wu, Junlin, et al.
Published: (2024)
by: Wu, Junlin, et al.
Published: (2024)
When Your AIs Deceive You: Challenges of Partial Observability in Reinforcement Learning from Human Feedback
by: Lang, Leon, et al.
Published: (2024)
by: Lang, Leon, et al.
Published: (2024)
Zero-Shot Reinforcement Learning Under Partial Observability
by: Jeen, Scott, et al.
Published: (2025)
by: Jeen, Scott, et al.
Published: (2025)
Multi-View Causal Representation Learning with Partial Observability
by: Yao, Dingling, et al.
Published: (2023)
by: Yao, Dingling, et al.
Published: (2023)
Online Learning for Multi-Layer Hierarchical Inference under Partial and Policy-Dependent Feedback
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
Guided Policy Optimization under Partial Observability
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
LLMs for Text-Based Exploration and Navigation Under Partial Observability
by: Sandfuchs, Stephan, et al.
Published: (2026)
by: Sandfuchs, Stephan, et al.
Published: (2026)
A Sparsity Principle for Partially Observable Causal Representation Learning
by: Xu, Danru, et al.
Published: (2024)
by: Xu, Danru, et al.
Published: (2024)
Conformal Reachability for Safe Control in Unknown Environments
by: Ma, Xinhang, et al.
Published: (2026)
by: Ma, Xinhang, et al.
Published: (2026)
Partially Observable Gaussian Process Network and Doubly Stochastic Variational Inference
by: Kiroriwal, Saksham, et al.
Published: (2025)
by: Kiroriwal, Saksham, et al.
Published: (2025)
Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability
by: Kiram, Firas Mohamed Elamine, et al.
Published: (2026)
by: Kiram, Firas Mohamed Elamine, et al.
Published: (2026)
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
by: Ahuja, Angad Singh
Published: (2026)
by: Ahuja, Angad Singh
Published: (2026)
Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy
by: Allen, Cameron, et al.
Published: (2024)
by: Allen, Cameron, et al.
Published: (2024)
Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning
by: Zhao, Yike, et al.
Published: (2026)
by: Zhao, Yike, et al.
Published: (2026)
Efficiently Aligning Language Models with Online Natural Language Feedback
by: Ye, Christine, et al.
Published: (2026)
by: Ye, Christine, et al.
Published: (2026)
Cost Efficient Fairness Audit Under Partial Feedback
by: Das, Nirjhar, et al.
Published: (2025)
by: Das, Nirjhar, et al.
Published: (2025)
Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains
by: Tao, Ruo Yu, et al.
Published: (2025)
by: Tao, Ruo Yu, et al.
Published: (2025)
Belief States for Cooperative Multi-Agent Reinforcement Learning under Partial Observability
by: Pritz, Paul J., et al.
Published: (2025)
by: Pritz, Paul J., et al.
Published: (2025)
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability
by: Wang, Wuhao, et al.
Published: (2025)
by: Wang, Wuhao, et al.
Published: (2025)
On the Role of Information Structure in Reinforcement Learning for Partially-Observable Sequential Teams and Games
by: Altabaa, Awni, et al.
Published: (2024)
by: Altabaa, Awni, et al.
Published: (2024)
A Scalable Approach to Solving Simulation-Based Network Security Games
by: Lanier, Michael, et al.
Published: (2026)
by: Lanier, Michael, et al.
Published: (2026)
Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability
by: Kim, Taewoon, et al.
Published: (2026)
by: Kim, Taewoon, et al.
Published: (2026)
Similar Items
-
Active Target Discovery under Uninformative Prior: The Power of Permanent and Transient Memory
by: Sarkar, Anindya, et al.
Published: (2025) -
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
by: Sarkar, Anindya, et al.
Published: (2026) -
Active Geospatial Search for Efficient Tenant Eviction Outreach
by: Sarkar, Anindya, et al.
Published: (2024) -
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
by: Lanier, Michael, et al.
Published: (2024) -
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
by: Ge, Luise, et al.
Published: (2025)