Task-Aware Exploration via a Predictive Bisimulation Metric
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Dayang, Liu, Ruihan, Wan, Lipeng, Liu, Yunlong, An, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Intrinsic Dynamics-Driven Generalizable Scene Representations for Vision-Oriented Decision-Making Applications
by: Liang, Dayang, et al.
Published: (2024)
by: Liang, Dayang, et al.
Published: (2024)
Episodic Reinforcement Learning with Expanded State-reward Space
by: Liang, Dayang, et al.
Published: (2024)
by: Liang, Dayang, et al.
Published: (2024)
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
by: Liang, Dayang, et al.
Published: (2026)
by: Liang, Dayang, et al.
Published: (2026)
ThanoRA: Task Heterogeneity-Aware Multi-Task Low-Rank Adaptation
by: Liang, Jian, et al.
Published: (2025)
by: Liang, Jian, et al.
Published: (2025)
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
by: Liu, Zeyang, et al.
Published: (2024)
by: Liu, Zeyang, et al.
Published: (2024)
A Generalized Bisimulation Metric of State Similarity between Markov Decision Processes: From Theoretical Propositions to Applications
by: Tao, Zhenyu, et al.
Published: (2025)
by: Tao, Zhenyu, et al.
Published: (2025)
Game Generation via Large Language Models
by: Hu, Chengpeng, et al.
Published: (2024)
by: Hu, Chengpeng, et al.
Published: (2024)
Stable Offline Value Function Learning with Bisimulation-based Representations
by: Pavse, Brahma S., et al.
Published: (2024)
by: Pavse, Brahma S., et al.
Published: (2024)
Scaling Synthetic Task Generation for Agents via Exploration
by: Ramrakhya, Ram, et al.
Published: (2025)
by: Ramrakhya, Ram, et al.
Published: (2025)
LASIL: Learner-Aware Supervised Imitation Learning For Long-term Microscopic Traffic Simulation
by: Guo, Ke, et al.
Published: (2024)
by: Guo, Ke, et al.
Published: (2024)
DeepThink3D: Enhancing Large Language Models with Programmatic Reasoning in Complex 3D Situated Reasoning Tasks
by: Song, Jiayi, et al.
Published: (2025)
by: Song, Jiayi, et al.
Published: (2025)
Boosting Meta-Learning for Few-Shot Text Classification via Label-guided Distance Scaling
by: Gao, Yunlong, et al.
Published: (2026)
by: Gao, Yunlong, et al.
Published: (2026)
SD-Net: Symmetric-Aware Keypoint Prediction and Domain Adaptation for 6D Pose Estimation In Bin-picking Scenarios
by: Huang, Ding-Tao, et al.
Published: (2024)
by: Huang, Ding-Tao, et al.
Published: (2024)
Selection, Reflection and Self-Refinement: Revisit Reasoning Tasks via a Causal Lens
by: Deng, Yunlong, et al.
Published: (2025)
by: Deng, Yunlong, et al.
Published: (2025)
Measuring and Analyzing Intelligence via Contextual Uncertainty in Large Language Models using Information-Theoretic Metrics
by: Shim, Jae Wan
Published: (2025)
by: Shim, Jae Wan
Published: (2025)
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
by: Chakroborti, Apu Kumar, et al.
Published: (2025)
by: Chakroborti, Apu Kumar, et al.
Published: (2025)
CALM: Consensus-Aware Localized Merging for Multi-Task Learning
by: Yan, Kunda, et al.
Published: (2025)
by: Yan, Kunda, et al.
Published: (2025)
Scene-Aware Explainable Multimodal Trajectory Prediction
by: Liu, Pei, et al.
Published: (2024)
by: Liu, Pei, et al.
Published: (2024)
Transport-Hub-Aware Spatial-Temporal Adaptive Graph Transformer for Traffic Flow Prediction
by: Xu, Xiao, et al.
Published: (2023)
by: Xu, Xiao, et al.
Published: (2023)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
by: Yue, Bo, et al.
Published: (2024)
by: Yue, Bo, et al.
Published: (2024)
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification
by: Qin, Minghao, et al.
Published: (2025)
by: Qin, Minghao, et al.
Published: (2025)
RDEx-CSOP: Feasibility-Aware Reconstructed Differential Evolution with Adaptive epsilon-Constraint Ranking
by: Tao, Sichen, et al.
Published: (2026)
by: Tao, Sichen, et al.
Published: (2026)
Autonomous Implicit Indoor Scene Reconstruction with Frontier Exploration
by: Zeng, Jing, et al.
Published: (2024)
by: Zeng, Jing, et al.
Published: (2024)
Learning to Explore: Scaling Agentic Reasoning via Exploration-Aware Policy Optimization
by: Hua, Xingyuan, et al.
Published: (2026)
by: Hua, Xingyuan, et al.
Published: (2026)
Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration
by: Chen, Weile, et al.
Published: (2026)
by: Chen, Weile, et al.
Published: (2026)
Joint Agent Memory and Exploration Learning via Novelty Signals
by: Tian, Shizuo, et al.
Published: (2026)
by: Tian, Shizuo, et al.
Published: (2026)
RDEx-CMOP: Feasibility-Aware Indicator-Guided Differential Evolution for Fixed-Budget Constrained Multiobjective Optimization
by: Tao, Sichen, et al.
Published: (2026)
by: Tao, Sichen, et al.
Published: (2026)
How Focused Are LLMs? A Quantitative Study via Repetitive Deterministic Prediction Tasks
by: Hou, Wanda, et al.
Published: (2025)
by: Hou, Wanda, et al.
Published: (2025)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
by: Wang, Jiaan, et al.
Published: (2024)
by: Wang, Jiaan, et al.
Published: (2024)
Exploration and Adaptation in Non-Stationary Tasks with Diffusion Policies
by: Baveja, Gunbir Singh
Published: (2025)
by: Baveja, Gunbir Singh
Published: (2025)
ProActor: Timing-Aware Reinforcement Learning for Proactive Task Scheduling Agents
by: Ding, Lei, et al.
Published: (2026)
by: Ding, Lei, et al.
Published: (2026)
GBPP: Grasp-Aware Base Placement Prediction for Robots via Two-Stage Learning
by: Chen, Jizhuo, et al.
Published: (2025)
by: Chen, Jizhuo, et al.
Published: (2025)
Towards Provably Unlearnable Examples via Bayes Error Optimisation
by: Zhang, Ruihan, et al.
Published: (2025)
by: Zhang, Ruihan, et al.
Published: (2025)
EchoTrail-GUI: Building Actionable Memory for GUI Agents via Critic-Guided Self-Exploration
by: Li, Runze, et al.
Published: (2025)
by: Li, Runze, et al.
Published: (2025)
ScreenSearch: Uncertainty-Aware OS Exploration
by: Solodko, Michael, et al.
Published: (2026)
by: Solodko, Michael, et al.
Published: (2026)
Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing
by: Li, Zhuoying, et al.
Published: (2025)
by: Li, Zhuoying, et al.
Published: (2025)
BEE: Metric-Adapted Explanations via Baseline Exploration-Exploitation
by: Barkan, Oren, et al.
Published: (2024)
by: Barkan, Oren, et al.
Published: (2024)
One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models
by: Shao, Hang, et al.
Published: (2023)
by: Shao, Hang, et al.
Published: (2023)
Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness
by: Fan, Fulong, et al.
Published: (2026)
by: Fan, Fulong, et al.
Published: (2026)
OPRIDE: Offline Preference-based Reinforcement Learning via In-Dataset Exploration
by: Yang, Yiqin, et al.
Published: (2026)
by: Yang, Yiqin, et al.
Published: (2026)
Similar Items
-
Intrinsic Dynamics-Driven Generalizable Scene Representations for Vision-Oriented Decision-Making Applications
by: Liang, Dayang, et al.
Published: (2024) -
Episodic Reinforcement Learning with Expanded State-reward Space
by: Liang, Dayang, et al.
Published: (2024) -
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
by: Liang, Dayang, et al.
Published: (2026) -
ThanoRA: Task Heterogeneity-Aware Multi-Task Low-Rank Adaptation
by: Liang, Jian, et al.
Published: (2025) -
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
by: Liu, Zeyang, et al.
Published: (2024)