Effective Exploration Based on the Structural Information Principles
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Xianghua, Peng, Hao, Li, Angsheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Decision Making Based on Structural Information Principles
by: Zeng, Xianghua, et al.
Published: (2024)
by: Zeng, Xianghua, et al.
Published: (2024)
Structural Information-based Hierarchical Diffusion for Offline Reinforcement Learning
by: Zeng, Xianghua, et al.
Published: (2025)
by: Zeng, Xianghua, et al.
Published: (2025)
RoBCtrl: Attacking GNN-Based Social Bot Detectors via Reinforced Manipulation of Bots Control Interaction
by: Yang, Yingguang, et al.
Published: (2025)
by: Yang, Yingguang, et al.
Published: (2025)
Robustness Evaluation of Graph-based News Detection Using Network Structural Information
by: Zeng, Xianghua, et al.
Published: (2025)
by: Zeng, Xianghua, et al.
Published: (2025)
Is Exploration All You Need? Effective Exploration Characteristics for Transfer in Reinforcement Learning
by: Balloch, Jonathan C., et al.
Published: (2024)
by: Balloch, Jonathan C., et al.
Published: (2024)
Empowering LLMs for Structure-Based Drug Design via Exploration-Augmented Latent Inference
by: Hu, Xuanning, et al.
Published: (2026)
by: Hu, Xuanning, et al.
Published: (2026)
A Minimum Variance Path Principle for Accurate and Stable Score-Based Density Ratio Estimation
by: Chen, Wei, et al.
Published: (2026)
by: Chen, Wei, et al.
Published: (2026)
Redundancy as a Structural Information Principle for Learning and Generalization
by: Bi, Yuda, et al.
Published: (2025)
by: Bi, Yuda, et al.
Published: (2025)
vLinear: A Powerful Linear Model for Multivariate Time Series Forecasting
by: Yue, Wenzhen, et al.
Published: (2026)
by: Yue, Wenzhen, et al.
Published: (2026)
Entropy-Aware Model Initialization for Effective Exploration in Deep Reinforcement Learning
by: Jang, Sooyoung, et al.
Published: (2021)
by: Jang, Sooyoung, et al.
Published: (2021)
The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning
by: Agarwal, Shivam, et al.
Published: (2025)
by: Agarwal, Shivam, et al.
Published: (2025)
Can LLMs Effectively Leverage Graph Structural Information through Prompts, and Why?
by: Huang, Jin, et al.
Published: (2023)
by: Huang, Jin, et al.
Published: (2023)
Dissecting the Failure of Invariant Learning on Graphs
by: Wang, Qixun, et al.
Published: (2024)
by: Wang, Qixun, et al.
Published: (2024)
Can In-context Learning Really Generalize to Out-of-distribution Tasks?
by: Wang, Qixun, et al.
Published: (2024)
by: Wang, Qixun, et al.
Published: (2024)
A Survey on Vulnerability of Federated Learning: A Learning Algorithm Perspective
by: Xie, Xianghua, et al.
Published: (2023)
by: Xie, Xianghua, et al.
Published: (2023)
SCALEX: Scalable Concept and Latent Exploration for Diffusion Models
by: Zeng, E. Zhixuan, et al.
Published: (2025)
by: Zeng, E. Zhixuan, et al.
Published: (2025)
From Implicit Exploration to Structured Reasoning: Leveraging Guideline and Refinement for LLMs
by: Chen, Jiaxiang, et al.
Published: (2025)
by: Chen, Jiaxiang, et al.
Published: (2025)
Local Causal Structure Learning in the Presence of Latent Variables
by: Xie, Feng, et al.
Published: (2024)
by: Xie, Feng, et al.
Published: (2024)
Value of Information-Enhanced Exploration in Bootstrapped DQN
by: Plataniotis, Stergios, et al.
Published: (2025)
by: Plataniotis, Stergios, et al.
Published: (2025)
Controllable Exploration in Hybrid-Policy RLVR for Multi-Modal Reasoning
by: Huang, Zhuoxu, et al.
Published: (2026)
by: Huang, Zhuoxu, et al.
Published: (2026)
Leveraging Invariant Principle for Heterophilic Graph Structure Distribution Shifts
by: Yang, Jinluan, et al.
Published: (2024)
by: Yang, Jinluan, et al.
Published: (2024)
Structured Exploration and Exploitation of Label Functions for Automated Data Annotation
by: Lam, Phong, et al.
Published: (2026)
by: Lam, Phong, et al.
Published: (2026)
Hypothesis Network Planned Exploration for Rapid Meta-Reinforcement Learning Adaptation
by: Jacobson, Maxwell Joseph, et al.
Published: (2023)
by: Jacobson, Maxwell Joseph, et al.
Published: (2023)
TreeAdv: Tree-Structured Advantage Redistribution for Group-Based RL
by: Cao, Lang, et al.
Published: (2026)
by: Cao, Lang, et al.
Published: (2026)
Offline Model-Based Reinforcement Learning with Anti-Exploration
by: Srinivasan, Padmanaba, et al.
Published: (2024)
by: Srinivasan, Padmanaba, et al.
Published: (2024)
A Pre-training Framework for Relational Data with Information-theoretic Principles
by: Truong, Quang, et al.
Published: (2025)
by: Truong, Quang, et al.
Published: (2025)
LLM-Explorer: A Plug-in Reinforcement Learning Policy Exploration Enhancement Driven by Large Language Models
by: Hao, Qianyue, et al.
Published: (2025)
by: Hao, Qianyue, et al.
Published: (2025)
Co-Exploration and Co-Exploitation via Shared Structure in Multi-Task Bandits
by: Mukherjee, Sumantrak, et al.
Published: (2025)
by: Mukherjee, Sumantrak, et al.
Published: (2025)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
M3-Net: A Cost-Effective Graph-Free MLP-Based Model for Traffic Prediction
by: Jin, Guangyin, et al.
Published: (2025)
by: Jin, Guangyin, et al.
Published: (2025)
LLMs for Text-Based Exploration and Navigation Under Partial Observability
by: Sandfuchs, Stephan, et al.
Published: (2026)
by: Sandfuchs, Stephan, et al.
Published: (2026)
Quality over Quantity: An Effective Large-Scale Data Reduction Strategy Based on Pointwise V-Information
by: Chen, Fei, et al.
Published: (2025)
by: Chen, Fei, et al.
Published: (2025)
PhySense: Principle-Based Physics Reasoning Benchmarking for Large Language Models
by: Xu, Yinggan, et al.
Published: (2025)
by: Xu, Yinggan, et al.
Published: (2025)
CacheClip: Accelerating RAG with Effective KV Cache Reuse
by: Yang, Bin, et al.
Published: (2025)
by: Yang, Bin, et al.
Published: (2025)
FreEformer: Frequency Enhanced Transformer for Multivariate Time Series Forecasting
by: Yue, Wenzhen, et al.
Published: (2025)
by: Yue, Wenzhen, et al.
Published: (2025)
MixKVQ: Query-Aware Mixed-Precision KV Cache Quantization for Long-Context Reasoning
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
Beyond High-Entropy Exploration: Correctness-Aware Low-Entropy Segment-Based Advantage Shaping for Reasoning LLMs
by: Chen, Xinzhu, et al.
Published: (2025)
by: Chen, Xinzhu, et al.
Published: (2025)
RED: Effective Trajectory Representation Learning with Comprehensive Information
by: Zhou, Silin, et al.
Published: (2024)
by: Zhou, Silin, et al.
Published: (2024)
Exploration Unbound
by: Arumugam, Dilip, et al.
Published: (2024)
by: Arumugam, Dilip, et al.
Published: (2024)
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning
by: Driss, Brahim, et al.
Published: (2025)
by: Driss, Brahim, et al.
Published: (2025)
Similar Items
-
Hierarchical Decision Making Based on Structural Information Principles
by: Zeng, Xianghua, et al.
Published: (2024) -
Structural Information-based Hierarchical Diffusion for Offline Reinforcement Learning
by: Zeng, Xianghua, et al.
Published: (2025) -
RoBCtrl: Attacking GNN-Based Social Bot Detectors via Reinforced Manipulation of Bots Control Interaction
by: Yang, Yingguang, et al.
Published: (2025) -
Robustness Evaluation of Graph-based News Detection Using Network Structural Information
by: Zeng, Xianghua, et al.
Published: (2025) -
Is Exploration All You Need? Effective Exploration Characteristics for Transfer in Reinforcement Learning
by: Balloch, Jonathan C., et al.
Published: (2024)