Curiosity is Knowledge: Self-Consistent Learning and No-Regret Optimization with Active Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yingke, Parashar, Anjali, Zhou, Enlu, Fan, Chuchu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference
by: Li, Yingke, et al.
Published: (2026)
by: Li, Yingke, et al.
Published: (2026)
Online Bayesian Risk-Averse Reinforcement Learning
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Bayesian Risk-Sensitive Policy Optimization For MDPs With General Loss Functions
by: Wang, Xiaoshuang, et al.
Published: (2025)
by: Wang, Xiaoshuang, et al.
Published: (2025)
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
by: Song, Meichen, et al.
Published: (2026)
by: Song, Meichen, et al.
Published: (2026)
Ranking and Selection with Simultaneous Input Data Collection
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Adaptive Simulation Experiment for LLM Policy Optimization
by: Hu, Mingjie, et al.
Published: (2026)
by: Hu, Mingjie, et al.
Published: (2026)
Bayesian Risk-averse Model Predictive Control with Consistency and Stability Guarantees
by: Li, Yingke, et al.
Published: (2025)
by: Li, Yingke, et al.
Published: (2025)
RADIUM: Predicting and Repairing End-to-End Robot Failures using Gradient-Accelerated Sampling
by: Dawson, Charles, et al.
Published: (2024)
by: Dawson, Charles, et al.
Published: (2024)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
by: Lin, Yifan, et al.
Published: (2024)
by: Lin, Yifan, et al.
Published: (2024)
Heuristic Search as Language-Guided Program Optimization
by: Yu, Mingxin, et al.
Published: (2026)
by: Yu, Mingxin, et al.
Published: (2026)
Low-Regret and Low-Complexity Learning for Hierarchical Inference
by: Chattopadhyay, Sameep, et al.
Published: (2025)
by: Chattopadhyay, Sameep, et al.
Published: (2025)
Curiosity Driven Exploration to Optimize Structure-Property Learning in Microscopy
by: Vatsavai, Aditya, et al.
Published: (2025)
by: Vatsavai, Aditya, et al.
Published: (2025)
SEED-SET: Scalable Evolving Experimental Design for System-level Ethical Testing
by: Parashar, Anjali, et al.
Published: (2026)
by: Parashar, Anjali, et al.
Published: (2026)
Solving Minimum-Cost Reach Avoid using Reinforcement Learning
by: So, Oswin, et al.
Published: (2024)
by: So, Oswin, et al.
Published: (2024)
Active Disruption Avoidance and Trajectory Design for Tokamak Ramp-downs with Neural Differential Equations and Reinforcement Learning
by: Wang, Allen M., et al.
Published: (2024)
by: Wang, Allen M., et al.
Published: (2024)
Alternating Regret for Online Convex Optimization
by: Hait, Soumita, et al.
Published: (2025)
by: Hait, Soumita, et al.
Published: (2025)
Curiosity-Driven Development of Action and Language in Robots Through Self-Exploration
by: Tinker, Theodore Jerome, et al.
Published: (2025)
by: Tinker, Theodore Jerome, et al.
Published: (2025)
Direct Regret Optimization in Bayesian Optimization
by: Zhang, Fengxue, et al.
Published: (2025)
by: Zhang, Fengxue, et al.
Published: (2025)
TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks
by: Meng, Yue, et al.
Published: (2025)
by: Meng, Yue, et al.
Published: (2025)
SGD with Dependent Data: Optimal Estimation, Regret, and Inference
by: Shen, Yinan, et al.
Published: (2026)
by: Shen, Yinan, et al.
Published: (2026)
Failure Prediction from Limited Hardware Demonstrations
by: Parashar, Anjali, et al.
Published: (2024)
by: Parashar, Anjali, et al.
Published: (2024)
Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability
by: Liu, Zihao, et al.
Published: (2025)
by: Liu, Zihao, et al.
Published: (2025)
Self-Normalized Martingales and Uniform Regret Bounds for Linear Regression
by: Chen, Fan, et al.
Published: (2026)
by: Chen, Fan, et al.
Published: (2026)
Self-Correcting Bayesian Optimization through Bayesian Active Learning
by: Hvarfner, Carl, et al.
Published: (2023)
by: Hvarfner, Carl, et al.
Published: (2023)
Active Context Selection Improves Simple Regret in Contextual Bandits
by: Shahverdikondori, Mohammad, et al.
Published: (2026)
by: Shahverdikondori, Mohammad, et al.
Published: (2026)
Meta-Learning in Self-Play Regret Minimization
by: Sychrovský, David, et al.
Published: (2025)
by: Sychrovský, David, et al.
Published: (2025)
Diversifying Policy Behaviors with Extrinsic Behavioral Curiosity
by: Wan, Zhenglin, et al.
Published: (2024)
by: Wan, Zhenglin, et al.
Published: (2024)
An Information-Geometric Approach to Artificial Curiosity
by: Nedergaard, Alexander, et al.
Published: (2025)
by: Nedergaard, Alexander, et al.
Published: (2025)
Rare event modeling with self-regularized normalizing flows: what can we learn from a single failure?
by: Dawson, Charles, et al.
Published: (2025)
by: Dawson, Charles, et al.
Published: (2025)
Learning U-Statistics with Active Inference
by: Wang, Xiaoning, et al.
Published: (2026)
by: Wang, Xiaoning, et al.
Published: (2026)
Active and Passive Causal Inference Learning
by: Im, Daniel Jiwoong, et al.
Published: (2023)
by: Im, Daniel Jiwoong, et al.
Published: (2023)
Robust Amortized Bayesian Inference with Self-Consistency Losses on Unlabeled Data
by: Mishra, Aayush, et al.
Published: (2025)
by: Mishra, Aayush, et al.
Published: (2025)
Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
by: Zhang, Songyuan, et al.
Published: (2025)
by: Zhang, Songyuan, et al.
Published: (2025)
Efficient Skill Discovery via Regret-Aware Optimization
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
Optimal Regret for Policy Optimization in Contextual Bandits
by: Levy, Orin, et al.
Published: (2026)
by: Levy, Orin, et al.
Published: (2026)
On Regret Bounds of Thompson Sampling for Bayesian Optimization
by: Takeno, Shion, et al.
Published: (2026)
by: Takeno, Shion, et al.
Published: (2026)
Regret Minimization via Saddle Point Optimization
by: Kirschner, Johannes, et al.
Published: (2024)
by: Kirschner, Johannes, et al.
Published: (2024)
Stopping Bayesian Optimization with Probabilistic Regret Bounds
by: Wilson, James T.
Published: (2024)
by: Wilson, James T.
Published: (2024)
Beyond Regrets: Geometric Metrics for Bayesian Optimization
by: Kim, Jungtaek
Published: (2024)
by: Kim, Jungtaek
Published: (2024)
TeLoGraF: Temporal Logic Planning via Graph-encoded Flow Matching
by: Meng, Yue, et al.
Published: (2025)
by: Meng, Yue, et al.
Published: (2025)
Similar Items
-
Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference
by: Li, Yingke, et al.
Published: (2026) -
Online Bayesian Risk-Averse Reinforcement Learning
by: Wang, Yuhao, et al.
Published: (2025) -
Bayesian Risk-Sensitive Policy Optimization For MDPs With General Loss Functions
by: Wang, Xiaoshuang, et al.
Published: (2025) -
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
by: Song, Meichen, et al.
Published: (2026) -
Ranking and Selection with Simultaneous Input Data Collection
by: Wang, Yuhao, et al.
Published: (2025)