On Generalization and Distributional Update for Mimicking Observations with Adequate Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Yirui, Jin, Yunfei, Liu, Xiaowei, Zhang, Xiaofeng, Zhang, Yangchun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Reward Transferability in Adversarial Inverse Reinforcement Learning: Insights from Random Matrix Theory
by: Zhang, Yangchun, et al.
Published: (2024)
by: Zhang, Yangchun, et al.
Published: (2024)
Rethinking Adversarial Inverse Reinforcement Learning: Policy Imitation, Transferable Reward Recovery and Algebraic Equilibrium Proof
by: Zhang, Yangchun, et al.
Published: (2024)
by: Zhang, Yangchun, et al.
Published: (2024)
LSAM: Asynchronous Distributed Training with Landscape-Smoothed Sharpness-Aware Minimization
by: Teng, Yunfei, et al.
Published: (2025)
by: Teng, Yunfei, et al.
Published: (2025)
MimiC: Combating Client Dropouts in Federated Learning by Mimicking Central Updates
by: Sun, Yuchang, et al.
Published: (2023)
by: Sun, Yuchang, et al.
Published: (2023)
Mimicking Better by Matching the Approximate Action Distribution
by: Ramos, João A. Cândido, et al.
Published: (2023)
by: Ramos, João A. Cândido, et al.
Published: (2023)
Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management
by: Zhang, Runze, et al.
Published: (2025)
by: Zhang, Runze, et al.
Published: (2025)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
by: Zhang, Junkai, et al.
Published: (2024)
by: Zhang, Junkai, et al.
Published: (2024)
EvolveGen: Algorithmic Level Hardware Model Checking Benchmark Generation through Reinforcement Learning
by: Hu, Guangyu, et al.
Published: (2026)
by: Hu, Guangyu, et al.
Published: (2026)
Towards efficient quantum algorithms for diffusion probabilistic models
by: Wang, Yunfei, et al.
Published: (2025)
by: Wang, Yunfei, et al.
Published: (2025)
ShiftKD: Benchmarking Knowledge Distillation under Distribution Shift
by: Zhang, Songming, et al.
Published: (2023)
by: Zhang, Songming, et al.
Published: (2023)
A-IC3: Learning-Guided Adaptive Inductive Generalization for Hardware Model Checking
by: Zhou, Xiaofeng, et al.
Published: (2026)
by: Zhou, Xiaofeng, et al.
Published: (2026)
Exploration and Anti-Exploration with Distributional Random Network Distillation
by: Yang, Kai, et al.
Published: (2024)
by: Yang, Kai, et al.
Published: (2024)
From Function to Distribution Modeling: A PAC-Generative Approach to Offline Optimization
by: Zhang, Qiang, et al.
Published: (2024)
by: Zhang, Qiang, et al.
Published: (2024)
Efficient Incremental Belief Updates Using Weighted Virtual Observations
by: Tolpin, David
Published: (2024)
by: Tolpin, David
Published: (2024)
Invariant Correlation of Representation with Label: Enhancing Domain Generalization in Noisy Environments
by: Jin, Gaojie, et al.
Published: (2024)
by: Jin, Gaojie, et al.
Published: (2024)
Exploring the Impact of Parameter Update Magnitude on Forgetting and Generalization of Continual Learning
by: He, JinLi, et al.
Published: (2026)
by: He, JinLi, et al.
Published: (2026)
Exploration by Random Distribution Distillation
by: Fang, Zhirui, et al.
Published: (2025)
by: Fang, Zhirui, et al.
Published: (2025)
Deep Functional Factor Models: Forecasting High-Dimensional Functional Time Series via Bayesian Nonparametric Factorization
by: Liu, Yirui, et al.
Published: (2023)
by: Liu, Yirui, et al.
Published: (2023)
Learning to Simulate: Generative Metamodeling via Quantile Regression
by: Hong, L. Jeff, et al.
Published: (2023)
by: Hong, L. Jeff, et al.
Published: (2023)
Three Forms of Stochastic Injection for Improved Distribution-to-Distribution Generative Modeling
by: Su, Shiye, et al.
Published: (2025)
by: Su, Shiye, et al.
Published: (2025)
On the Suboptimality of GP-UCB under Polynomial Effective Optimism
by: Wang, Wenjia, et al.
Published: (2023)
by: Wang, Wenjia, et al.
Published: (2023)
Towards A Unified PAC-Bayesian Framework for Norm-based Generalization Bounds
by: Yi, Xinping, et al.
Published: (2026)
by: Yi, Xinping, et al.
Published: (2026)
Generalized Linear Bandits: Almost Optimal Regret with One-Pass Update
by: Zhang, Yu-Jie, et al.
Published: (2025)
by: Zhang, Yu-Jie, et al.
Published: (2025)
Optimal rates of approximation by shallow ReLU$^k$ neural networks and applications to nonparametric regression
by: Yang, Yunfei, et al.
Published: (2023)
by: Yang, Yunfei, et al.
Published: (2023)
Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction
by: Zu, Lipeng, et al.
Published: (2025)
by: Zu, Lipeng, et al.
Published: (2025)
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
Active Exploration via Autoregressive Generation of Missing Data
by: Cai, Tiffany Tianhui, et al.
Published: (2024)
by: Cai, Tiffany Tianhui, et al.
Published: (2024)
Semi-Supervised End-To-End Contrastive Learning For Time Series Classification
by: Cai, Huili, et al.
Published: (2023)
by: Cai, Huili, et al.
Published: (2023)
ODICE: Revealing the Mystery of Distribution Correction Estimation via Orthogonal-gradient Update
by: Mao, Liyuan, et al.
Published: (2024)
by: Mao, Liyuan, et al.
Published: (2024)
Guiding Diffusion Models with Reinforcement Learning for Stable Molecule Generation
by: Zhou, Zhijian, et al.
Published: (2025)
by: Zhou, Zhijian, et al.
Published: (2025)
Cold-Start Forecasting of New Product Life-Cycles via Conditional Diffusion Models
by: Zhou, Ruihan, et al.
Published: (2026)
by: Zhou, Ruihan, et al.
Published: (2026)
Distribution-Centric Policy Optimization Dominates Exploration-Exploitation Trade-off
by: Li, Zhaochun, et al.
Published: (2026)
by: Li, Zhaochun, et al.
Published: (2026)
Imitation from Observations with Trajectory-Level Generative Embeddings
by: Qu, Yongtao, et al.
Published: (2026)
by: Qu, Yongtao, et al.
Published: (2026)
Finite Volume-Informed Neural Network Framework for 2D Shallow Water Equations: Rugged Loss Landscapes and the Importance of Data Guidance
by: Liu, Xiaofeng
Published: (2026)
by: Liu, Xiaofeng
Published: (2026)
Sobolev norm inconsistency of kernel interpolation
by: Yang, Yunfei
Published: (2025)
by: Yang, Yunfei
Published: (2025)
Out-of-Distribution Generalization in Climate-Aware Yield Prediction with Earth Observation Data
by: Chakravarty, Aditya
Published: (2025)
by: Chakravarty, Aditya
Published: (2025)
Adversarial Label Invariant Graph Data Augmentations for Out-of-Distribution Generalization
by: Zhang, Simon, et al.
Published: (2026)
by: Zhang, Simon, et al.
Published: (2026)
Heavy-Tailed Linear Bandits: Huber Regression with One-Pass Update
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
UniGeM: Unifying Data Mixing and Selection via Geometric Exploration and Mining
by: Wang, Changhao, et al.
Published: (2026)
by: Wang, Changhao, et al.
Published: (2026)
A Survey on Evaluation of Out-of-Distribution Generalization
by: Yu, Han, et al.
Published: (2024)
by: Yu, Han, et al.
Published: (2024)
Similar Items
-
On Reward Transferability in Adversarial Inverse Reinforcement Learning: Insights from Random Matrix Theory
by: Zhang, Yangchun, et al.
Published: (2024) -
Rethinking Adversarial Inverse Reinforcement Learning: Policy Imitation, Transferable Reward Recovery and Algebraic Equilibrium Proof
by: Zhang, Yangchun, et al.
Published: (2024) -
LSAM: Asynchronous Distributed Training with Landscape-Smoothed Sharpness-Aware Minimization
by: Teng, Yunfei, et al.
Published: (2025) -
MimiC: Combating Client Dropouts in Federated Learning by Mimicking Central Updates
by: Sun, Yuchang, et al.
Published: (2023) -
Mimicking Better by Matching the Approximate Action Distribution
by: Ramos, João A. Cândido, et al.
Published: (2023)