Saved in:
| Main Authors: | Chen, Yuhui, Li, Haoran, Zhao, Dongbin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2310.06343 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026)
by: Liu, Xin, et al.
Published: (2026)
Videos are Sample-Efficient Supervisions: Behavior Cloning from Videos via Latent Representations
by: Liu, Xin, et al.
Published: (2025)
by: Liu, Xin, et al.
Published: (2025)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
Cross-domain Random Pre-training with Prototypes for Reinforcement Learning
by: Liu, Xin, et al.
Published: (2023)
by: Liu, Xin, et al.
Published: (2023)
ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
by: Chen, Yuhui, et al.
Published: (2025)
by: Chen, Yuhui, et al.
Published: (2025)
Learning Future Representation with Synthetic Observations for Sample-efficient Reinforcement Learning
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
ComSD: Balancing Behavioral Quality and Diversity in Unsupervised Skill Discovery
by: Liu, Xin, et al.
Published: (2023)
by: Liu, Xin, et al.
Published: (2023)
SELU: Self-Learning Embodied MLLMs in Unknown Environments
by: Li, Boyu, et al.
Published: (2024)
by: Li, Boyu, et al.
Published: (2024)
Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement
by: Wang, Zhi, et al.
Published: (2024)
by: Wang, Zhi, et al.
Published: (2024)
ARAC: Adaptive Regularized Multi-Agent Soft Actor-Critic in Graph-Structured Adversarial Games
by: Shi, Ruochuan, et al.
Published: (2025)
by: Shi, Ruochuan, et al.
Published: (2025)
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability
by: Lu, Runyu, et al.
Published: (2025)
by: Lu, Runyu, et al.
Published: (2025)
Reward Models Identify Consistency, Not Causality
by: Xu, Yuhui, et al.
Published: (2025)
by: Xu, Yuhui, et al.
Published: (2025)
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
by: Fu, Yuqian, et al.
Published: (2026)
by: Fu, Yuqian, et al.
Published: (2026)
Equilibrium Policy Generalization: A Reinforcement Learning Framework for Cross-Graph Zero-Shot Generalization in Pursuit-Evasion Games
by: Lu, Runyu, et al.
Published: (2025)
by: Lu, Runyu, et al.
Published: (2025)
Deep learning for model correction of dynamical systems with data scarcity
by: Tatsuoka, Caroline, et al.
Published: (2024)
by: Tatsuoka, Caroline, et al.
Published: (2024)
Neural Predictive Control to Coordinate Discrete- and Continuous-Time Models for Time-Series Analysis with Control-Theoretical Improvements
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Consistency for Large Neural Networks: Regression and Classification
by: Zhan, Haoran, et al.
Published: (2024)
by: Zhan, Haoran, et al.
Published: (2024)
DUE: A Deep Learning Framework and Library for Modeling Unknown Equations
by: Chen, Junfeng, et al.
Published: (2025)
by: Chen, Junfeng, et al.
Published: (2025)
COPO: Consistency-Aware Policy Optimization
by: Han, Jinghang, et al.
Published: (2025)
by: Han, Jinghang, et al.
Published: (2025)
Categorical Policies: Multimodal Policy Learning and Exploration in Continuous Control
by: Islam, SM Mazharul, et al.
Published: (2025)
by: Islam, SM Mazharul, et al.
Published: (2025)
Benchmarking Smoothness and Reducing High-Frequency Oscillations in Continuous Control Policies
by: Christmann, Guilherme, et al.
Published: (2024)
by: Christmann, Guilherme, et al.
Published: (2024)
Fixed Design Analysis of Regularization-Based Continual Learning
by: Li, Haoran, et al.
Published: (2023)
by: Li, Haoran, et al.
Published: (2023)
Random Policy Evaluation Uncovers Policies of Generative Flow Networks
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
GaLore$+$: Boosting Low-Rank Adaptation for LLMs with Cross-Head Projection
by: Liao, Xutao, et al.
Published: (2024)
by: Liao, Xutao, et al.
Published: (2024)
A Training-Free Conditional Diffusion Model for Learning Stochastic Dynamical Systems
by: Liu, Yanfang, et al.
Published: (2024)
by: Liu, Yanfang, et al.
Published: (2024)
Multi-view Clustering via Bi-level Decoupling and Consistency Learning
by: Dong, Shihao, et al.
Published: (2025)
by: Dong, Shihao, et al.
Published: (2025)
Cross-Domain Policy Optimization via Bellman Consistency and Hybrid Critics
by: Chen, Ming-Hong, et al.
Published: (2026)
by: Chen, Ming-Hong, et al.
Published: (2026)
Temperature as a Meta-Policy: Adaptive Temperature in LLM Reinforcement Learning
by: Dang, Haoran, et al.
Published: (2026)
by: Dang, Haoran, et al.
Published: (2026)
CPIG: Leveraging Consistency Policy with Intention Guidance for Multi-agent Exploration
by: Fu, Yuqian, et al.
Published: (2024)
by: Fu, Yuqian, et al.
Published: (2024)
DiffEditor: Enhancing Speech Editing with Semantic Enrichment and Acoustic Consistency
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy
by: Xu, Kaixuan, et al.
Published: (2025)
by: Xu, Kaixuan, et al.
Published: (2025)
Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model
by: Tu, Songjun, et al.
Published: (2024)
by: Tu, Songjun, et al.
Published: (2024)
Multi-fidelity Parameter Estimation Using Conditional Diffusion Models
by: Tatsuoka, Caroline, et al.
Published: (2025)
by: Tatsuoka, Caroline, et al.
Published: (2025)
Continual Policy Distillation from Distributed Reinforcement Learning Teachers
by: Li, Yuxuan, et al.
Published: (2026)
by: Li, Yuxuan, et al.
Published: (2026)
Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control
by: Rahn, Nate, et al.
Published: (2023)
by: Rahn, Nate, et al.
Published: (2023)
Posterior Optimization with Clipped Objective for Bridging Efficiency and Stability in Generative Policy Learning
by: Chen, Yuhui, et al.
Published: (2026)
by: Chen, Yuhui, et al.
Published: (2026)
Data-driven Effective Modeling of Multiscale Stochastic Dynamical Systems
by: Chen, Yuan, et al.
Published: (2024)
by: Chen, Yuan, et al.
Published: (2024)
On-Policy Consistency Training Improves LLM Safety with Minimal Capability Degradation
by: Han, Andy, et al.
Published: (2026)
by: Han, Andy, et al.
Published: (2026)
Memory-Statistics Tradeoff in Continual Learning with Structural Regularization
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Similar Items
-
Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
by: Li, Haoran, et al.
Published: (2024) -
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026) -
Videos are Sample-Efficient Supervisions: Behavior Cloning from Videos via Latent Representations
by: Liu, Xin, et al.
Published: (2025) -
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024) -
Cross-domain Random Pre-training with Prototypes for Reinforcement Learning
by: Liu, Xin, et al.
Published: (2023)