Progressive-Resolution Policy Distillation: Leveraging Coarse-Resolution Simulations for Time-Efficient Fine-Resolution Policy Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kadokawa, Yuki, Tahara, Hirotaka, Matsubara, Takamitsu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BEAC: Imitating Complex Exploration and Task-oriented Behaviors for Invisible Object Nonprehensile Manipulation
by: Tahara, Hirotaka, et al.
Published: (2025)
by: Tahara, Hirotaka, et al.
Published: (2025)
Prolonging Tool Life: Learning Skillful Use of General-purpose Tools through Lifespan-guided Reinforcement Learning
by: Wu, Po-Yen, et al.
Published: (2025)
by: Wu, Po-Yen, et al.
Published: (2025)
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
by: Kadokawa, Yuki, et al.
Published: (2025)
by: Kadokawa, Yuki, et al.
Published: (2025)
Robust Sim-to-Real Cloth Untangling through Reduced-Resolution Observations via Adaptive Force-Difference Quantization
by: Tsurumine, Yoshihisa, et al.
Published: (2026)
by: Tsurumine, Yoshihisa, et al.
Published: (2026)
Composite Gaussian Processes Flows for Learning Discontinuous Multimodal Policies
by: Wang, Shu-yuan, et al.
Published: (2025)
by: Wang, Shu-yuan, et al.
Published: (2025)
Robust Iterative Value Conversion: Deep Reinforcement Learning for Neurochip-driven Edge Robots
by: Kadokawa, Yuki, et al.
Published: (2024)
by: Kadokawa, Yuki, et al.
Published: (2024)
Feasibility-aware Imitation Learning from Observations through a Hand-mounted Demonstration Interface
by: Takahashi, Kei, et al.
Published: (2025)
by: Takahashi, Kei, et al.
Published: (2025)
Self-Supervised Learning of Grasping Arbitrary Objects On-the-Move
by: Kiyokawa, Takuya, et al.
Published: (2024)
by: Kiyokawa, Takuya, et al.
Published: (2024)
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
by: Takahashi, Keiichiro, et al.
Published: (2024)
by: Takahashi, Keiichiro, et al.
Published: (2024)
DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport
by: Shibata, Kazuki, et al.
Published: (2026)
by: Shibata, Kazuki, et al.
Published: (2026)
Real Time Semantic Segmentation of High Resolution Automotive LiDAR Scans
by: Reichert, Hannes, et al.
Published: (2025)
by: Reichert, Hannes, et al.
Published: (2025)
Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications
by: Hatanaka, Wataru, et al.
Published: (2025)
by: Hatanaka, Wataru, et al.
Published: (2025)
ASBI: Leveraging Informative Real-World Data for Active Black-Box Simulator Tuning
by: Kim, Gahee, et al.
Published: (2025)
by: Kim, Gahee, et al.
Published: (2025)
Data-driven Probabilistic Trajectory Learning with High Temporal Resolution in Terminal Airspace
by: Xiang, Jun, et al.
Published: (2024)
by: Xiang, Jun, et al.
Published: (2024)
Scalable Multi-Objective Robot Reinforcement Learning through Gradient Conflict Resolution
by: Munn, Humphrey, et al.
Published: (2025)
by: Munn, Humphrey, et al.
Published: (2025)
Leveraging Demonstrator-perceived Precision for Safe Interactive Imitation Learning of Clearance-limited Tasks
by: Oh, Hanbit, et al.
Published: (2024)
by: Oh, Hanbit, et al.
Published: (2024)
Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning
by: Yin, Patrick, et al.
Published: (2025)
by: Yin, Patrick, et al.
Published: (2025)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
by: Wang, Zhendong, et al.
Published: (2024)
by: Wang, Zhendong, et al.
Published: (2024)
Score and Distribution Matching Policy: Advanced Accelerated Visuomotor Policies via Matched Distillation
by: Jia, Bofang, et al.
Published: (2024)
by: Jia, Bofang, et al.
Published: (2024)
MRS: Multi-Resolution Skills for HRL Agents
by: Sharma, Shashank, et al.
Published: (2025)
by: Sharma, Shashank, et al.
Published: (2025)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
by: Nakamura, Issa, et al.
Published: (2026)
by: Nakamura, Issa, et al.
Published: (2026)
Body Transformer: Leveraging Robot Embodiment for Policy Learning
by: Sferrazza, Carmelo, et al.
Published: (2024)
by: Sferrazza, Carmelo, et al.
Published: (2024)
Accelerating Visual-Policy Learning through Parallel Differentiable Simulation
by: You, Haoxiang, et al.
Published: (2025)
by: You, Haoxiang, et al.
Published: (2025)
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
by: Zhang, Tonghe, et al.
Published: (2025)
by: Zhang, Tonghe, et al.
Published: (2025)
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
by: Xu, Charles, et al.
Published: (2024)
by: Xu, Charles, et al.
Published: (2024)
Learning Hip Exoskeleton Control Policy via Predictive Neuromusculoskeletal Simulation
by: Park, Ilseung, et al.
Published: (2026)
by: Park, Ilseung, et al.
Published: (2026)
FDPP: Fine-tune Diffusion Policy with Human Preference
by: Chen, Yuxin, et al.
Published: (2025)
by: Chen, Yuxin, et al.
Published: (2025)
CoLF: Learning Consistent Leader-Follower Policies for Vision-Language-Guided Multi-Robot Cooperative Transport
by: Despature, Joachim Yann, et al.
Published: (2026)
by: Despature, Joachim Yann, et al.
Published: (2026)
Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies
by: Longhini, Alberta, et al.
Published: (2026)
by: Longhini, Alberta, et al.
Published: (2026)
CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies
by: Chen, Keyu, et al.
Published: (2026)
by: Chen, Keyu, et al.
Published: (2026)
Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning
by: Park, Byoungwoo, et al.
Published: (2026)
by: Park, Byoungwoo, et al.
Published: (2026)
MResT: Multi-Resolution Sensing for Real-Time Control with Vision-Language Models
by: Saxena, Saumya, et al.
Published: (2024)
by: Saxena, Saumya, et al.
Published: (2024)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
Flow-Based Single-Step Completion for Efficient and Expressive Policy Learning
by: Koirala, Prajwal, et al.
Published: (2025)
by: Koirala, Prajwal, et al.
Published: (2025)
Task-priority Intermediated Hierarchical Distributed Policies: Reinforcement Learning of Adaptive Multi-robot Cooperative Transport
by: Naito, Yusei, et al.
Published: (2024)
by: Naito, Yusei, et al.
Published: (2024)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
Similar Items
-
BEAC: Imitating Complex Exploration and Task-oriented Behaviors for Invisible Object Nonprehensile Manipulation
by: Tahara, Hirotaka, et al.
Published: (2025) -
Prolonging Tool Life: Learning Skillful Use of General-purpose Tools through Lifespan-guided Reinforcement Learning
by: Wu, Po-Yen, et al.
Published: (2025) -
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
by: Kadokawa, Yuki, et al.
Published: (2025) -
Robust Sim-to-Real Cloth Untangling through Reduced-Resolution Observations via Adaptive Force-Difference Quantization
by: Tsurumine, Yoshihisa, et al.
Published: (2026) -
Composite Gaussian Processes Flows for Learning Discontinuous Multimodal Policies
by: Wang, Shu-yuan, et al.
Published: (2025)