Robust Iterative Value Conversion: Deep Reinforcement Learning for Neurochip-driven Edge Robots
Fuente:
arXiv
Saved in:
| Main Authors: | Kadokawa, Yuki, Kodera, Tomohito, Tsurumine, Yoshihisa, Nishimura, Shinya, Matsubara, Takamitsu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Sim-to-Real Cloth Untangling through Reduced-Resolution Observations via Adaptive Force-Difference Quantization
by: Tsurumine, Yoshihisa, et al.
Published: (2026)
by: Tsurumine, Yoshihisa, et al.
Published: (2026)
DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport
by: Shibata, Kazuki, et al.
Published: (2026)
by: Shibata, Kazuki, et al.
Published: (2026)
Progressive-Resolution Policy Distillation: Leveraging Coarse-Resolution Simulations for Time-Efficient Fine-Resolution Policy Learning
by: Kadokawa, Yuki, et al.
Published: (2024)
by: Kadokawa, Yuki, et al.
Published: (2024)
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
by: Kadokawa, Yuki, et al.
Published: (2025)
by: Kadokawa, Yuki, et al.
Published: (2025)
Cooperative Grasping and Transportation using Multi-agent Reinforcement Learning with Ternary Force Representation
by: Bernard-Tiong, Ing-Sheng, et al.
Published: (2024)
by: Bernard-Tiong, Ing-Sheng, et al.
Published: (2024)
Self-Supervised Learning of Grasping Arbitrary Objects On-the-Move
by: Kiyokawa, Takuya, et al.
Published: (2024)
by: Kiyokawa, Takuya, et al.
Published: (2024)
Prolonging Tool Life: Learning Skillful Use of General-purpose Tools through Lifespan-guided Reinforcement Learning
by: Wu, Po-Yen, et al.
Published: (2025)
by: Wu, Po-Yen, et al.
Published: (2025)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
by: Nakamura, Issa, et al.
Published: (2026)
by: Nakamura, Issa, et al.
Published: (2026)
Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications
by: Hatanaka, Wataru, et al.
Published: (2025)
by: Hatanaka, Wataru, et al.
Published: (2025)
Leveraging Demonstrator-perceived Precision for Safe Interactive Imitation Learning of Clearance-limited Tasks
by: Oh, Hanbit, et al.
Published: (2024)
by: Oh, Hanbit, et al.
Published: (2024)
CoLF: Learning Consistent Leader-Follower Policies for Vision-Language-Guided Multi-Robot Cooperative Transport
by: Despature, Joachim Yann, et al.
Published: (2026)
by: Despature, Joachim Yann, et al.
Published: (2026)
Reinforcement Learning of Multi-robot Task Allocation for Multi-object Transportation with Infeasible Tasks
by: Shida, Yuma, et al.
Published: (2024)
by: Shida, Yuma, et al.
Published: (2024)
Task-priority Intermediated Hierarchical Distributed Policies: Reinforcement Learning of Adaptive Multi-robot Cooperative Transport
by: Naito, Yusei, et al.
Published: (2024)
by: Naito, Yusei, et al.
Published: (2024)
Incipient Slip Detection by Vibration Injection into Soft Sensor
by: Komeno, Naoto, et al.
Published: (2024)
by: Komeno, Naoto, et al.
Published: (2024)
Disentangled Iterative Surface Fitting for Contact-stable Grasp Planning
by: Yamanokuchi, Tomoya, et al.
Published: (2025)
by: Yamanokuchi, Tomoya, et al.
Published: (2025)
Feasibility-aware Imitation Learning from Observation with Multimodal Feedback
by: Takahashi, Kei, et al.
Published: (2026)
by: Takahashi, Kei, et al.
Published: (2026)
DecompGrind: A Decomposition Framework for Robotic Grinding via Cutting-Surface Planning and Contact-Force Adaptation
by: Araki, Shunsuke, et al.
Published: (2026)
by: Araki, Shunsuke, et al.
Published: (2026)
ASBI: Leveraging Informative Real-World Data for Active Black-Box Simulator Tuning
by: Kim, Gahee, et al.
Published: (2025)
by: Kim, Gahee, et al.
Published: (2025)
BEAC: Imitating Complex Exploration and Task-oriented Behaviors for Invisible Object Nonprehensile Manipulation
by: Tahara, Hirotaka, et al.
Published: (2025)
by: Tahara, Hirotaka, et al.
Published: (2025)
Where Do We Look When We Teach? Analyzing Human Gaze Behavior Across Demonstration Devices in Robot Imitation Learning
by: Ishida, Yutaro, et al.
Published: (2025)
by: Ishida, Yutaro, et al.
Published: (2025)
Feasibility-aware Imitation Learning from Observations through a Hand-mounted Demonstration Interface
by: Takahashi, Kei, et al.
Published: (2025)
by: Takahashi, Kei, et al.
Published: (2025)
Tracing Energy Flow: Learning Tactile-based Grasping Force Control to Prevent Slippage in Dynamic Object Interaction
by: Kuo, Cheng-Yu, et al.
Published: (2025)
by: Kuo, Cheng-Yu, et al.
Published: (2025)
Learning Quiet Walking for a Small Home Robot
by: Watanabe, Ryo, et al.
Published: (2025)
by: Watanabe, Ryo, et al.
Published: (2025)
Cutting Sequence Diffuser: Sim-to-Real Transferable Planning for Object Shaping by Grinding
by: Hachimine, Takumi, et al.
Published: (2024)
by: Hachimine, Takumi, et al.
Published: (2024)
Composite Gaussian Processes Flows for Learning Discontinuous Multimodal Policies
by: Wang, Shu-yuan, et al.
Published: (2025)
by: Wang, Shu-yuan, et al.
Published: (2025)
Domains as Objectives: Domain-Uncertainty-Aware Policy Optimization through Explicit Multi-Domain Convex Coverage Set Learning
by: Ilboudo, Wendyam Eric Lionel, et al.
Published: (2024)
by: Ilboudo, Wendyam Eric Lionel, et al.
Published: (2024)
DISF: Disentangled Iterative Surface Fitting for Contact-stable Grasp Planning with Grasp Pose Alignment to the Object Center of Mass
by: Yamanokuchi, Tomoya, et al.
Published: (2025)
by: Yamanokuchi, Tomoya, et al.
Published: (2025)
OIPP: Object-Adaptive Impact Point Predictor for Catching Diverse In-Flight Objects
by: Nguyen, Ngoc Huy, et al.
Published: (2025)
by: Nguyen, Ngoc Huy, et al.
Published: (2025)
ICCO: Learning an Instruction-conditioned Coordinator for Language-guided Task-aligned Multi-robot Control
by: Yano, Yoshiki, et al.
Published: (2025)
by: Yano, Yoshiki, et al.
Published: (2025)
Task-Relevant and Irrelevant Region-Aware Augmentation for Generalizable Vision-Based Imitation Learning in Agricultural Manipulation
by: Hattori, Shun, et al.
Published: (2026)
by: Hattori, Shun, et al.
Published: (2026)
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
by: Takahashi, Keiichiro, et al.
Published: (2024)
by: Takahashi, Keiichiro, et al.
Published: (2024)
When to Replan? An Adaptive Replanning Strategy for Autonomous Navigation using Deep Reinforcement Learning
by: Honda, Kohei, et al.
Published: (2023)
by: Honda, Kohei, et al.
Published: (2023)
Unsupervised Neural Motion Retargeting for Humanoid Teleoperation
by: Yagi, Satoshi, et al.
Published: (2024)
by: Yagi, Satoshi, et al.
Published: (2024)
Value Iteration for Learning Concurrently Executable Robotic Control Tasks
by: Tahmid, Sheikh A., et al.
Published: (2025)
by: Tahmid, Sheikh A., et al.
Published: (2025)
Robotic Strawberry Harvesting with Robust Vision and Deep Reinforcement Learning based Sim-to-Real Control
by: Bashir, Al, et al.
Published: (2026)
by: Bashir, Al, et al.
Published: (2026)
Efficient Reinforcement Learning of Task Planners for Robotic Palletization through Iterative Action Masking Learning
by: Wu, Zheng, et al.
Published: (2024)
by: Wu, Zheng, et al.
Published: (2024)
Semantically-driven Deep Reinforcement Learning for Inspection Path Planning
by: Malczyk, Grzegorz, et al.
Published: (2025)
by: Malczyk, Grzegorz, et al.
Published: (2025)
Deep Reinforcement Learning for Mobile Robot Path Planning
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
Deep Reinforcement Learning-based Large-scale Robot Exploration
by: Cao, Yuhong, et al.
Published: (2024)
by: Cao, Yuhong, et al.
Published: (2024)
SocialNav-MoE: A Mixture-of-Experts Vision Language Model for Socially Compliant Navigation with Reinforcement Fine-Tuning
by: Kawabata, Tomohito, et al.
Published: (2025)
by: Kawabata, Tomohito, et al.
Published: (2025)
Similar Items
-
Robust Sim-to-Real Cloth Untangling through Reduced-Resolution Observations via Adaptive Force-Difference Quantization
by: Tsurumine, Yoshihisa, et al.
Published: (2026) -
DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport
by: Shibata, Kazuki, et al.
Published: (2026) -
Progressive-Resolution Policy Distillation: Leveraging Coarse-Resolution Simulations for Time-Efficient Fine-Resolution Policy Learning
by: Kadokawa, Yuki, et al.
Published: (2024) -
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
by: Kadokawa, Yuki, et al.
Published: (2025) -
Cooperative Grasping and Transportation using Multi-agent Reinforcement Learning with Ternary Force Representation
by: Bernard-Tiong, Ing-Sheng, et al.
Published: (2024)