Domains as Objectives: Domain-Uncertainty-Aware Policy Optimization through Explicit Multi-Domain Convex Coverage Set Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ilboudo, Wendyam Eric Lionel, Kobayashi, Taisuke, Matsubara, Takamitsu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
An Inference-Based Architecture for Intent and Affordance Saturation in Decision-Making
von: Ilboudo, Wendyam Eric Lionel, et al.
Veröffentlicht: (2025)
von: Ilboudo, Wendyam Eric Lionel, et al.
Veröffentlicht: (2025)
CoLF: Learning Consistent Leader-Follower Policies for Vision-Language-Guided Multi-Robot Cooperative Transport
von: Despature, Joachim Yann, et al.
Veröffentlicht: (2026)
von: Despature, Joachim Yann, et al.
Veröffentlicht: (2026)
Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications
von: Hatanaka, Wataru, et al.
Veröffentlicht: (2025)
von: Hatanaka, Wataru, et al.
Veröffentlicht: (2025)
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2025)
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2025)
Task-priority Intermediated Hierarchical Distributed Policies: Reinforcement Learning of Adaptive Multi-robot Cooperative Transport
von: Naito, Yusei, et al.
Veröffentlicht: (2024)
von: Naito, Yusei, et al.
Veröffentlicht: (2024)
Progressive-Resolution Policy Distillation: Leveraging Coarse-Resolution Simulations for Time-Efficient Fine-Resolution Policy Learning
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2024)
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2024)
Leveraging Demonstrator-perceived Precision for Safe Interactive Imitation Learning of Clearance-limited Tasks
von: Oh, Hanbit, et al.
Veröffentlicht: (2024)
von: Oh, Hanbit, et al.
Veröffentlicht: (2024)
Composite Gaussian Processes Flows for Learning Discontinuous Multimodal Policies
von: Wang, Shu-yuan, et al.
Veröffentlicht: (2025)
von: Wang, Shu-yuan, et al.
Veröffentlicht: (2025)
Feasibility-aware Imitation Learning from Observations through a Hand-mounted Demonstration Interface
von: Takahashi, Kei, et al.
Veröffentlicht: (2025)
von: Takahashi, Kei, et al.
Veröffentlicht: (2025)
Search at Scale: Improving Numerical Conditioning of Ergodic Coverage Optimization for Multi-Scale Domains
von: Lahrach, Yanis, et al.
Veröffentlicht: (2025)
von: Lahrach, Yanis, et al.
Veröffentlicht: (2025)
Reinforcement Learning of Multi-robot Task Allocation for Multi-object Transportation with Infeasible Tasks
von: Shida, Yuma, et al.
Veröffentlicht: (2024)
von: Shida, Yuma, et al.
Veröffentlicht: (2024)
Incipient Slip Detection by Vibration Injection into Soft Sensor
von: Komeno, Naoto, et al.
Veröffentlicht: (2024)
von: Komeno, Naoto, et al.
Veröffentlicht: (2024)
LiRA: Light-Robust Adversary for Model-based Reinforcement Learning in Real World
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
Feasibility-aware Imitation Learning from Observation with Multimodal Feedback
von: Takahashi, Kei, et al.
Veröffentlicht: (2026)
von: Takahashi, Kei, et al.
Veröffentlicht: (2026)
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
ASBI: Leveraging Informative Real-World Data for Active Black-Box Simulator Tuning
von: Kim, Gahee, et al.
Veröffentlicht: (2025)
von: Kim, Gahee, et al.
Veröffentlicht: (2025)
BEAC: Imitating Complex Exploration and Task-oriented Behaviors for Invisible Object Nonprehensile Manipulation
von: Tahara, Hirotaka, et al.
Veröffentlicht: (2025)
von: Tahara, Hirotaka, et al.
Veröffentlicht: (2025)
Intentionally-underestimated Value Function at Terminal State for Temporal-difference Learning with Mis-designed Reward
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
Task-Relevant and Irrelevant Region-Aware Augmentation for Generalizable Vision-Based Imitation Learning in Agricultural Manipulation
von: Hattori, Shun, et al.
Veröffentlicht: (2026)
von: Hattori, Shun, et al.
Veröffentlicht: (2026)
ICCO: Learning an Instruction-conditioned Coordinator for Language-guided Task-aligned Multi-robot Control
von: Yano, Yoshiki, et al.
Veröffentlicht: (2025)
von: Yano, Yoshiki, et al.
Veröffentlicht: (2025)
Design of Restricted Normalizing Flow towards Arbitrary Stochastic Policy with Computational Efficiency
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
Tracing Energy Flow: Learning Tactile-based Grasping Force Control to Prevent Slippage in Dynamic Object Interaction
von: Kuo, Cheng-Yu, et al.
Veröffentlicht: (2025)
von: Kuo, Cheng-Yu, et al.
Veröffentlicht: (2025)
Cutting Sequence Diffuser: Sim-to-Real Transferable Planning for Object Shaping by Grinding
von: Hachimine, Takumi, et al.
Veröffentlicht: (2024)
von: Hachimine, Takumi, et al.
Veröffentlicht: (2024)
Prolonging Tool Life: Learning Skillful Use of General-purpose Tools through Lifespan-guided Reinforcement Learning
von: Wu, Po-Yen, et al.
Veröffentlicht: (2025)
von: Wu, Po-Yen, et al.
Veröffentlicht: (2025)
OIPP: Object-Adaptive Impact Point Predictor for Catching Diverse In-Flight Objects
von: Nguyen, Ngoc Huy, et al.
Veröffentlicht: (2025)
von: Nguyen, Ngoc Huy, et al.
Veröffentlicht: (2025)
Cooperative Grasping and Transportation using Multi-agent Reinforcement Learning with Ternary Force Representation
von: Bernard-Tiong, Ing-Sheng, et al.
Veröffentlicht: (2024)
von: Bernard-Tiong, Ing-Sheng, et al.
Veröffentlicht: (2024)
Robust Sim-to-Real Cloth Untangling through Reduced-Resolution Observations via Adaptive Force-Difference Quantization
von: Tsurumine, Yoshihisa, et al.
Veröffentlicht: (2026)
von: Tsurumine, Yoshihisa, et al.
Veröffentlicht: (2026)
Spiral Complete Coverage Path Planning Based on Conformal Slit Mapping in Multi-connected Domains
von: Shen, Changqing, et al.
Veröffentlicht: (2023)
von: Shen, Changqing, et al.
Veröffentlicht: (2023)
Point Bridge: 3D Representations for Cross Domain Policy Learning
von: Haldar, Siddhant, et al.
Veröffentlicht: (2026)
von: Haldar, Siddhant, et al.
Veröffentlicht: (2026)
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory
von: Yang, Changchuan, et al.
Veröffentlicht: (2025)
von: Yang, Changchuan, et al.
Veröffentlicht: (2025)
DADP: Domain Adaptive Diffusion Policy
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
Domain Adaptation of Visual Policies with a Single Demonstration
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
Multi-Agent Off-World Exploration for Sparse Evidence Discovery via Gaussian Belief Mapping and Dual-Domain Coverage
von: Qiao, Zhuoran, et al.
Veröffentlicht: (2026)
von: Qiao, Zhuoran, et al.
Veröffentlicht: (2026)
Robust Iterative Value Conversion: Deep Reinforcement Learning for Neurochip-driven Edge Robots
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2024)
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2024)
Real-time Sampling-based Model Predictive Control based on Reverse Kullback-Leibler Divergence and Its Adaptive Acceleration
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2022)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2022)
DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport
von: Shibata, Kazuki, et al.
Veröffentlicht: (2026)
von: Shibata, Kazuki, et al.
Veröffentlicht: (2026)
Where Do We Look When We Teach? Analyzing Human Gaze Behavior Across Demonstration Devices in Robot Imitation Learning
von: Ishida, Yutaro, et al.
Veröffentlicht: (2025)
von: Ishida, Yutaro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024) -
An Inference-Based Architecture for Intent and Affordance Saturation in Decision-Making
von: Ilboudo, Wendyam Eric Lionel, et al.
Veröffentlicht: (2025) -
CoLF: Learning Consistent Leader-Follower Policies for Vision-Language-Guided Multi-Robot Cooperative Transport
von: Despature, Joachim Yann, et al.
Veröffentlicht: (2026) -
Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications
von: Hatanaka, Wataru, et al.
Veröffentlicht: (2025) -
DAPPER: Discriminability-Aware Policy-to-Policy Preference-Based Reinforcement Learning for Query-Efficient Robot Skill Acquisition
von: Kadokawa, Yuki, et al.
Veröffentlicht: (2025)