Few-shot Vision-based Human Activity Recognition with MLLM-based Visual Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Wenqi, Arakawa, Yutaka |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-shot Sim-to-Real Transfer for Reinforcement Learning-based Visual Servoing of Soft Continuum Arms
von: Yang, Hsin-Jung, et al.
Veröffentlicht: (2025)
von: Yang, Hsin-Jung, et al.
Veröffentlicht: (2025)
Lifelong Ensemble Learning based on Multiple Representations for Few-Shot Object Recognition
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2022)
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2022)
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
von: Li, Puhao, et al.
Veröffentlicht: (2025)
von: Li, Puhao, et al.
Veröffentlicht: (2025)
Millimeter Wave Radar-based Human Activity Recognition for Healthcare Monitoring Robot
von: Gu, Zhanzhong, et al.
Veröffentlicht: (2024)
von: Gu, Zhanzhong, et al.
Veröffentlicht: (2024)
Learning Vision-based Robotic Manipulation Tasks Sequentially in Offline Reinforcement Learning Settings
von: Yadav, Sudhir Pratap, et al.
Veröffentlicht: (2023)
von: Yadav, Sudhir Pratap, et al.
Veröffentlicht: (2023)
Composing Diffusion Policies for Few-shot Learning of Movement Trajectories
von: Patil, Omkar, et al.
Veröffentlicht: (2024)
von: Patil, Omkar, et al.
Veröffentlicht: (2024)
ContactRL: Safe Reinforcement Learning based Motion Planning for Contact based Human Robot Collaboration
von: Mulkana, Sundas Rafat, et al.
Veröffentlicht: (2025)
von: Mulkana, Sundas Rafat, et al.
Veröffentlicht: (2025)
Few Shot System Identification for Reinforcement Learning
von: Farid, Karim, et al.
Veröffentlicht: (2021)
von: Farid, Karim, et al.
Veröffentlicht: (2021)
Safe Navigation for Robotic Digestive Endoscopy via Human Intervention-based Reinforcement Learning
von: Tan, Min, et al.
Veröffentlicht: (2024)
von: Tan, Min, et al.
Veröffentlicht: (2024)
MOTIF: Learning Action Motifs for Few-shot Cross-Embodiment Transfer
von: Zhi, Heng, et al.
Veröffentlicht: (2026)
von: Zhi, Heng, et al.
Veröffentlicht: (2026)
AED: Adaptable Error Detection for Few-shot Imitation Policy
von: Yeh, Jia-Fong, et al.
Veröffentlicht: (2024)
von: Yeh, Jia-Fong, et al.
Veröffentlicht: (2024)
PrefMMT: Modeling Human Preferences in Preference-based Reinforcement Learning with Multimodal Transformers
von: Zhao, Dezhong, et al.
Veröffentlicht: (2024)
von: Zhao, Dezhong, et al.
Veröffentlicht: (2024)
Robotic Strawberry Harvesting with Robust Vision and Deep Reinforcement Learning based Sim-to-Real Control
von: Bashir, Al, et al.
Veröffentlicht: (2026)
von: Bashir, Al, et al.
Veröffentlicht: (2026)
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models
von: Chen, Yuxuan, et al.
Veröffentlicht: (2025)
von: Chen, Yuxuan, et al.
Veröffentlicht: (2025)
FlowVLA: Visual Chain of Thought-based Motion Reasoning for Vision-Language-Action Models
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
Commonsense Scene Graph-based Target Localization for Object Search
von: Ge, Wenqi, et al.
Veröffentlicht: (2024)
von: Ge, Wenqi, et al.
Veröffentlicht: (2024)
SP-VINS: A Hybrid Stereo Visual Inertial Navigation System based on Implicit Environmental Map
von: Du, Xueyu, et al.
Veröffentlicht: (2025)
von: Du, Xueyu, et al.
Veröffentlicht: (2025)
A Novel MLLM-based Approach for Autonomous Driving in Different Weather Conditions
von: Fourati, Sonda, et al.
Veröffentlicht: (2024)
von: Fourati, Sonda, et al.
Veröffentlicht: (2024)
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
Show and Grasp: Few-shot Semantic Segmentation for Robot Grasping through Zero-shot Foundation Models
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
TraKDis: A Transformer-based Knowledge Distillation Approach for Visual Reinforcement Learning with Application to Cloth Manipulation
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
ImLPR: Image-based LiDAR Place Recognition using Vision Foundation Models
von: Jung, Minwoo, et al.
Veröffentlicht: (2025)
von: Jung, Minwoo, et al.
Veröffentlicht: (2025)
Leveraging GCN-based Action Recognition for Teleoperation in Daily Activity Assistance
von: Kwok, Thomas M., et al.
Veröffentlicht: (2025)
von: Kwok, Thomas M., et al.
Veröffentlicht: (2025)
Dual-Agent Reinforcement Learning for Adaptive and Cost-Aware Visual-Inertial Odometry
von: Pan, Feiyang, et al.
Veröffentlicht: (2025)
von: Pan, Feiyang, et al.
Veröffentlicht: (2025)
Enhancing Robotic Arm Activity Recognition with Vision Transformers and Wavelet-Transformed Channel State Information
von: Zandi, Rojin, et al.
Veröffentlicht: (2024)
von: Zandi, Rojin, et al.
Veröffentlicht: (2024)
Learning-based Dynamic Robot-to-Human Handover
von: Kim, Hyeonseong, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonseong, et al.
Veröffentlicht: (2025)
UAV Control with Vision-based Hand Gesture Recognition over Edge-Computing
von: Abdalla, Sousannah, et al.
Veröffentlicht: (2025)
von: Abdalla, Sousannah, et al.
Veröffentlicht: (2025)
H-Zero: Cross-Humanoid Locomotion Pretraining Enables Few-shot Novel Embodiment Transfer
von: Lin, Yunfeng, et al.
Veröffentlicht: (2025)
von: Lin, Yunfeng, et al.
Veröffentlicht: (2025)
Few-shot Sim2Real Based on High Fidelity Rendering with Force Feedback Teleoperation
von: Zou, Yanwen, et al.
Veröffentlicht: (2025)
von: Zou, Yanwen, et al.
Veröffentlicht: (2025)
Few-shot transfer of tool-use skills using human demonstrations with proximity and tactile sensing
von: Aoyama, Marina Y., et al.
Veröffentlicht: (2025)
von: Aoyama, Marina Y., et al.
Veröffentlicht: (2025)
Few-Shot Vision-Language Action-Incremental Policy Learning
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
Knowledge-Decoupled Synergetic Learning: An MLLM based Collaborative Approach to Few-shot Multimodal Dialogue Intention Recognition
von: Chen, Bin, et al.
Veröffentlicht: (2025)
von: Chen, Bin, et al.
Veröffentlicht: (2025)
Visual-Geometry GP-based Navigable Space for Autonomous Navigation
von: Ali, Mahmoud, et al.
Veröffentlicht: (2024)
von: Ali, Mahmoud, et al.
Veröffentlicht: (2024)
DriveMind: A Dual Visual Language Model-based Reinforcement Learning Framework for Autonomous Driving
von: Wasif, Dawood, et al.
Veröffentlicht: (2025)
von: Wasif, Dawood, et al.
Veröffentlicht: (2025)
VMTS: Vision-Assisted Teacher-Student Reinforcement Learning for Multi-Terrain Locomotion in Bipedal Robots
von: Chen, Fu, et al.
Veröffentlicht: (2025)
von: Chen, Fu, et al.
Veröffentlicht: (2025)
HierKick: Hierarchical Reinforcement Learning for Vision-Guided Soccer Robot Control
von: Chen, Yizhi, et al.
Veröffentlicht: (2026)
von: Chen, Yizhi, et al.
Veröffentlicht: (2026)
VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2024)
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning-based Large-scale Robot Exploration
von: Cao, Yuhong, et al.
Veröffentlicht: (2024)
von: Cao, Yuhong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Zero-shot Sim-to-Real Transfer for Reinforcement Learning-based Visual Servoing of Soft Continuum Arms
von: Yang, Hsin-Jung, et al.
Veröffentlicht: (2025) -
Lifelong Ensemble Learning based on Multiple Representations for Few-Shot Object Recognition
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2022) -
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
von: Li, Puhao, et al.
Veröffentlicht: (2025) -
Millimeter Wave Radar-based Human Activity Recognition for Healthcare Monitoring Robot
von: Gu, Zhanzhong, et al.
Veröffentlicht: (2024) -
Learning Vision-based Robotic Manipulation Tasks Sequentially in Offline Reinforcement Learning Settings
von: Yadav, Sudhir Pratap, et al.
Veröffentlicht: (2023)