LDTrack: Dynamic People Tracking by Service Robots using Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fung, Angus, Benhabib, Beno, Nejat, Goldie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robots Autonomously Detecting People: A Multimodal Deep Contrastive Learning Method Robust to Intraclass Variations
von: Fung, Angus, et al.
Veröffentlicht: (2022)
von: Fung, Angus, et al.
Veröffentlicht: (2022)
Embodied AI with Foundation Models for Mobile Service Robots: A Systematic Review
von: Lisondra, Matthew, et al.
Veröffentlicht: (2025)
von: Lisondra, Matthew, et al.
Veröffentlicht: (2025)
Mobile Robot Navigation Using Hand-Drawn Maps: A Vision Language Model Approach
von: Tan, Aaron Hao, et al.
Veröffentlicht: (2025)
von: Tan, Aaron Hao, et al.
Veröffentlicht: (2025)
MLLM-Search: A Zero-Shot Approach to Finding People using Multimodal Large Language Models
von: Fung, Angus, et al.
Veröffentlicht: (2024)
von: Fung, Angus, et al.
Veröffentlicht: (2024)
PovNet+: A Deep Learning Architecture for Socially Assistive Robots to Learn and Assist with Multiple Activities of Daily Living
von: Robinson, Fraser, et al.
Veröffentlicht: (2026)
von: Robinson, Fraser, et al.
Veröffentlicht: (2026)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
von: Wen, Youpeng, et al.
Veröffentlicht: (2024)
von: Wen, Youpeng, et al.
Veröffentlicht: (2024)
PathFinder: Attention-Driven Dynamic Non-Line-of-Sight Tracking with a Mobile Robot
von: Kannapiran, Shenbagaraj, et al.
Veröffentlicht: (2024)
von: Kannapiran, Shenbagaraj, et al.
Veröffentlicht: (2024)
X-Nav: Learning End-to-End Cross-Embodiment Navigation for Mobile Robots
von: Wang, Haitong, et al.
Veröffentlicht: (2025)
von: Wang, Haitong, et al.
Veröffentlicht: (2025)
LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
Adaptive Anomaly Recovery for Telemanipulation: A Diffusion Model Approach to Vision-Based Tracking
von: Wang, Haoyang, et al.
Veröffentlicht: (2025)
von: Wang, Haoyang, et al.
Veröffentlicht: (2025)
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
6D Object Pose Tracking in Internet Videos for Robotic Manipulation
von: Ponimatkin, Georgy, et al.
Veröffentlicht: (2025)
von: Ponimatkin, Georgy, et al.
Veröffentlicht: (2025)
Dynamic-Dark SLAM: RGB-Thermal Cooperative Robot Vision Strategy for Multi-Person Tracking in Both Well-Lit and Low-Light Scenes
von: Sakai, Tatsuro, et al.
Veröffentlicht: (2025)
von: Sakai, Tatsuro, et al.
Veröffentlicht: (2025)
Exploring Conditions for Diffusion models in Robotic Control
von: Shin, Heeseong, et al.
Veröffentlicht: (2025)
von: Shin, Heeseong, et al.
Veröffentlicht: (2025)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
GS-LTS: 3D Gaussian Splatting-Based Adaptive Modeling for Long-Term Service Robots
von: Fu, Bin, et al.
Veröffentlicht: (2025)
von: Fu, Bin, et al.
Veröffentlicht: (2025)
Recognition of Dynamic Hand Gestures in Long Distance using a Web-Camera for Robot Guidance
von: Beeri, Eran Bamani, et al.
Veröffentlicht: (2024)
von: Beeri, Eran Bamani, et al.
Veröffentlicht: (2024)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
Fast LiDAR Upsampling using Conditional Diffusion Models
von: Helgesen, Sander Elias Magnussen, et al.
Veröffentlicht: (2024)
von: Helgesen, Sander Elias Magnussen, et al.
Veröffentlicht: (2024)
Pixel Motion Diffusion is What We Need for Robot Control
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2025)
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2025)
Online Calibration of a Single-Track Ground Vehicle Dynamics Model by Tight Fusion with Visual-Inertial Odometry
von: Li, Haolong, et al.
Veröffentlicht: (2023)
von: Li, Haolong, et al.
Veröffentlicht: (2023)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
Towards Generalizable Robotic Manipulation in Dynamic Environments
von: Fang, Heng, et al.
Veröffentlicht: (2026)
von: Fang, Heng, et al.
Veröffentlicht: (2026)
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
von: Zhang, Qiyao, et al.
Veröffentlicht: (2026)
von: Zhang, Qiyao, et al.
Veröffentlicht: (2026)
Context-Aware Risk Estimation in Home Environments: A Probabilistic Framework for Service Robots
von: Ishii, Sena, et al.
Veröffentlicht: (2025)
von: Ishii, Sena, et al.
Veröffentlicht: (2025)
Robotic VLA Benefits from Joint Learning with Motion Image Diffusion
von: Fang, Yu, et al.
Veröffentlicht: (2025)
von: Fang, Yu, et al.
Veröffentlicht: (2025)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
von: Chai, Ying, et al.
Veröffentlicht: (2025)
von: Chai, Ying, et al.
Veröffentlicht: (2025)
MonoSE(3)-Diffusion: A Monocular SE(3) Diffusion Framework for Robust Camera-to-Robot Pose Estimation
von: Zhu, Kangjian, et al.
Veröffentlicht: (2025)
von: Zhu, Kangjian, et al.
Veröffentlicht: (2025)
AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
TrackVLA: Embodied Visual Tracking in the Wild
von: Wang, Shaoan, et al.
Veröffentlicht: (2025)
von: Wang, Shaoan, et al.
Veröffentlicht: (2025)
Occupancy World Model for Robots
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
RobotSeg: A Model and Dataset for Segmenting Robots in Image and Video
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
TouchAnything: Diffusion-Guided 3D Reconstruction from Sparse Robot Touches
von: Gu, Langzhe, et al.
Veröffentlicht: (2026)
von: Gu, Langzhe, et al.
Veröffentlicht: (2026)
SmartFlow: Robotic Process Automation using LLMs
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
Lost & Found: Tracking Changes from Egocentric Observations in 3D Dynamic Scene Graphs
von: Behrens, Tjark, et al.
Veröffentlicht: (2024)
von: Behrens, Tjark, et al.
Veröffentlicht: (2024)
Keypoint-based Dynamic Object 6-DoF Pose Tracking via Event Camera
von: Wang, Zhe, et al.
Veröffentlicht: (2026)
von: Wang, Zhe, et al.
Veröffentlicht: (2026)
CloudTrack: Scalable UAV Tracking with Cloud Semantics
von: Blei, Yannik, et al.
Veröffentlicht: (2024)
von: Blei, Yannik, et al.
Veröffentlicht: (2024)
GarmentTracking: Category-Level Garment Pose Tracking
von: Xue, Han, et al.
Veröffentlicht: (2023)
von: Xue, Han, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Robots Autonomously Detecting People: A Multimodal Deep Contrastive Learning Method Robust to Intraclass Variations
von: Fung, Angus, et al.
Veröffentlicht: (2022) -
Embodied AI with Foundation Models for Mobile Service Robots: A Systematic Review
von: Lisondra, Matthew, et al.
Veröffentlicht: (2025) -
Mobile Robot Navigation Using Hand-Drawn Maps: A Vision Language Model Approach
von: Tan, Aaron Hao, et al.
Veröffentlicht: (2025) -
MLLM-Search: A Zero-Shot Approach to Finding People using Multimodal Large Language Models
von: Fung, Angus, et al.
Veröffentlicht: (2024) -
PovNet+: A Deep Learning Architecture for Socially Assistive Robots to Learn and Assist with Multiple Activities of Daily Living
von: Robinson, Fraser, et al.
Veröffentlicht: (2026)