JRDB-PanoTrack: An Open-world Panoptic Segmentation and Tracking Robotic Dataset in Crowded Human Environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Le, Duy-Tho, Gou, Chenhui, Datta, Stavya, Shi, Hengcan, Reid, Ian, Cai, Jianfei, Rezatofighi, Hamid |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DifFUSER: Diffusion Model for Robust Multi-Sensor Fusion in 3D Object Detection and BEV Segmentation
di: Le, Duy-Tho, et al.
Pubblicazione: (2024)
di: Le, Duy-Tho, et al.
Pubblicazione: (2024)
ARIS: Agentic and Relationship Intelligence System for Social Robots
di: Datta, Stavya, et al.
Pubblicazione: (2026)
di: Datta, Stavya, et al.
Pubblicazione: (2026)
Marginalized Generalized IoU (MGIoU): A Unified Objective Function for Optimizing Any Convex Parametric Shapes
di: Le, Duy-Tho, et al.
Pubblicazione: (2025)
di: Le, Duy-Tho, et al.
Pubblicazione: (2025)
DrVideo: Document Retrieval Based Long Video Understanding
di: Ma, Ziyu, et al.
Pubblicazione: (2024)
di: Ma, Ziyu, et al.
Pubblicazione: (2024)
JRDB-Pose3D: A Multi-person 3D Human Pose and Shape Estimation Dataset for Robotics
di: Biswas, Sandika, et al.
Pubblicazione: (2026)
di: Biswas, Sandika, et al.
Pubblicazione: (2026)
JRDB-Social: A Multifaceted Robotic Dataset for Understanding of Context and Dynamics of Human Interactions Within Social Groups
di: Jahangard, Simindokht, et al.
Pubblicazione: (2024)
di: Jahangard, Simindokht, et al.
Pubblicazione: (2024)
Improving Visual Perception of a Social Robot for Controlled and In-the-wild Human-robot Interaction
di: Zhong, Wangjie, et al.
Pubblicazione: (2024)
di: Zhong, Wangjie, et al.
Pubblicazione: (2024)
JRDB-Reasoning: A Difficulty-Graded Benchmark for Visual Reasoning in Robotics
di: Jahangard, Simindokht, et al.
Pubblicazione: (2025)
di: Jahangard, Simindokht, et al.
Pubblicazione: (2025)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
di: Ehsanpour, Mahsa, et al.
Pubblicazione: (2024)
di: Ehsanpour, Mahsa, et al.
Pubblicazione: (2024)
How Well Can Vision Language Models See Image Details?
di: Gou, Chenhui, et al.
Pubblicazione: (2024)
di: Gou, Chenhui, et al.
Pubblicazione: (2024)
PanoGS: Gaussian-based Panoptic Segmentation for 3D Open Vocabulary Scene Understanding
di: Zhai, Hongjia, et al.
Pubblicazione: (2025)
di: Zhai, Hongjia, et al.
Pubblicazione: (2025)
dinov3.seg: Open-Vocabulary Semantic Segmentation with DINOv3
di: Dutta, Saikat, et al.
Pubblicazione: (2026)
di: Dutta, Saikat, et al.
Pubblicazione: (2026)
GyroCopter: Differential Bearing Measuring Trajectory Planner for Tracking and Localizing Radio Frequency Sources
di: Chen, Fei, et al.
Pubblicazione: (2024)
di: Chen, Fei, et al.
Pubblicazione: (2024)
ETHcavation: A Dataset and Pipeline for Panoptic Scene Understanding and Object Tracking in Dynamic Construction Environments
di: Terenzi, Lorenzo, et al.
Pubblicazione: (2024)
di: Terenzi, Lorenzo, et al.
Pubblicazione: (2024)
An Empirical Study on How Video-LLMs Answer Video Questions
di: Gou, Chenhui, et al.
Pubblicazione: (2025)
di: Gou, Chenhui, et al.
Pubblicazione: (2025)
Mobile-VideoGPT: Fast and Accurate Model for Mobile Video Understanding
di: Shaker, Abdelrahman, et al.
Pubblicazione: (2025)
di: Shaker, Abdelrahman, et al.
Pubblicazione: (2025)
Hier-SLAM: Scaling-up Semantics in SLAM with a Hierarchically Categorical Gaussian Splatting
di: Li, Boying, et al.
Pubblicazione: (2024)
di: Li, Boying, et al.
Pubblicazione: (2024)
AerOSeg: Harnessing SAM for Open-Vocabulary Segmentation in Remote Sensing Images
di: Dutta, Saikat, et al.
Pubblicazione: (2025)
di: Dutta, Saikat, et al.
Pubblicazione: (2025)
Open-World Panoptic Segmentation
di: Sodano, Matteo, et al.
Pubblicazione: (2024)
di: Sodano, Matteo, et al.
Pubblicazione: (2024)
TrackOcc: Camera-based 4D Panoptic Occupancy Tracking
di: Chen, Zhuoguang, et al.
Pubblicazione: (2025)
di: Chen, Zhuoguang, et al.
Pubblicazione: (2025)
SPORTS: Simultaneous Panoptic Odometry, Rendering, Tracking and Segmentation for Urban Scenes Understanding
di: Yang, Zhiliu, et al.
Pubblicazione: (2025)
di: Yang, Zhiliu, et al.
Pubblicazione: (2025)
Learning Appearance and Motion Cues for Panoptic Tracking
di: Hurtado, Juana Valeria, et al.
Pubblicazione: (2025)
di: Hurtado, Juana Valeria, et al.
Pubblicazione: (2025)
Panoptic-SLAM: Visual SLAM in Dynamic Environments using Panoptic Segmentation
di: Abati, Gabriel Fischer, et al.
Pubblicazione: (2024)
di: Abati, Gabriel Fischer, et al.
Pubblicazione: (2024)
IndoorCrowd: A Multi-Scene Dataset for Human Detection, Segmentation, and Tracking with an Automated Annotation Pipeline
di: Nae, Sebastian-Ion, et al.
Pubblicazione: (2026)
di: Nae, Sebastian-Ion, et al.
Pubblicazione: (2026)
CalibAnyView: Beyond Single-View Camera Calibration in the Wild
di: Li, Boying, et al.
Pubblicazione: (2026)
di: Li, Boying, et al.
Pubblicazione: (2026)
Hier-SLAM++: Neuro-Symbolic Semantic SLAM with a Hierarchically Categorical Gaussian Splatting
di: Li, Boying, et al.
Pubblicazione: (2025)
di: Li, Boying, et al.
Pubblicazione: (2025)
DynamicTrack: Advancing Gigapixel Tracking in Crowded Scenes
di: Zhao, Yunqi, et al.
Pubblicazione: (2024)
di: Zhao, Yunqi, et al.
Pubblicazione: (2024)
Lidar Panoptic Segmentation in an Open World
di: Chakravarthy, Anirudh S, et al.
Pubblicazione: (2024)
di: Chakravarthy, Anirudh S, et al.
Pubblicazione: (2024)
ASAP-Textured Gaussians: Enhancing Textured Gaussians with Adaptive Sampling and Anisotropic Parameterization
di: Wei, Meng, et al.
Pubblicazione: (2025)
di: Wei, Meng, et al.
Pubblicazione: (2025)
Normal-GS: 3D Gaussian Splatting with Normal-Involved Rendering
di: Wei, Meng, et al.
Pubblicazione: (2024)
di: Wei, Meng, et al.
Pubblicazione: (2024)
Multi-Objective Multi-Agent Planning for Discovering and Tracking Multiple Mobile Objects
di: Van Nguyen, Hoa, et al.
Pubblicazione: (2022)
di: Van Nguyen, Hoa, et al.
Pubblicazione: (2022)
PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM
di: Chen, Runnan, et al.
Pubblicazione: (2024)
di: Chen, Runnan, et al.
Pubblicazione: (2024)
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
OpenAnimalTracks: A Dataset for Animal Track Recognition
di: Shinoda, Risa, et al.
Pubblicazione: (2024)
di: Shinoda, Risa, et al.
Pubblicazione: (2024)
Open Vocabulary Panoptic Segmentation With Retrieval Augmentation
di: Sadeq, Nafis, et al.
Pubblicazione: (2026)
di: Sadeq, Nafis, et al.
Pubblicazione: (2026)
PanoSSC: Exploring Monocular Panoptic 3D Scene Reconstruction for Autonomous Driving
di: Shi, Yining, et al.
Pubblicazione: (2024)
di: Shi, Yining, et al.
Pubblicazione: (2024)
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
di: Yu, Xuan, et al.
Pubblicazione: (2024)
di: Yu, Xuan, et al.
Pubblicazione: (2024)
Latent Gaussian Splatting for 4D Panoptic Occupancy Tracking
di: Luz, Maximilian, et al.
Pubblicazione: (2026)
di: Luz, Maximilian, et al.
Pubblicazione: (2026)
UAV (Unmanned Aerial Vehicles): Diverse Applications of UAV Datasets in Segmentation, Classification, Detection, and Tracking
di: Rahman, Md. Mahfuzur, et al.
Pubblicazione: (2024)
di: Rahman, Md. Mahfuzur, et al.
Pubblicazione: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
di: Jahangard, Simindokht, et al.
Pubblicazione: (2025)
di: Jahangard, Simindokht, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DifFUSER: Diffusion Model for Robust Multi-Sensor Fusion in 3D Object Detection and BEV Segmentation
di: Le, Duy-Tho, et al.
Pubblicazione: (2024) -
ARIS: Agentic and Relationship Intelligence System for Social Robots
di: Datta, Stavya, et al.
Pubblicazione: (2026) -
Marginalized Generalized IoU (MGIoU): A Unified Objective Function for Optimizing Any Convex Parametric Shapes
di: Le, Duy-Tho, et al.
Pubblicazione: (2025) -
DrVideo: Document Retrieval Based Long Video Understanding
di: Ma, Ziyu, et al.
Pubblicazione: (2024) -
JRDB-Pose3D: A Multi-person 3D Human Pose and Shape Estimation Dataset for Robotics
di: Biswas, Sandika, et al.
Pubblicazione: (2026)