IDD-X: A Multi-View Dataset for Ego-relative Important Object Localization and Explanation in Dense and Unstructured Traffic
Fuente:
arXiv
Saved in:
| Main Authors: | Parikh, Chirag, Saluja, Rohit, Jawahar, C. V., Sarvadevabhatla, Ravi Kiran |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transfer-LMR: Heavy-Tail Driving Behavior Recognition in Diverse Traffic Scenarios
by: Parikh, Chirag, et al.
Published: (2024)
by: Parikh, Chirag, et al.
Published: (2024)
RoadTones: Tone Controllable Text Generation from Road Event Videos
by: Parikh, Chirag, et al.
Published: (2026)
by: Parikh, Chirag, et al.
Published: (2026)
Pedestrian Intention and Trajectory Prediction in Unstructured Traffic Using IDD-PeD
by: Bokkasam, Ruthvik, et al.
Published: (2025)
by: Bokkasam, Ruthvik, et al.
Published: (2025)
RoadSocial: A Diverse VideoQA Dataset and Benchmark for Road Event Understanding from Social Video Narratives
by: Parikh, Chirag, et al.
Published: (2025)
by: Parikh, Chirag, et al.
Published: (2025)
IndicDLP: A Foundational Dataset for Multi-Lingual and Multi-Domain Document Layout Parsing
by: Nath, Oikantik, et al.
Published: (2025)
by: Nath, Oikantik, et al.
Published: (2025)
Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains
by: Deng, Zhiyun, et al.
Published: (2025)
by: Deng, Zhiyun, et al.
Published: (2025)
The S3LI Vulcano Dataset: A Dataset for Multi-Modal SLAM in Unstructured Planetary Environments
by: Giubilato, Riccardo, et al.
Published: (2026)
by: Giubilato, Riccardo, et al.
Published: (2026)
TexTAR : Textual Attribute Recognition in Multi-domain and Multi-lingual Document Images
by: Kumar, Rohan, et al.
Published: (2025)
by: Kumar, Rohan, et al.
Published: (2025)
Survey on Datasets for Perception in Unstructured Outdoor Environments
by: Mortimer, Peter, et al.
Published: (2024)
by: Mortimer, Peter, et al.
Published: (2024)
Keyframe-based Dense Mapping with the Graph of View-Dependent Local Maps
by: Zielinski, Krzysztof, et al.
Published: (2026)
by: Zielinski, Krzysztof, et al.
Published: (2026)
DashCop: Automated E-ticket Generation for Two-Wheeler Traffic Violations Using Dashcam Videos
by: Rawat, Deepti, et al.
Published: (2025)
by: Rawat, Deepti, et al.
Published: (2025)
DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects
by: Chen, Jiahong, et al.
Published: (2025)
by: Chen, Jiahong, et al.
Published: (2025)
MoRAG -- Multi-Fusion Retrieval Augmented Generation for Human Motion
by: Kalakonda, Sai Shashank, et al.
Published: (2024)
by: Kalakonda, Sai Shashank, et al.
Published: (2024)
The GOOSE Dataset for Perception in Unstructured Environments
by: Mortimer, Peter, et al.
Published: (2023)
by: Mortimer, Peter, et al.
Published: (2023)
Monocular Person Localization under Camera Ego-motion
by: Zhan, Yu, et al.
Published: (2025)
by: Zhan, Yu, et al.
Published: (2025)
V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views
by: You, Junwei, et al.
Published: (2026)
by: You, Junwei, et al.
Published: (2026)
Sparse3DTrack: Monocular 3D Object Tracking Using Sparse Supervision
by: Gosala, Nikhil, et al.
Published: (2026)
by: Gosala, Nikhil, et al.
Published: (2026)
Event-Free Moving Object Segmentation from Moving Ego Vehicle
by: Zhou, Zhuyun, et al.
Published: (2023)
by: Zhou, Zhuyun, et al.
Published: (2023)
UEVAVD: A Dataset for Developing UAV's Eye View Active Object Detection
by: Jiang, Xinhua, et al.
Published: (2024)
by: Jiang, Xinhua, et al.
Published: (2024)
BEVCar: Camera-Radar Fusion for BEV Map and Object Segmentation
by: Schramm, Jonas, et al.
Published: (2024)
by: Schramm, Jonas, et al.
Published: (2024)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
by: Wang, Jiasen, et al.
Published: (2024)
by: Wang, Jiasen, et al.
Published: (2024)
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)
by: Punamiya, Ryan, et al.
Published: (2026)
NuPlanQA: A Large-Scale Dataset and Benchmark for Multi-View Driving Scene Understanding in Multi-Modal Large Language Models
by: Park, Sung-Yeon, et al.
Published: (2025)
by: Park, Sung-Yeon, et al.
Published: (2025)
ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting
by: Papais, Sandro, et al.
Published: (2025)
by: Papais, Sandro, et al.
Published: (2025)
Benchmarking Multi-View BEV Object Detection with Mixed Pinhole and Fisheye Cameras
by: Liu, Xiangzhong, et al.
Published: (2026)
by: Liu, Xiangzhong, et al.
Published: (2026)
SOS-Match: Segmentation for Open-Set Robust Correspondence Search and Robot Localization in Unstructured Environments
by: Thomas, Annika, et al.
Published: (2024)
by: Thomas, Annika, et al.
Published: (2024)
OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation
by: Jawaid, Ahad, et al.
Published: (2025)
by: Jawaid, Ahad, et al.
Published: (2025)
Active 6D Pose Estimation for Textureless Objects using Multi-View RGB Frames
by: Yang, Jun, et al.
Published: (2025)
by: Yang, Jun, et al.
Published: (2025)
StreamMOS: Streaming Moving Object Segmentation with Multi-View Perception and Dual-Span Memory
by: Li, Zhiheng, et al.
Published: (2024)
by: Li, Zhiheng, et al.
Published: (2024)
UncertaintyTrack: Exploiting Detection and Localization Uncertainty in Multi-Object Tracking
by: Lee, Chang Won, et al.
Published: (2024)
by: Lee, Chang Won, et al.
Published: (2024)
Design and Identification of Keypoint Patches in Unstructured Environments
by: Park, Taewook, et al.
Published: (2024)
by: Park, Taewook, et al.
Published: (2024)
Online Adaptive Traversability Estimation through Interaction for Unstructured, Densely Vegetated Environments
by: Ruetz, Fabio A., et al.
Published: (2025)
by: Ruetz, Fabio A., et al.
Published: (2025)
A Multi-Level Similarity Approach for Single-View Object Grasping: Matching, Planning, and Fine-Tuning
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
LiDAR-BEVMTN: Real-Time LiDAR Bird's-Eye View Multi-Task Perception Network for Autonomous Driving
by: Mohapatra, Sambit, et al.
Published: (2023)
by: Mohapatra, Sambit, et al.
Published: (2023)
EgoFSD: Ego-Centric Fully Sparse Paradigm with Uncertainty Denoising and Iterative Refinement for Efficient End-to-End Self-Driving
by: Su, Haisheng, et al.
Published: (2024)
by: Su, Haisheng, et al.
Published: (2024)
SeePerSea: Multi-modal Perception Dataset of In-water Objects for Autonomous Surface Vehicles
by: Jeong, Mingi, et al.
Published: (2024)
by: Jeong, Mingi, et al.
Published: (2024)
Synset Signset Germany: a Synthetic Dataset for German Traffic Sign Recognition
by: Sielemann, Anne, et al.
Published: (2025)
by: Sielemann, Anne, et al.
Published: (2025)
CLIP-Loc: Multi-modal Landmark Association for Global Localization in Object-based Maps
by: Matsuzaki, Shigemichi, et al.
Published: (2024)
by: Matsuzaki, Shigemichi, et al.
Published: (2024)
PAct: Part-Decomposed Single-View Articulated Object Generation
by: Liu, Qingming, et al.
Published: (2026)
by: Liu, Qingming, et al.
Published: (2026)
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025)
by: Kim, Young Hun, et al.
Published: (2025)
Similar Items
-
Transfer-LMR: Heavy-Tail Driving Behavior Recognition in Diverse Traffic Scenarios
by: Parikh, Chirag, et al.
Published: (2024) -
RoadTones: Tone Controllable Text Generation from Road Event Videos
by: Parikh, Chirag, et al.
Published: (2026) -
Pedestrian Intention and Trajectory Prediction in Unstructured Traffic Using IDD-PeD
by: Bokkasam, Ruthvik, et al.
Published: (2025) -
RoadSocial: A Diverse VideoQA Dataset and Benchmark for Road Event Understanding from Social Video Narratives
by: Parikh, Chirag, et al.
Published: (2025) -
IndicDLP: A Foundational Dataset for Multi-Lingual and Multi-Domain Document Layout Parsing
by: Nath, Oikantik, et al.
Published: (2025)