On-Road Object Importance Estimation: A New Dataset and A Model with Multi-Fold Top-Down Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nan, Zhixiong, Chen, Yilong, Zhou, Tianfei, Xiang, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FBLNet: FeedBack Loop Network for Driver Attention Prediction
von: Chen, Yilong, et al.
Veröffentlicht: (2022)
von: Chen, Yilong, et al.
Veröffentlicht: (2022)
BiFold: Bimanual Cloth Folding with Language Guidance
von: Barbany, Oriol, et al.
Veröffentlicht: (2025)
von: Barbany, Oriol, et al.
Veröffentlicht: (2025)
MI-DETR: An Object Detection Model with Multi-time Inquiries Mechanism
von: Nan, Zhixiong, et al.
Veröffentlicht: (2025)
von: Nan, Zhixiong, et al.
Veröffentlicht: (2025)
Top-Down Guidance for Learning Object-Centric Representations
von: Zou, Junhong, et al.
Veröffentlicht: (2024)
von: Zou, Junhong, et al.
Veröffentlicht: (2024)
DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model
von: Nan, Zhixiong, et al.
Veröffentlicht: (2024)
von: Nan, Zhixiong, et al.
Veröffentlicht: (2024)
MineInsight: A Multi-sensor Dataset for Humanitarian Demining Robotics in Off-Road Environments
von: Malizia, Mario, et al.
Veröffentlicht: (2025)
von: Malizia, Mario, et al.
Veröffentlicht: (2025)
Flat'n'Fold: A Diverse Multi-Modal Dataset for Garment Perception and Manipulation
von: Zhuang, Lipeng, et al.
Veröffentlicht: (2024)
von: Zhuang, Lipeng, et al.
Veröffentlicht: (2024)
TAG: Target-Agnostic Guidance for Stable Object-Centric Inference in Vision-Language-Action Models
von: Zhou, Jiaying, et al.
Veröffentlicht: (2026)
von: Zhou, Jiaying, et al.
Veröffentlicht: (2026)
Target-Oriented Object Grasping via Multimodal Human Guidance
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
GenTrack: A New Generation of Multi-Object Tracking
von: Van Nguyen, Toan, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Toan, et al.
Veröffentlicht: (2025)
Road Obstacle Detection based on Unknown Objectness Scores
von: Noguchi, Chihiro, et al.
Veröffentlicht: (2024)
von: Noguchi, Chihiro, et al.
Veröffentlicht: (2024)
FRED: A Multi-Modal Autonomous Driving Dataset for Flooded Road Environments
von: Malone, Connor, et al.
Veröffentlicht: (2026)
von: Malone, Connor, et al.
Veröffentlicht: (2026)
FoldNet: Learning Generalizable Closed-Loop Policy for Garment Folding via Keypoint-Driven Asset and Demonstration Synthesis
von: Chen, Yuxing, et al.
Veröffentlicht: (2025)
von: Chen, Yuxing, et al.
Veröffentlicht: (2025)
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
Label-Efficient 3D Object Detection For Road-Side Units
von: Dao, Minh-Quan, et al.
Veröffentlicht: (2024)
von: Dao, Minh-Quan, et al.
Veröffentlicht: (2024)
Online Estimation of Table-Top Grown Strawberry Mass in Field Conditions with Occlusions
von: Zhen, Jinshan, et al.
Veröffentlicht: (2025)
von: Zhen, Jinshan, et al.
Veröffentlicht: (2025)
OnSiteVRU: A High-Resolution Trajectory Dataset for High-Density Vulnerable Road Users
von: Yan, Zhangcun, et al.
Veröffentlicht: (2025)
von: Yan, Zhangcun, et al.
Veröffentlicht: (2025)
S3E: A Multi-Robot Multimodal Dataset for Collaborative SLAM
von: Feng, Dapeng, et al.
Veröffentlicht: (2022)
von: Feng, Dapeng, et al.
Veröffentlicht: (2022)
Geo-locating Road Objects using Inverse Haversine Formula with NVIDIA Driveworks
von: Shami, Mamoona Birkhez, et al.
Veröffentlicht: (2024)
von: Shami, Mamoona Birkhez, et al.
Veröffentlicht: (2024)
MASSTAR: A Multi-Modal and Large-Scale Scene Dataset with a Versatile Toolchain for Surface Prediction and Completion
von: Zheng, Guiyong, et al.
Veröffentlicht: (2024)
von: Zheng, Guiyong, et al.
Veröffentlicht: (2024)
Learning a General Model: Folding Clothing with Topological Dynamics
von: Liu, Yiming, et al.
Veröffentlicht: (2025)
von: Liu, Yiming, et al.
Veröffentlicht: (2025)
SeePerSea: Multi-modal Perception Dataset of In-water Objects for Autonomous Surface Vehicles
von: Jeong, Mingi, et al.
Veröffentlicht: (2024)
von: Jeong, Mingi, et al.
Veröffentlicht: (2024)
MARVL: Multi-Stage Guidance for Robotic Manipulation via Vision-Language Models
von: Zhou, Xunlan, et al.
Veröffentlicht: (2026)
von: Zhou, Xunlan, et al.
Veröffentlicht: (2026)
Object Pose Estimation through Dexterous Touch
von: Shahidzadeh, Amir-Hossein, et al.
Veröffentlicht: (2025)
von: Shahidzadeh, Amir-Hossein, et al.
Veröffentlicht: (2025)
DVMNet++: Rethinking Relative Pose Estimation for Unseen Objects
von: Zhao, Chen, et al.
Veröffentlicht: (2024)
von: Zhao, Chen, et al.
Veröffentlicht: (2024)
SAM-Enhanced Segmentation on Road Datasets: Balancing Critical Classes in Autonomous Driving
von: Tahves, Toomas, et al.
Veröffentlicht: (2026)
von: Tahves, Toomas, et al.
Veröffentlicht: (2026)
Articulated Object Estimation in the Wild
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
DegustaBot: Zero-Shot Visual Preference Estimation for Personalized Multi-Object Rearrangement
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
One-Shot Dual-Arm Imitation Learning
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
Realtime Robust Shape Estimation of Deformable Linear Object
von: Zhang, Jiaming, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2024)
EvTTC: An Event Camera Dataset for Time-to-Collision Estimation
von: Sun, Kaizhen, et al.
Veröffentlicht: (2024)
von: Sun, Kaizhen, et al.
Veröffentlicht: (2024)
PokeFlex: A Real-World Dataset of Volumetric Deformable Objects for Robotics
von: Obrist, Jan, et al.
Veröffentlicht: (2024)
von: Obrist, Jan, et al.
Veröffentlicht: (2024)
UEVAVD: A Dataset for Developing UAV's Eye View Active Object Detection
von: Jiang, Xinhua, et al.
Veröffentlicht: (2024)
von: Jiang, Xinhua, et al.
Veröffentlicht: (2024)
AerialMind: Towards Referring Multi-Object Tracking in UAV Scenarios
von: Chen, Chenglizhao, et al.
Veröffentlicht: (2025)
von: Chen, Chenglizhao, et al.
Veröffentlicht: (2025)
AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs
von: Chang, Boyu, et al.
Veröffentlicht: (2026)
von: Chang, Boyu, et al.
Veröffentlicht: (2026)
DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects
von: Chen, Jiahong, et al.
Veröffentlicht: (2025)
von: Chen, Jiahong, et al.
Veröffentlicht: (2025)
RoadRunner -- Learning Traversability Estimation for Autonomous Off-road Driving
von: Frey, Jonas, et al.
Veröffentlicht: (2024)
von: Frey, Jonas, et al.
Veröffentlicht: (2024)
DOFS: A Real-world 3D Deformable Object Dataset with Full Spatial Information for Dynamics Model Learning
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
HOH: Markerless Multimodal Human-Object-Human Handover Dataset with Large Object Count
von: Wiederhold, Noah, et al.
Veröffentlicht: (2023)
von: Wiederhold, Noah, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
FBLNet: FeedBack Loop Network for Driver Attention Prediction
von: Chen, Yilong, et al.
Veröffentlicht: (2022) -
BiFold: Bimanual Cloth Folding with Language Guidance
von: Barbany, Oriol, et al.
Veröffentlicht: (2025) -
MI-DETR: An Object Detection Model with Multi-time Inquiries Mechanism
von: Nan, Zhixiong, et al.
Veröffentlicht: (2025) -
Top-Down Guidance for Learning Object-Centric Representations
von: Zou, Junhong, et al.
Veröffentlicht: (2024) -
DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model
von: Nan, Zhixiong, et al.
Veröffentlicht: (2024)