NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Kai, Wang, Xu, Fan, Rui, Yang, Kailun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation
von: Qin, Yu, et al.
Veröffentlicht: (2026)
von: Qin, Yu, et al.
Veröffentlicht: (2026)
DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
Omnidirectional Multi-Object Tracking
von: Luo, Kai, et al.
Veröffentlicht: (2025)
von: Luo, Kai, et al.
Veröffentlicht: (2025)
OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback
von: Luo, Kai, et al.
Veröffentlicht: (2025)
von: Luo, Kai, et al.
Veröffentlicht: (2025)
O3N: Omnidirectional Open-Vocabulary Occupancy Prediction
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
OccTrack360: 4D Panoptic Occupancy Tracking from Surround-View Fisheye Cameras
von: Lin, Yongzhi, et al.
Veröffentlicht: (2026)
von: Lin, Yongzhi, et al.
Veröffentlicht: (2026)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
Offboard Occupancy Refinement with Hybrid Propagation for Autonomous Driving
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Towards Consistent Object Detection via LiDAR-Camera Synergy
von: Luo, Kai, et al.
Veröffentlicht: (2024)
von: Luo, Kai, et al.
Veröffentlicht: (2024)
Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
von: He, Xuan, et al.
Veröffentlicht: (2023)
von: He, Xuan, et al.
Veröffentlicht: (2023)
Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2024)
von: Yang, Fan, et al.
Veröffentlicht: (2024)
DeProPose: Deficiency-Proof 3D Human Pose Estimation via Adaptive Multi-View Fusion
von: Jiao, Jianbin, et al.
Veröffentlicht: (2025)
von: Jiao, Jianbin, et al.
Veröffentlicht: (2025)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
von: Shi, Hao, et al.
Veröffentlicht: (2023)
von: Shi, Hao, et al.
Veröffentlicht: (2023)
Multi-Keypoint Affordance Representation for Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
Towards Precise 3D Human Pose Estimation with Multi-Perspective Spatial-Temporal Relational Transformers
von: Jiao, Jianbin, et al.
Veröffentlicht: (2024)
von: Jiao, Jianbin, et al.
Veröffentlicht: (2024)
Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise
von: Li, Wenxin, et al.
Veröffentlicht: (2026)
von: Li, Wenxin, et al.
Veröffentlicht: (2026)
Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance
von: Zhao, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiayi, et al.
Veröffentlicht: (2025)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
von: Jia, Wanjun, et al.
Veröffentlicht: (2025)
von: Jia, Wanjun, et al.
Veröffentlicht: (2025)
CoBEVMoE: Heterogeneity-aware Feature Fusion with Dynamic Mixture-of-Experts for Collaborative Perception
von: Kong, Lingzhao, et al.
Veröffentlicht: (2025)
von: Kong, Lingzhao, et al.
Veröffentlicht: (2025)
LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection
von: Teng, Fei, et al.
Veröffentlicht: (2025)
von: Teng, Fei, et al.
Veröffentlicht: (2025)
NRSeg: Noise-Resilient Learning for BEV Semantic Segmentation via Driving World Models
von: Li, Siyu, et al.
Veröffentlicht: (2025)
von: Li, Siyu, et al.
Veröffentlicht: (2025)
EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras
von: Wang, Luming, et al.
Veröffentlicht: (2026)
von: Wang, Luming, et al.
Veröffentlicht: (2026)
Language-Driven Dual Style Mixing for Single-Domain Generalized Object Detection
von: Qin, Hongda, et al.
Veröffentlicht: (2025)
von: Qin, Hongda, et al.
Veröffentlicht: (2025)
A Compendium of Autonomous Navigation using Object Detection and Tracking in Unmanned Aerial Vehicles
von: Arora, Mohit, et al.
Veröffentlicht: (2025)
von: Arora, Mohit, et al.
Veröffentlicht: (2025)
QuaDreamer: Controllable Panoramic Video Generation for Quadruped Robots
von: Wu, Sheng, et al.
Veröffentlicht: (2025)
von: Wu, Sheng, et al.
Veröffentlicht: (2025)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots
von: Zhao, Guoqiang, et al.
Veröffentlicht: (2026)
von: Zhao, Guoqiang, et al.
Veröffentlicht: (2026)
Event-guided 3D Gaussian Splatting for Dynamic Human and Scene Reconstruction
von: Yin, Xiaoting, et al.
Veröffentlicht: (2025)
von: Yin, Xiaoting, et al.
Veröffentlicht: (2025)
InterEdit: Navigating Text-Guided Multi-Human 3D Motion Editing
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras
von: Teng, Fei, et al.
Veröffentlicht: (2024)
von: Teng, Fei, et al.
Veröffentlicht: (2024)
Panoramic Out-of-Distribution Segmentation
von: Duan, Mengfei, et al.
Veröffentlicht: (2025)
von: Duan, Mengfei, et al.
Veröffentlicht: (2025)
Hallucinating 360°: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting
von: Teng, Fei, et al.
Veröffentlicht: (2025)
von: Teng, Fei, et al.
Veröffentlicht: (2025)
Out-of-Distribution Semantic Occupancy Prediction
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
P2U-SLAM: A Monocular Wide-FoV SLAM System Based on Point Uncertainty and Pose Uncertainty
von: Zhang, Yufan, et al.
Veröffentlicht: (2024)
von: Zhang, Yufan, et al.
Veröffentlicht: (2024)
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather
von: Li, Haoyuan, et al.
Veröffentlicht: (2024)
von: Li, Haoyuan, et al.
Veröffentlicht: (2024)
Exploring Event-based Human Pose Estimation with 3D Event Representations
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
DriveNetBench: An Affordable and Configurable Single-Camera Benchmarking System for Autonomous Driving Networks
von: Al-Bustami, Ali, et al.
Veröffentlicht: (2025)
von: Al-Bustami, Ali, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024) -
Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation
von: Qin, Yu, et al.
Veröffentlicht: (2026) -
DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking
von: Deng, Buyin, et al.
Veröffentlicht: (2025) -
Omnidirectional Multi-Object Tracking
von: Luo, Kai, et al.
Veröffentlicht: (2025) -
OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback
von: Luo, Kai, et al.
Veröffentlicht: (2025)