Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Xuan, Xie, Yuxuan, Liu, Yili, Lu, Haojian, Xiong, Rong, Liao, Yiyi, Wang, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
von: Yu, Xuan, et al.
Veröffentlicht: (2024)
von: Yu, Xuan, et al.
Veröffentlicht: (2024)
PanopticSplatting: End-to-End Panoptic Gaussian Splatting
von: Xie, Yuxuan, et al.
Veröffentlicht: (2025)
von: Xie, Yuxuan, et al.
Veröffentlicht: (2025)
UnIRe: Unsupervised Instance Decomposition for Dynamic Urban Scene Reconstruction
von: Mao, Yunxuan, et al.
Veröffentlicht: (2025)
von: Mao, Yunxuan, et al.
Veröffentlicht: (2025)
RING#: PR-by-PE Global Localization with Roto-translation Equivariant Gram Learning
von: Lu, Sha, et al.
Veröffentlicht: (2024)
von: Lu, Sha, et al.
Veröffentlicht: (2024)
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
von: Patel, Naman, et al.
Veröffentlicht: (2025)
von: Patel, Naman, et al.
Veröffentlicht: (2025)
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
von: Dong, Runpei, et al.
Veröffentlicht: (2026)
von: Dong, Runpei, et al.
Veröffentlicht: (2026)
$ν$-DBA: Neural Implicit Dense Bundle Adjustment Enables Image-Only Driving Scene Reconstruction
von: Mao, Yunxuan, et al.
Veröffentlicht: (2024)
von: Mao, Yunxuan, et al.
Veröffentlicht: (2024)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
Future Success Prediction in Open-Vocabulary Object Manipulation Tasks Based on End-Effector Trajectories
von: Kambara, Motonari, et al.
Veröffentlicht: (2024)
von: Kambara, Motonari, et al.
Veröffentlicht: (2024)
ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving
von: Lu, Han, et al.
Veröffentlicht: (2024)
von: Lu, Han, et al.
Veröffentlicht: (2024)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
Open-World Panoptic Segmentation
von: Sodano, Matteo, et al.
Veröffentlicht: (2024)
von: Sodano, Matteo, et al.
Veröffentlicht: (2024)
Collision Risk Estimation via Loss Prediction in End-to-End Autonomous Driving
von: Xiong, Ziliang, et al.
Veröffentlicht: (2025)
von: Xiong, Ziliang, et al.
Veröffentlicht: (2025)
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
An End-to-End Decision-Aware Multi-Scale Attention-Based Model for Explainable Autonomous Driving
von: Azad, Maryam Sadat Hosseini, et al.
Veröffentlicht: (2026)
von: Azad, Maryam Sadat Hosseini, et al.
Veröffentlicht: (2026)
MapTRv2: An End-to-End Framework for Online Vectorized HD Map Construction
von: Liao, Bencheng, et al.
Veröffentlicht: (2023)
von: Liao, Bencheng, et al.
Veröffentlicht: (2023)
DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving
von: Song, Ziying, et al.
Veröffentlicht: (2025)
von: Song, Ziying, et al.
Veröffentlicht: (2025)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
von: Wang, Tianqi, et al.
Veröffentlicht: (2024)
von: Wang, Tianqi, et al.
Veröffentlicht: (2024)
OpenEMMA: Open-Source Multimodal Model for End-to-End Autonomous Driving
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
Cognitive-Hierarchy Guided End-to-End Planning for Autonomous Driving
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
Exploring the Causality of End-to-End Autonomous Driving
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
BEVDiffLoc: End-to-End LiDAR Global Localization in BEV View based on Diffusion Model
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
Semantic Segmentation and Scene Reconstruction of RGB-D Image Frames: An End-to-End Modular Pipeline for Robotic Applications
von: Zheng, Zhiwu, et al.
Veröffentlicht: (2024)
von: Zheng, Zhiwu, et al.
Veröffentlicht: (2024)
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
RAP: 3D Rasterization Augmented End-to-End Planning
von: Feng, Lan, et al.
Veröffentlicht: (2025)
von: Feng, Lan, et al.
Veröffentlicht: (2025)
Unraveling the Effects of Synthetic Data on End-to-End Autonomous Driving
von: Ge, Junhao, et al.
Veröffentlicht: (2025)
von: Ge, Junhao, et al.
Veröffentlicht: (2025)
MeanFuser: Fast One-Step Multi-Modal Trajectory Generation and Adaptive Reconstruction via MeanFlow for End-to-End Autonomous Driving
von: Wang, Junli, et al.
Veröffentlicht: (2026)
von: Wang, Junli, et al.
Veröffentlicht: (2026)
UniUncer: Unified Dynamic Static Uncertainty for End to End Driving
von: Gao, Yu, et al.
Veröffentlicht: (2026)
von: Gao, Yu, et al.
Veröffentlicht: (2026)
PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement Learning
von: Gao, Hao, et al.
Veröffentlicht: (2025)
von: Gao, Hao, et al.
Veröffentlicht: (2025)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2026)
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2026)
SUPER-AD: Semantic Uncertainty-aware Planning for End-to-End Robust Autonomous Driving
von: Ryu, Wonjeong, et al.
Veröffentlicht: (2025)
von: Ryu, Wonjeong, et al.
Veröffentlicht: (2025)
GaussianFusion: Gaussian-Based Multi-Sensor Fusion for End-to-End Autonomous Driving
von: Liu, Shuai, et al.
Veröffentlicht: (2025)
von: Liu, Shuai, et al.
Veröffentlicht: (2025)
DriveSafer: End-to-End Autonomous Driving with Safety Guidance
von: Sural, Shounak, et al.
Veröffentlicht: (2026)
von: Sural, Shounak, et al.
Veröffentlicht: (2026)
ComDrive: Comfort-Oriented End-to-End Autonomous Driving
von: Wang, Junming, et al.
Veröffentlicht: (2024)
von: Wang, Junming, et al.
Veröffentlicht: (2024)
ScrewSplat: An End-to-End Method for Articulated Object Recognition
von: Kim, Seungyeon, et al.
Veröffentlicht: (2025)
von: Kim, Seungyeon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
von: Yu, Xuan, et al.
Veröffentlicht: (2024) -
PanopticSplatting: End-to-End Panoptic Gaussian Splatting
von: Xie, Yuxuan, et al.
Veröffentlicht: (2025) -
UnIRe: Unsupervised Instance Decomposition for Dynamic Urban Scene Reconstruction
von: Mao, Yunxuan, et al.
Veröffentlicht: (2025) -
RING#: PR-by-PE Global Localization with Roto-translation Equivariant Gram Learning
von: Lu, Sha, et al.
Veröffentlicht: (2024) -
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
von: Patel, Naman, et al.
Veröffentlicht: (2025)