Saved in:
| Main Authors: | Song, Zhihang, Yao, Dingyi, Ming, Ruibo, Peng, Lihui, Yao, Danya, Zhang, Yi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.16848 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Timealign: A multi-modal object detection method for time misalignment fusing in autonomous driving
by: Song, Zhihang, et al.
Published: (2024)
by: Song, Zhihang, et al.
Published: (2024)
Synthetic Dataset Evaluation Based on Generalized Cross Validation
by: Song, Zhihang, et al.
Published: (2025)
by: Song, Zhihang, et al.
Published: (2025)
A Style-Based Profiling Framework for Quantifying the Synthetic-to-Real Gap in Autonomous Driving Datasets
by: Yao, Dingyi, et al.
Published: (2025)
by: Yao, Dingyi, et al.
Published: (2025)
Synthetic Datasets for Autonomous Driving: A Survey
by: Song, Zhihang, et al.
Published: (2023)
by: Song, Zhihang, et al.
Published: (2023)
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving
by: Lai, Zhihao, et al.
Published: (2024)
by: Lai, Zhihao, et al.
Published: (2024)
OMeGa: Joint Optimization of Explicit Meshes and Gaussian Splats for Robust Scene-Level Surface Reconstruction
by: Cao, Yuhang, et al.
Published: (2025)
by: Cao, Yuhang, et al.
Published: (2025)
An interactive enhanced driving dataset for autonomous driving
by: Feng, Haojie, et al.
Published: (2026)
by: Feng, Haojie, et al.
Published: (2026)
Optical aberrations in autonomous driving: Physics-informed parameterized temperature scaling for neural network uncertainty calibration
by: Wolf, Dominik Werner, et al.
Published: (2024)
by: Wolf, Dominik Werner, et al.
Published: (2024)
ARCON: Advancing Auto-Regressive Continuation for Driving Videos
by: Ming, Ruibo, et al.
Published: (2024)
by: Ming, Ruibo, et al.
Published: (2024)
X-Drive: Cross-modality consistent multi-sensor data synthesis for driving scenarios
by: Xie, Yichen, et al.
Published: (2024)
by: Xie, Yichen, et al.
Published: (2024)
COLA: COarse-LAbel multi-source LiDAR semantic segmentation for autonomous driving
by: Sanchez, Jules, et al.
Published: (2023)
by: Sanchez, Jules, et al.
Published: (2023)
A Survey on Future Frame Synthesis: Bridging Deterministic and Generative Approaches
by: Ming, Ruibo, et al.
Published: (2024)
by: Ming, Ruibo, et al.
Published: (2024)
Can Multi-modal (reasoning) LLMs detect document manipulation?
by: Liang, Zisheng, et al.
Published: (2025)
by: Liang, Zisheng, et al.
Published: (2025)
Visual Grounding with Multi-modal Conditional Adaptation
by: Yao, Ruilin, et al.
Published: (2024)
by: Yao, Ruilin, et al.
Published: (2024)
Learning autonomous driving from aerial imagery
by: Murali, Varun, et al.
Published: (2024)
by: Murali, Varun, et al.
Published: (2024)
SToRM: Supervised Token Reduction for Multi-modal LLMs toward efficient end-to-end autonomous driving
by: Kim, Seo Hyun, et al.
Published: (2026)
by: Kim, Seo Hyun, et al.
Published: (2026)
SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment
by: Mao, Xinyu, et al.
Published: (2025)
by: Mao, Xinyu, et al.
Published: (2025)
On depth prediction for autonomous driving using self-supervised learning
by: Boulahbal, Houssem
Published: (2024)
by: Boulahbal, Houssem
Published: (2024)
Vision-based 3D occupancy prediction in autonomous driving: a review and outlook
by: Zhang, Yanan, et al.
Published: (2024)
by: Zhang, Yanan, et al.
Published: (2024)
Towards multi-modal forgery representation learning for AI-generated video detection and localization
by: Le, Dat, et al.
Published: (2026)
by: Le, Dat, et al.
Published: (2026)
Model alignment using inter-modal bridges
by: Gholamzadeh, Ali, et al.
Published: (2025)
by: Gholamzadeh, Ali, et al.
Published: (2025)
A method for detecting dead fish on large water surfaces based on improved YOLOv10
by: Tian, Qingbin, et al.
Published: (2024)
by: Tian, Qingbin, et al.
Published: (2024)
Intelligent driving vehicle front multi-target tracking and detection based on YOLOv5 and point cloud 3D projection
by: Liu, Dayong, et al.
Published: (2025)
by: Liu, Dayong, et al.
Published: (2025)
Multi-model approach for autonomous driving: A comprehensive study on traffic sign-, vehicle- and lane detection and behavioral cloning
by: Jaisankar, Kanishkha, et al.
Published: (2026)
by: Jaisankar, Kanishkha, et al.
Published: (2026)
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
by: Jeon, Byungwoo, et al.
Published: (2026)
by: Jeon, Byungwoo, et al.
Published: (2026)
BuckTales : A multi-UAV dataset for multi-object tracking and re-identification of wild antelopes
by: Naik, Hemal, et al.
Published: (2024)
by: Naik, Hemal, et al.
Published: (2024)
A Prediction-as-Perception Framework for 3D Object Detection
by: Zhang, Song, et al.
Published: (2026)
by: Zhang, Song, et al.
Published: (2026)
MV-DETR: Multi-modality indoor object detection by Multi-View DEtecton TRansformers
by: Dong, Zichao, et al.
Published: (2024)
by: Dong, Zichao, et al.
Published: (2024)
GaussianCross: Cross-modal Self-supervised 3D Representation Learning via Gaussian Splatting
by: Yao, Lei, et al.
Published: (2025)
by: Yao, Lei, et al.
Published: (2025)
On the dynamic evolution of CLIP texture-shape bias and its relationship to human alignment and model robustness
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
Are NeRFs ready for autonomous driving? Towards closing the real-to-simulation gap
by: Lindström, Carl, et al.
Published: (2024)
by: Lindström, Carl, et al.
Published: (2024)
Research on target detection method of distracted driving behavior based on improved YOLOv8
by: Shen, Shiquan, et al.
Published: (2024)
by: Shen, Shiquan, et al.
Published: (2024)
YOLOv1 to YOLOv10: The fastest and most accurate real-time object detection systems
by: Wang, Chien-Yao, et al.
Published: (2024)
by: Wang, Chien-Yao, et al.
Published: (2024)
Mirasol3B: A Multimodal Autoregressive model for time-aligned and contextual modalities
by: Piergiovanni, AJ, et al.
Published: (2023)
by: Piergiovanni, AJ, et al.
Published: (2023)
AffordanceGrasp-R1:Leveraging Reasoning-Based Affordance Segmentation with Reinforcement Learning for Robotic Grasping
by: Zhou, Dingyi, et al.
Published: (2026)
by: Zhou, Dingyi, et al.
Published: (2026)
Weakly supervised alignment and registration of MR-CT for cervical cancer radiotherapy
by: Zhang, Jjahao, et al.
Published: (2024)
by: Zhang, Jjahao, et al.
Published: (2024)
Towards learning-based planning:The nuPlan benchmark for real-world autonomous driving
by: Karnchanachari, Napat, et al.
Published: (2024)
by: Karnchanachari, Napat, et al.
Published: (2024)
Domain adaptive pose estimation via multi-level alignment
by: Chen, Yugan, et al.
Published: (2024)
by: Chen, Yugan, et al.
Published: (2024)
One target to align them all: LiDAR, RGB and event cameras extrinsic calibration for Autonomous Driving
by: Bertogalli, Andrea, et al.
Published: (2025)
by: Bertogalli, Andrea, et al.
Published: (2025)
Fusion-then-Distillation: Toward Cross-modal Positive Distillation for Domain Adaptive 3D Semantic Segmentation
by: Wu, Yao, et al.
Published: (2024)
by: Wu, Yao, et al.
Published: (2024)
Similar Items
-
Timealign: A multi-modal object detection method for time misalignment fusing in autonomous driving
by: Song, Zhihang, et al.
Published: (2024) -
Synthetic Dataset Evaluation Based on Generalized Cross Validation
by: Song, Zhihang, et al.
Published: (2025) -
A Style-Based Profiling Framework for Quantifying the Synthetic-to-Real Gap in Autonomous Driving Datasets
by: Yao, Dingyi, et al.
Published: (2025) -
Synthetic Datasets for Autonomous Driving: A Survey
by: Song, Zhihang, et al.
Published: (2023) -
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving
by: Lai, Zhihao, et al.
Published: (2024)