DriveGEN: Generalized and Robust 3D Detection in Driving via Controllable Text-to-Image Diffusion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Hongbin, Guo, Zilu, Zhang, Yifan, Niu, Shuaicheng, Li, Yafeng, Zhang, Ruimao, Cui, Shuguang, Li, Zhen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DriveFlow: Rectified Flow Adaptation for Robust 3D Object Detection in Autonomous Driving
von: Lin, Hongbin, et al.
Veröffentlicht: (2025)
von: Lin, Hongbin, et al.
Veröffentlicht: (2025)
Fully Test-Time Adaptation for Monocular 3D Object Detection
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
FutureX: Enhance End-to-End Autonomous Driving via Latent Chain-of-Thought World Model
von: Lin, Hongbin, et al.
Veröffentlicht: (2025)
von: Lin, Hongbin, et al.
Veröffentlicht: (2025)
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
von: Yang, Yiming, et al.
Veröffentlicht: (2025)
von: Yang, Yiming, et al.
Veröffentlicht: (2025)
Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation
von: Zhan, Yifan, et al.
Veröffentlicht: (2026)
von: Zhan, Yifan, et al.
Veröffentlicht: (2026)
Text2Traffic: A Text-to-Image Generation and Editing Method for Traffic Scenes
von: Lv, Feng, et al.
Veröffentlicht: (2025)
von: Lv, Feng, et al.
Veröffentlicht: (2025)
Unlock the Power of Unlabeled Data in Language Driving Model
von: Wang, Chaoqun, et al.
Veröffentlicht: (2025)
von: Wang, Chaoqun, et al.
Veröffentlicht: (2025)
VistaGEN: Consistent Driving Video Generation with Fine-Grained Control Using Multiview Visual-Language Reasoning
von: Chen, Li-Heng, et al.
Veröffentlicht: (2026)
von: Chen, Li-Heng, et al.
Veröffentlicht: (2026)
Risk-Controllable Multi-View Diffusion for Driving Scenario Generation
von: Lin, Hongyi, et al.
Veröffentlicht: (2026)
von: Lin, Hongyi, et al.
Veröffentlicht: (2026)
DV-3DLane: End-to-end Multi-modal 3D Lane Detection with Dual-view Representation
von: Luo, Yueru, et al.
Veröffentlicht: (2024)
von: Luo, Yueru, et al.
Veröffentlicht: (2024)
LLM4GEN: Leveraging Semantic Representation of LLMs for Text-to-Image Generation
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous Driving
von: Zhang, Haiming, et al.
Veröffentlicht: (2025)
von: Zhang, Haiming, et al.
Veröffentlicht: (2025)
WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation
von: Lu, Jiachen, et al.
Veröffentlicht: (2023)
von: Lu, Jiachen, et al.
Veröffentlicht: (2023)
Reproducibility Study of "ITI-GEN: Inclusive Text-to-Image Generation"
von: Fernández, Daniel Gallo, et al.
Veröffentlicht: (2024)
von: Fernández, Daniel Gallo, et al.
Veröffentlicht: (2024)
VALL-T: Decoder-Only Generative Transducer for Robust and Decoding-Controllable Text-to-Speech
von: Du, Chenpeng, et al.
Veröffentlicht: (2024)
von: Du, Chenpeng, et al.
Veröffentlicht: (2024)
DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
RelTopo: Multi-Level Relational Modeling for Driving Scene Topology Reasoning
von: Luo, Yueru, et al.
Veröffentlicht: (2025)
von: Luo, Yueru, et al.
Veröffentlicht: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
von: Ma, Jiahua, et al.
Veröffentlicht: (2025)
von: Ma, Jiahua, et al.
Veröffentlicht: (2025)
How Learning Dynamics Drive Adversarially Robust Generalization?
von: Xu, Yuelin, et al.
Veröffentlicht: (2024)
von: Xu, Yuelin, et al.
Veröffentlicht: (2024)
Robustness-Aware 3D Object Detection in Autonomous Driving: A Review and Outlook
von: Song, Ziying, et al.
Veröffentlicht: (2024)
von: Song, Ziying, et al.
Veröffentlicht: (2024)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
DriveFine: Refining-Augmented Masked Diffusion VLA for Precise and Robust Driving
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Text2Scenario: Text-Driven Scenario Generation for Autonomous Driving Test
von: Cai, Xuan, et al.
Veröffentlicht: (2025)
von: Cai, Xuan, et al.
Veröffentlicht: (2025)
VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving
von: Zhang, Haiming, et al.
Veröffentlicht: (2024)
von: Zhang, Haiming, et al.
Veröffentlicht: (2024)
Semantic-Supervised Spatial-Temporal Fusion for LiDAR-based 3D Object Detection
von: Wang, Chaoqun, et al.
Veröffentlicht: (2025)
von: Wang, Chaoqun, et al.
Veröffentlicht: (2025)
Single Image Rolling Shutter Removal with Diffusion Models
von: Yang, Zhanglei, et al.
Veröffentlicht: (2024)
von: Yang, Zhanglei, et al.
Veröffentlicht: (2024)
Diffusion-Based Generative Models for 3D Occupancy Prediction in Autonomous Driving
von: Wang, Yunshen, et al.
Veröffentlicht: (2025)
von: Wang, Yunshen, et al.
Veröffentlicht: (2025)
4D Driving Scene Generation With Stereo Forcing
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
One-Shot Diffusion Mimicker for Handwritten Text Generation
von: Dai, Gang, et al.
Veröffentlicht: (2024)
von: Dai, Gang, et al.
Veröffentlicht: (2024)
DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation
von: Guo, Jiazhe, et al.
Veröffentlicht: (2025)
von: Guo, Jiazhe, et al.
Veröffentlicht: (2025)
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
von: Wu, Wei, et al.
Veröffentlicht: (2024)
von: Wu, Wei, et al.
Veröffentlicht: (2024)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
von: Zheng, Chaoda, et al.
Veröffentlicht: (2024)
von: Zheng, Chaoda, et al.
Veröffentlicht: (2024)
Multi-modal Traffic Scenario Generation for Autonomous Driving System Testing
von: Tu, Zhi, et al.
Veröffentlicht: (2025)
von: Tu, Zhi, et al.
Veröffentlicht: (2025)
A Generative Car-following Model Conditioned On Driving Styles
von: Zhang, Yifan, et al.
Veröffentlicht: (2021)
von: Zhang, Yifan, et al.
Veröffentlicht: (2021)
HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation
von: Zhang, Conglang, et al.
Veröffentlicht: (2026)
von: Zhang, Conglang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DriveFlow: Rectified Flow Adaptation for Robust 3D Object Detection in Autonomous Driving
von: Lin, Hongbin, et al.
Veröffentlicht: (2025) -
Fully Test-Time Adaptation for Monocular 3D Object Detection
von: Lin, Hongbin, et al.
Veröffentlicht: (2024) -
FutureX: Enhance End-to-End Autonomous Driving via Latent Chain-of-Thought World Model
von: Lin, Hongbin, et al.
Veröffentlicht: (2025) -
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
von: Yang, Yiming, et al.
Veröffentlicht: (2025) -
Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation
von: Zhan, Yifan, et al.
Veröffentlicht: (2026)