DriveGEN: Generalized and Robust 3D Detection in Driving via Controllable Text-to-Image Diffusion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Hongbin, Guo, Zilu, Zhang, Yifan, Niu, Shuaicheng, Li, Yafeng, Zhang, Ruimao, Cui, Shuguang, Li, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DriveFlow: Rectified Flow Adaptation for Robust 3D Object Detection in Autonomous Driving
by: Lin, Hongbin, et al.
Published: (2025)
by: Lin, Hongbin, et al.
Published: (2025)
Fully Test-Time Adaptation for Monocular 3D Object Detection
by: Lin, Hongbin, et al.
Published: (2024)
by: Lin, Hongbin, et al.
Published: (2024)
FutureX: Enhance End-to-End Autonomous Driving via Latent Chain-of-Thought World Model
by: Lin, Hongbin, et al.
Published: (2025)
by: Lin, Hongbin, et al.
Published: (2025)
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
by: Yang, Yiming, et al.
Published: (2025)
by: Yang, Yiming, et al.
Published: (2025)
Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation
by: Zhan, Yifan, et al.
Published: (2026)
by: Zhan, Yifan, et al.
Published: (2026)
Text2Traffic: A Text-to-Image Generation and Editing Method for Traffic Scenes
by: Lv, Feng, et al.
Published: (2025)
by: Lv, Feng, et al.
Published: (2025)
Unlock the Power of Unlabeled Data in Language Driving Model
by: Wang, Chaoqun, et al.
Published: (2025)
by: Wang, Chaoqun, et al.
Published: (2025)
VistaGEN: Consistent Driving Video Generation with Fine-Grained Control Using Multiview Visual-Language Reasoning
by: Chen, Li-Heng, et al.
Published: (2026)
by: Chen, Li-Heng, et al.
Published: (2026)
Risk-Controllable Multi-View Diffusion for Driving Scenario Generation
by: Lin, Hongyi, et al.
Published: (2026)
by: Lin, Hongyi, et al.
Published: (2026)
DV-3DLane: End-to-end Multi-modal 3D Lane Detection with Dual-view Representation
by: Luo, Yueru, et al.
Published: (2024)
by: Luo, Yueru, et al.
Published: (2024)
LLM4GEN: Leveraging Semantic Representation of LLMs for Text-to-Image Generation
by: Liu, Mushui, et al.
Published: (2024)
by: Liu, Mushui, et al.
Published: (2024)
SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2025)
by: Zhang, Haiming, et al.
Published: (2025)
WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation
by: Lu, Jiachen, et al.
Published: (2023)
by: Lu, Jiachen, et al.
Published: (2023)
Reproducibility Study of "ITI-GEN: Inclusive Text-to-Image Generation"
by: Fernández, Daniel Gallo, et al.
Published: (2024)
by: Fernández, Daniel Gallo, et al.
Published: (2024)
VALL-T: Decoder-Only Generative Transducer for Robust and Decoding-Controllable Text-to-Speech
by: Du, Chenpeng, et al.
Published: (2024)
by: Du, Chenpeng, et al.
Published: (2024)
DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion
by: Wang, Weijie, et al.
Published: (2025)
by: Wang, Weijie, et al.
Published: (2025)
RelTopo: Multi-Level Relational Modeling for Driving Scene Topology Reasoning
by: Luo, Yueru, et al.
Published: (2025)
by: Luo, Yueru, et al.
Published: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
by: Cai, Kaiwen, et al.
Published: (2025)
by: Cai, Kaiwen, et al.
Published: (2025)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
by: Ma, Jiahua, et al.
Published: (2025)
by: Ma, Jiahua, et al.
Published: (2025)
How Learning Dynamics Drive Adversarially Robust Generalization?
by: Xu, Yuelin, et al.
Published: (2024)
by: Xu, Yuelin, et al.
Published: (2024)
Robustness-Aware 3D Object Detection in Autonomous Driving: A Review and Outlook
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
by: Liao, Bencheng, et al.
Published: (2024)
by: Liao, Bencheng, et al.
Published: (2024)
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
by: Huang, Binyuan, et al.
Published: (2024)
by: Huang, Binyuan, et al.
Published: (2024)
DriveFine: Refining-Augmented Masked Diffusion VLA for Precise and Robust Driving
by: Dang, Chenxu, et al.
Published: (2026)
by: Dang, Chenxu, et al.
Published: (2026)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Text2Scenario: Text-Driven Scenario Generation for Autonomous Driving Test
by: Cai, Xuan, et al.
Published: (2025)
by: Cai, Xuan, et al.
Published: (2025)
VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2024)
by: Zhang, Haiming, et al.
Published: (2024)
Semantic-Supervised Spatial-Temporal Fusion for LiDAR-based 3D Object Detection
by: Wang, Chaoqun, et al.
Published: (2025)
by: Wang, Chaoqun, et al.
Published: (2025)
Single Image Rolling Shutter Removal with Diffusion Models
by: Yang, Zhanglei, et al.
Published: (2024)
by: Yang, Zhanglei, et al.
Published: (2024)
Diffusion-Based Generative Models for 3D Occupancy Prediction in Autonomous Driving
by: Wang, Yunshen, et al.
Published: (2025)
by: Wang, Yunshen, et al.
Published: (2025)
4D Driving Scene Generation With Stereo Forcing
by: Lu, Hao, et al.
Published: (2025)
by: Lu, Hao, et al.
Published: (2025)
One-Shot Diffusion Mimicker for Handwritten Text Generation
by: Dai, Gang, et al.
Published: (2024)
by: Dai, Gang, et al.
Published: (2024)
DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation
by: Guo, Jiazhe, et al.
Published: (2025)
by: Guo, Jiazhe, et al.
Published: (2025)
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
by: Zhang, Haiming, et al.
Published: (2026)
by: Zhang, Haiming, et al.
Published: (2026)
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
by: Wu, Wei, et al.
Published: (2024)
by: Wu, Wei, et al.
Published: (2024)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
by: Zheng, Chaoda, et al.
Published: (2024)
by: Zheng, Chaoda, et al.
Published: (2024)
Multi-modal Traffic Scenario Generation for Autonomous Driving System Testing
by: Tu, Zhi, et al.
Published: (2025)
by: Tu, Zhi, et al.
Published: (2025)
A Generative Car-following Model Conditioned On Driving Styles
by: Zhang, Yifan, et al.
Published: (2021)
by: Zhang, Yifan, et al.
Published: (2021)
HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation
by: Zhang, Conglang, et al.
Published: (2026)
by: Zhang, Conglang, et al.
Published: (2026)
Similar Items
-
DriveFlow: Rectified Flow Adaptation for Robust 3D Object Detection in Autonomous Driving
by: Lin, Hongbin, et al.
Published: (2025) -
Fully Test-Time Adaptation for Monocular 3D Object Detection
by: Lin, Hongbin, et al.
Published: (2024) -
FutureX: Enhance End-to-End Autonomous Driving via Latent Chain-of-Thought World Model
by: Lin, Hongbin, et al.
Published: (2025) -
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
by: Yang, Yiming, et al.
Published: (2025) -
Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation
by: Zhan, Yifan, et al.
Published: (2026)