LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Yun, Pittaluga, Francesco, Jiang, Ziyu, Zwicker, Matthias, Chandraker, Manmohan, Tasneem, Zaid |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes
von: Soroco, Mauricio, et al.
Veröffentlicht: (2026)
von: Soroco, Mauricio, et al.
Veröffentlicht: (2026)
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
von: Sharan, S P, et al.
Veröffentlicht: (2023)
von: Sharan, S P, et al.
Veröffentlicht: (2023)
Drive-1-to-3: Enriching Diffusion Priors for Novel View Synthesis of Real Vehicles
von: Lin, Chuang, et al.
Veröffentlicht: (2024)
von: Lin, Chuang, et al.
Veröffentlicht: (2024)
SAFE-SIM: Safety-Critical Closed-Loop Traffic Simulation with Diffusion-Controllable Adversaries
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2023)
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2023)
PhyCo: Learning Controllable Physical Priors for Generative Motion
von: Narayanan, Sriram, et al.
Veröffentlicht: (2026)
von: Narayanan, Sriram, et al.
Veröffentlicht: (2026)
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
von: Khan, Zaid, et al.
Veröffentlicht: (2024)
von: Khan, Zaid, et al.
Veröffentlicht: (2024)
LidaRF: Delving into Lidar for Neural Radiance Field on Street Scenes
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
SceneCrafter: Controllable Multi-View Driving Scene Editing
von: Zhu, Zehao, et al.
Veröffentlicht: (2025)
von: Zhu, Zehao, et al.
Veröffentlicht: (2025)
CTRL-D: Controllable Dynamic 3D Scene Editing with Personalized 2D Diffusion
von: He, Kai, et al.
Veröffentlicht: (2024)
von: He, Kai, et al.
Veröffentlicht: (2024)
Tuned Contrastive Learning
von: Animesh, Chaitanya, et al.
Veröffentlicht: (2023)
von: Animesh, Chaitanya, et al.
Veröffentlicht: (2023)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation
von: He, Yun, et al.
Veröffentlicht: (2026)
von: He, Yun, et al.
Veröffentlicht: (2026)
LANGTRAJ: Diffusion Model and Dataset for Language-Conditioned Trajectory Simulation
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2025)
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2025)
AutoScape: Geometry-Consistent Long-Horizon Scene Generation
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
NERFIFY: A Multi-Agent Framework for Turning NeRF Papers into Code
von: Jain, Seemandhar, et al.
Veröffentlicht: (2026)
von: Jain, Seemandhar, et al.
Veröffentlicht: (2026)
AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving
von: Liang, Mingfu, et al.
Veröffentlicht: (2024)
von: Liang, Mingfu, et al.
Veröffentlicht: (2024)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
Locally Orderless Images for Optimization in Differentiable Rendering
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
UDA-Bench: Revisiting Common Assumptions in Unsupervised Domain Adaptation Using a Standardized Framework
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
LangCoop: Collaborative Driving with Language
von: Gao, Xiangbo, et al.
Veröffentlicht: (2025)
von: Gao, Xiangbo, et al.
Veröffentlicht: (2025)
DriveEditor: A Unified 3D Information-Guided Framework for Controllable Object Editing in Driving Scenes
von: Liang, Yiyuan, et al.
Veröffentlicht: (2024)
von: Liang, Yiyuan, et al.
Veröffentlicht: (2024)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
Physics-Aware 3D Gaussian Editing for Driving Scene Generation
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
von: Qian, Tianwen, et al.
Veröffentlicht: (2023)
von: Qian, Tianwen, et al.
Veröffentlicht: (2023)
Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion
von: Li, Haodong, et al.
Veröffentlicht: (2026)
von: Li, Haodong, et al.
Veröffentlicht: (2026)
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
von: Aich, Abhishek, et al.
Veröffentlicht: (2024)
von: Aich, Abhishek, et al.
Veröffentlicht: (2024)
SIMSplat: Predictive Driving Scene Editing with Language-aligned 4D Gaussian Splatting
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
Distilling Multi-modal Large Language Models for Autonomous Driving
von: Hegde, Deepti, et al.
Veröffentlicht: (2025)
von: Hegde, Deepti, et al.
Veröffentlicht: (2025)
Instantaneous Perception of Moving Objects in 3D
von: Liu, Di, et al.
Veröffentlicht: (2024)
von: Liu, Di, et al.
Veröffentlicht: (2024)
DrivingGPT: Unifying Driving World Modeling and Planning with Multi-modal Autoregressive Transformers
von: Chen, Yuntao, et al.
Veröffentlicht: (2024)
von: Chen, Yuntao, et al.
Veröffentlicht: (2024)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
CTRL-O: Language-Controllable Object-Centric Visual Representation Learning
von: Didolkar, Aniket, et al.
Veröffentlicht: (2025)
von: Didolkar, Aniket, et al.
Veröffentlicht: (2025)
InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes
von: Liu, Hongyuan, et al.
Veröffentlicht: (2025)
von: Liu, Hongyuan, et al.
Veröffentlicht: (2025)
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
von: Vobecky, Antonin, et al.
Veröffentlicht: (2022)
von: Vobecky, Antonin, et al.
Veröffentlicht: (2022)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
von: Shi, Chen, et al.
Veröffentlicht: (2025)
von: Shi, Chen, et al.
Veröffentlicht: (2025)
Sce2DriveX: A Generalized MLLM Framework for Scene-to-Drive Learning
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
Mirage: One-Step Video Diffusion for Photorealistic and Coherent Asset Editing in Driving Scenes
von: Wang, Shuyun, et al.
Veröffentlicht: (2025)
von: Wang, Shuyun, et al.
Veröffentlicht: (2025)
OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving
von: Zhang, Zhenguo, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenguo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
von: Wang, Yifan, et al.
Veröffentlicht: (2026) -
HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes
von: Soroco, Mauricio, et al.
Veröffentlicht: (2026) -
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026) -
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
von: Sharan, S P, et al.
Veröffentlicht: (2023) -
Drive-1-to-3: Enriching Diffusion Priors for Novel View Synthesis of Real Vehicles
von: Lin, Chuang, et al.
Veröffentlicht: (2024)