MADrive: Memory-Augmented Driving Scene Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Karpikova, Polina, Selikhanovych, Daniil, Struminsky, Kirill, Musaev, Ruslan, Golitsyna, Maria, Baranchuk, Dmitry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CasTex: Cascaded Text-to-Texture Synthesis via Explicit Texture Maps and Physically-Based Shading
by: Aliev, Mishan, et al.
Published: (2025)
by: Aliev, Mishan, et al.
Published: (2025)
Inverse Bridge Matching Distillation
by: Gushchin, Nikita, et al.
Published: (2025)
by: Gushchin, Nikita, et al.
Published: (2025)
Differentiable Rendering with Reparameterized Volume Sampling
by: Morozov, Nikita, et al.
Published: (2023)
by: Morozov, Nikita, et al.
Published: (2023)
Accurate Compression of Text-to-Image Diffusion Models via Vector Quantization
by: Egiazarian, Vage, et al.
Published: (2024)
by: Egiazarian, Vage, et al.
Published: (2024)
Revisiting Autoregressive Models for Generative Image Classification
by: Sudakov, Ilia, et al.
Published: (2026)
by: Sudakov, Ilia, et al.
Published: (2026)
Your Student is Better Than Expected: Adaptive Teacher-Student Collaboration for Text-Conditional Diffusion Models
by: Starodubcev, Nikita, et al.
Published: (2023)
by: Starodubcev, Nikita, et al.
Published: (2023)
Scale-wise Distillation of Diffusion Models
by: Starodubcev, Nikita, et al.
Published: (2025)
by: Starodubcev, Nikita, et al.
Published: (2025)
Invertible Consistency Distillation for Text-Guided Image Editing in Around 7 Steps
by: Starodubcev, Nikita, et al.
Published: (2024)
by: Starodubcev, Nikita, et al.
Published: (2024)
SUPER: Selfie Undistortion and Head Pose Editing with Identity Preservation
by: Karpikova, Polina, et al.
Published: (2024)
by: Karpikova, Polina, et al.
Published: (2024)
Rethinking Global Text Conditioning in Diffusion Transformers
by: Starodubcev, Nikita, et al.
Published: (2026)
by: Starodubcev, Nikita, et al.
Published: (2026)
Alchemist: Turning Public Text-to-Image Data into Generative Gold
by: Startsev, Valerii, et al.
Published: (2025)
by: Startsev, Valerii, et al.
Published: (2025)
Registers Matter for Pixel-Space Diffusion Transformers
by: Starodubcev, Nikita, et al.
Published: (2026)
by: Starodubcev, Nikita, et al.
Published: (2026)
Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis
by: Voronov, Anton, et al.
Published: (2024)
by: Voronov, Anton, et al.
Published: (2024)
OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
A3D: Does Diffusion Dream about 3D Alignment?
by: Ignatyev, Savva, et al.
Published: (2024)
by: Ignatyev, Savva, et al.
Published: (2024)
One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation
by: Selikhanovych, Daniil, et al.
Published: (2025)
by: Selikhanovych, Daniil, et al.
Published: (2025)
Visual Implicit Geometry Transformer for Autonomous Driving
by: Shirokov, Arsenii, et al.
Published: (2026)
by: Shirokov, Arsenii, et al.
Published: (2026)
SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models
by: Makarov, Vladislav, et al.
Published: (2026)
by: Makarov, Vladislav, et al.
Published: (2026)
SuperPrimitive: Scene Reconstruction at a Primitive Level
by: Mazur, Kirill, et al.
Published: (2023)
by: Mazur, Kirill, et al.
Published: (2023)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
by: Shi, Chen, et al.
Published: (2025)
by: Shi, Chen, et al.
Published: (2025)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Effective Data Augmentation With Diffusion Models
by: Trabucco, Brandon, et al.
Published: (2023)
by: Trabucco, Brandon, et al.
Published: (2023)
SceneCrafter: Controllable Multi-View Driving Scene Editing
by: Zhu, Zehao, et al.
Published: (2025)
by: Zhu, Zehao, et al.
Published: (2025)
Generating Multimodal Driving Scenes via Next-Scene Prediction
by: Wu, Yanhao, et al.
Published: (2025)
by: Wu, Yanhao, et al.
Published: (2025)
UniScene: Unified Occupancy-centric Driving Scene Generation
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
FlowAD: Ego-Scene Interactive Modeling for Autonomous Driving
by: Guo, Mingzhe, et al.
Published: (2026)
by: Guo, Mingzhe, et al.
Published: (2026)
Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning
by: Li, Han, et al.
Published: (2026)
by: Li, Han, et al.
Published: (2026)
3DGraphLLM: Combining Semantic Graphs and Large Language Models for 3D Scene Understanding
by: Zemskova, Tatiana, et al.
Published: (2024)
by: Zemskova, Tatiana, et al.
Published: (2024)
DriveFix: Spatio-Temporally Coherent Driving Scene Restoration
by: Si, Heyu, et al.
Published: (2026)
by: Si, Heyu, et al.
Published: (2026)
HarmonicNeRF: Geometry-Informed Synthetic View Augmentation for 3D Scene Reconstruction in Driving Scenarios
by: Pan, Xiaochao, et al.
Published: (2023)
by: Pan, Xiaochao, et al.
Published: (2023)
Automatic Teaching Platform on Vision Language Retrieval Augmented Generation
by: Gokhman, Ruslan, et al.
Published: (2025)
by: Gokhman, Ruslan, et al.
Published: (2025)
DriveWorld: 4D Pre-trained Scene Understanding via World Models for Autonomous Driving
by: Min, Chen, et al.
Published: (2024)
by: Min, Chen, et al.
Published: (2024)
InsightDrive: Insight Scene Representation for End-to-End Autonomous Driving
by: Song, Ruiqi, et al.
Published: (2025)
by: Song, Ruiqi, et al.
Published: (2025)
DriveSplat: Unified Neural Gaussian Reconstruction for Dynamic Driving Scenes
by: Wang, Cong, et al.
Published: (2025)
by: Wang, Cong, et al.
Published: (2025)
FlexDrive: Toward Trajectory Flexibility in Driving Scene Reconstruction and Rendering
by: Zhou, Jingqiu, et al.
Published: (2025)
by: Zhou, Jingqiu, et al.
Published: (2025)
ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
by: Zhao, Guosheng, et al.
Published: (2025)
by: Zhao, Guosheng, et al.
Published: (2025)
4D Primitive-Mâché: Glueing Primitives for Persistent 4D Scene Reconstruction
by: Mazur, Kirill, et al.
Published: (2025)
by: Mazur, Kirill, et al.
Published: (2025)
What's in Common? Multimodal Models Hallucinate When Reasoning Across Scenes
by: Ross, Candace, et al.
Published: (2025)
by: Ross, Candace, et al.
Published: (2025)
DriveDreamer4D: World Models Are Effective Data Machines for 4D Driving Scene Representation
by: Zhao, Guosheng, et al.
Published: (2024)
by: Zhao, Guosheng, et al.
Published: (2024)
Weakly Supervised Semantic Segmentation for Driving Scenes
by: Kim, Dongseob, et al.
Published: (2023)
by: Kim, Dongseob, et al.
Published: (2023)
Similar Items
-
CasTex: Cascaded Text-to-Texture Synthesis via Explicit Texture Maps and Physically-Based Shading
by: Aliev, Mishan, et al.
Published: (2025) -
Inverse Bridge Matching Distillation
by: Gushchin, Nikita, et al.
Published: (2025) -
Differentiable Rendering with Reparameterized Volume Sampling
by: Morozov, Nikita, et al.
Published: (2023) -
Accurate Compression of Text-to-Image Diffusion Models via Vector Quantization
by: Egiazarian, Vage, et al.
Published: (2024) -
Revisiting Autoregressive Models for Generative Image Classification
by: Sudakov, Ilia, et al.
Published: (2026)