Visual Implicit Geometry Transformer for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shirokov, Arsenii, Kuznetsov, Mikhail, Stepochkin, Danila, Evdokimov, Egor, Glazkov, Daniil, Patakin, Nikolay, Konushin, Anton, Senushkin, Dmitry |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unified Sensor Simulation for Autonomous Driving
von: Patakin, Nikolay, et al.
Veröffentlicht: (2026)
von: Patakin, Nikolay, et al.
Veröffentlicht: (2026)
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task
von: Gabdullin, Bulat, et al.
Veröffentlicht: (2024)
von: Gabdullin, Bulat, et al.
Veröffentlicht: (2024)
A3D: Does Diffusion Dream about 3D Alignment?
von: Ignatyev, Savva, et al.
Veröffentlicht: (2024)
von: Ignatyev, Savva, et al.
Veröffentlicht: (2024)
Z3D: Zero-Shot 3D Visual Grounding from Images
von: Drozdov, Nikita, et al.
Veröffentlicht: (2026)
von: Drozdov, Nikita, et al.
Veröffentlicht: (2026)
Cooperative Face Liveness Detection from Optical Flow
von: Sokolov, Artem, et al.
Veröffentlicht: (2025)
von: Sokolov, Artem, et al.
Veröffentlicht: (2025)
UniDet3D: Multi-dataset Indoor 3D Object Detection
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2024)
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2024)
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
von: Lemeshko, Andrey, et al.
Veröffentlicht: (2025)
von: Lemeshko, Andrey, et al.
Veröffentlicht: (2025)
DynaMix: Generalizable Person Re-identification via Dynamic Relabeling and Mixed Data Sampling
von: Mamedov, Timur, et al.
Veröffentlicht: (2025)
von: Mamedov, Timur, et al.
Veröffentlicht: (2025)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
von: Mamedov, Timur, et al.
Veröffentlicht: (2024)
von: Mamedov, Timur, et al.
Veröffentlicht: (2024)
DriveVGGT: Calibration-Constrained Visual Geometry Transformers for Multi-Camera Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
TUN3D: Towards Real-World Scene Understanding from Unposed Images
von: Konushin, Anton, et al.
Veröffentlicht: (2025)
von: Konushin, Anton, et al.
Veröffentlicht: (2025)
DVGT: Driving Visual Geometry Transformer
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis
von: Voronov, Anton, et al.
Veröffentlicht: (2024)
von: Voronov, Anton, et al.
Veröffentlicht: (2024)
HairFastGAN: Realistic and Robust Hair Transfer with a Fast Encoder-Based Approach
von: Nikolaev, Maxim, et al.
Veröffentlicht: (2024)
von: Nikolaev, Maxim, et al.
Veröffentlicht: (2024)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
von: Kolodiazhnyi, Maksim, et al.
Veröffentlicht: (2025)
MADrive: Memory-Augmented Driving Scene Modeling
von: Karpikova, Polina, et al.
Veröffentlicht: (2025)
von: Karpikova, Polina, et al.
Veröffentlicht: (2025)
Towards properties of adversarial image perturbations
von: Kuznetsov, Egor, et al.
Veröffentlicht: (2025)
von: Kuznetsov, Egor, et al.
Veröffentlicht: (2025)
DriveDiTFit: Fine-tuning Diffusion Transformers for Autonomous Driving
von: Tu, Jiahang, et al.
Veröffentlicht: (2024)
von: Tu, Jiahang, et al.
Veröffentlicht: (2024)
Unsupervised Monocular Road Segmentation for Autonomous Driving via Scene Geometry
von: Rostami, Sara Hatami, et al.
Veröffentlicht: (2025)
von: Rostami, Sara Hatami, et al.
Veröffentlicht: (2025)
Geo-EVS: Geometry-Conditioned Extrapolative View Synthesis for Autonomous Driving
von: Lan, Yatong, et al.
Veröffentlicht: (2026)
von: Lan, Yatong, et al.
Veröffentlicht: (2026)
VGGT: Visual Geometry Grounded Transformer
von: Wang, Jianyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jianyuan, et al.
Veröffentlicht: (2025)
Quantized Visual Geometry Grounded Transformer
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
Lightweight Temporal Transformer Decomposition for Federated Autonomous Driving
von: Do, Tuong, et al.
Veröffentlicht: (2025)
von: Do, Tuong, et al.
Veröffentlicht: (2025)
TinyDrive: Multiscale Visual Question Answering with Selective Token Routing for Autonomous Driving
von: Hassani, Hossein, et al.
Veröffentlicht: (2025)
von: Hassani, Hossein, et al.
Veröffentlicht: (2025)
CADReasoner: Iterative Program Editing for CAD Reverse Engineering
von: Kabisov, Soslan, et al.
Veröffentlicht: (2026)
von: Kabisov, Soslan, et al.
Veröffentlicht: (2026)
Modulate and Reconstruct: Learning Hyperspectral Imaging from Misaligned Smartphone Views
von: Reutsky, Daniil, et al.
Veröffentlicht: (2025)
von: Reutsky, Daniil, et al.
Veröffentlicht: (2025)
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
von: Yang, Zetong, et al.
Veröffentlicht: (2023)
von: Yang, Zetong, et al.
Veröffentlicht: (2023)
Visual Adversarial Attack on Vision-Language Models for Autonomous Driving
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
Multiscale Video Transformers for Class Agnostic Segmentation in Autonomous Driving
von: Cheshmi, Leila, et al.
Veröffentlicht: (2025)
von: Cheshmi, Leila, et al.
Veröffentlicht: (2025)
Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving
von: Kim, Donghyun, et al.
Veröffentlicht: (2026)
von: Kim, Donghyun, et al.
Veröffentlicht: (2026)
DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
LIX: Implicitly Infusing Spatial Geometric Prior Knowledge into Visual Semantic Segmentation for Autonomous Driving
von: Guo, Sicen, et al.
Veröffentlicht: (2024)
von: Guo, Sicen, et al.
Veröffentlicht: (2024)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
HD-VGGT: High-Resolution Visual Geometry Transformer
von: Chen, Tianrun, et al.
Veröffentlicht: (2026)
von: Chen, Tianrun, et al.
Veröffentlicht: (2026)
Hints of Prompt: Enhancing Visual Representation for Multimodal LLMs in Autonomous Driving
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
ControlLoc: Physical-World Hijacking Attack on Visual Perception in Autonomous Driving
von: Ma, Chen, et al.
Veröffentlicht: (2024)
von: Ma, Chen, et al.
Veröffentlicht: (2024)
QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Unified Sensor Simulation for Autonomous Driving
von: Patakin, Nikolay, et al.
Veröffentlicht: (2026) -
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task
von: Gabdullin, Bulat, et al.
Veröffentlicht: (2024) -
A3D: Does Diffusion Dream about 3D Alignment?
von: Ignatyev, Savva, et al.
Veröffentlicht: (2024) -
Z3D: Zero-Shot 3D Visual Grounding from Images
von: Drozdov, Nikita, et al.
Veröffentlicht: (2026) -
Cooperative Face Liveness Detection from Optical Flow
von: Sokolov, Artem, et al.
Veröffentlicht: (2025)