Translating Images to Road Network: A Sequence-to-Sequence Perspective
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Jiachen, Nie, Ming, Zhang, Bozhou, Peng, Reyuan, Cai, Xinyue, Xu, Hang, Wen, Feng, Zhang, Wei, Zhang, Li |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LaneGraph2Seq: Lane Topology Extraction with Language Model via Vertex-Edge Encoding and Connectivity Enhancement
por: Peng, Renyuan, et al.
Publicado: (2024)
por: Peng, Renyuan, et al.
Publicado: (2024)
LaneCorrect: Self-supervised Lane Detection
por: Nie, Ming, et al.
Publicado: (2024)
por: Nie, Ming, et al.
Publicado: (2024)
Deep Non-rigid Structure-from-Motion: A Sequence-to-Sequence Translation Perspective
por: Deng, Hui, et al.
Publicado: (2022)
por: Deng, Hui, et al.
Publicado: (2022)
Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving
por: Nie, Ming, et al.
Publicado: (2023)
por: Nie, Ming, et al.
Publicado: (2023)
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
por: Yuan, Runtian, et al.
Publicado: (2025)
por: Yuan, Runtian, et al.
Publicado: (2025)
ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models
por: Li, Bozhou, et al.
Publicado: (2025)
por: Li, Bozhou, et al.
Publicado: (2025)
DeMo: Decoupling Motion Forecasting into Directional Intentions and Dynamic States
por: Zhang, Bozhou, et al.
Publicado: (2024)
por: Zhang, Bozhou, et al.
Publicado: (2024)
KFFocus: Highlighting Keyframes for Enhanced Video Understanding
por: Nie, Ming, et al.
Publicado: (2025)
por: Nie, Ming, et al.
Publicado: (2025)
USF-Net: A Unified Spatiotemporal Fusion Network for Ground-Based Remote Sensing Cloud Image Sequence Extrapolation
por: Niu, Penghui, et al.
Publicado: (2025)
por: Niu, Penghui, et al.
Publicado: (2025)
Insertion Network for Image Sequence Correspondence
por: Su, Dingjie, et al.
Publicado: (2026)
por: Su, Dingjie, et al.
Publicado: (2026)
Perception in Plan: Coupled Perception and Planning for End-to-End Autonomous Driving
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
ColorFlow: Retrieval-Augmented Image Sequence Colorization
por: Zhuang, Junhao, et al.
Publicado: (2024)
por: Zhuang, Junhao, et al.
Publicado: (2024)
Dual-frequency Selected Knowledge Distillation with Statistical-based Sample Rectification for PolSAR Image Classification
por: Xin, Xinyue, et al.
Publicado: (2025)
por: Xin, Xinyue, et al.
Publicado: (2025)
DeMo++: Motion Decoupling for Autonomous Driving
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
Motion Forecasting in Continuous Driving
por: Song, Nan, et al.
Publicado: (2024)
por: Song, Nan, et al.
Publicado: (2024)
Unified Sequence-to-Sequence Learning for Single- and Multi-Modal Visual Object Tracking
por: Chen, Xin, et al.
Publicado: (2023)
por: Chen, Xin, et al.
Publicado: (2023)
Uncertainty-Aware Normal-Guided Gaussian Splatting for Surface Reconstruction from Sparse Image Sequences
por: Tan, Zhen, et al.
Publicado: (2025)
por: Tan, Zhen, et al.
Publicado: (2025)
MedSG-Bench: A Benchmark for Medical Image Sequences Grounding
por: Yue, Jingkun, et al.
Publicado: (2025)
por: Yue, Jingkun, et al.
Publicado: (2025)
PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking
por: Zou, Quanchen, et al.
Publicado: (2025)
por: Zou, Quanchen, et al.
Publicado: (2025)
Towards Unified Multimodal Interleaved Generation via Group Relative Policy Optimization
por: Nie, Ming, et al.
Publicado: (2026)
por: Nie, Ming, et al.
Publicado: (2026)
Relative Position Matters: Trajectory Prediction and Planning with Polar Representation
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
Learning Sequence Descriptor based on Spatio-Temporal Attention for Visual Place Recognition
por: Zhao, Junqiao, et al.
Publicado: (2023)
por: Zhao, Junqiao, et al.
Publicado: (2023)
Sequence Length Scaling in Vision Transformers for Scientific Images on Frontier
por: Tsaris, Aristeidis, et al.
Publicado: (2024)
por: Tsaris, Aristeidis, et al.
Publicado: (2024)
Autoregressive Sequence Modeling for 3D Medical Image Representation
por: Wang, Siwen, et al.
Publicado: (2024)
por: Wang, Siwen, et al.
Publicado: (2024)
Unsupervised Online 3D Instance Segmentation with Synthetic Sequences and Dynamic Loss
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving
por: Song, Nan, et al.
Publicado: (2025)
por: Song, Nan, et al.
Publicado: (2025)
Personalized Image Descriptions from Attention Sequences
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
Mode-as-Sequence: Translating Multimodal Motion Prediction into Unified Sequential Mode Modeling
por: Zhou, Zikang, et al.
Publicado: (2026)
por: Zhou, Zikang, et al.
Publicado: (2026)
Bridging the Gap between Text, Audio, Image, and Any Sequence: A Novel Approach using Gloss-based Annotation
por: Fang, Sen, et al.
Publicado: (2024)
por: Fang, Sen, et al.
Publicado: (2024)
OptiCorNet: Optimizing Sequence-Based Context Correlation for Visual Place Recognition
por: Li, Zhenyu, et al.
Publicado: (2025)
por: Li, Zhenyu, et al.
Publicado: (2025)
ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving
por: Li, Jingyu, et al.
Publicado: (2025)
por: Li, Jingyu, et al.
Publicado: (2025)
Future-Aware End-to-End Driving: Bidirectional Modeling of Trajectory Planning and Scene Evolution
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
LiSTAR: Ray-Centric World Models for 4D LiDAR Sequences in Autonomous Driving
por: Liu, Pei, et al.
Publicado: (2025)
por: Liu, Pei, et al.
Publicado: (2025)
Skywork UniPic 3.0: Unified Multi-Image Composition via Sequence Modeling
por: Wei, Hongyang, et al.
Publicado: (2026)
por: Wei, Hongyang, et al.
Publicado: (2026)
MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequence
por: Zhao, Canyu, et al.
Publicado: (2024)
por: Zhao, Canyu, et al.
Publicado: (2024)
KeyGS: A Keyframe-Centric Gaussian Splatting Method for Monocular Image Sequences
por: Chang, Keng-Wei, et al.
Publicado: (2024)
por: Chang, Keng-Wei, et al.
Publicado: (2024)
Asynchronous Multimodal Video Sequence Fusion via Learning Modality-Exclusive and -Agnostic Representations
por: Yang, Dingkang, et al.
Publicado: (2024)
por: Yang, Dingkang, et al.
Publicado: (2024)
Drawing2CAD: Sequence-to-Sequence Learning for CAD Generation from Vector Drawings
por: Qin, Feiwei, et al.
Publicado: (2025)
por: Qin, Feiwei, et al.
Publicado: (2025)
See Tomorrow, Act Today: Foresight-Driven Autonomous Driving
por: Zhang, Bozhou, et al.
Publicado: (2026)
por: Zhang, Bozhou, et al.
Publicado: (2026)
Ejemplares similares
-
LaneGraph2Seq: Lane Topology Extraction with Language Model via Vertex-Edge Encoding and Connectivity Enhancement
por: Peng, Renyuan, et al.
Publicado: (2024) -
LaneCorrect: Self-supervised Lane Detection
por: Nie, Ming, et al.
Publicado: (2024) -
Deep Non-rigid Structure-from-Motion: A Sequence-to-Sequence Translation Perspective
por: Deng, Hui, et al.
Publicado: (2022) -
Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving
por: Nie, Ming, et al.
Publicado: (2023) -
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
por: Yuan, Runtian, et al.
Publicado: (2025)