Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
Fuente:
arXiv
Guardado en:
| Autores principales: | Zheng, Mi, Yang, Guanglei, Huang, Zitong, Guo, Zhenhua, Han, Kevin, Zuo, Wangmeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
por: Zou, Jian, et al.
Publicado: (2023)
por: Zou, Jian, et al.
Publicado: (2023)
MR-GDINO: Efficient Open-World Continual Object Detection
por: Dong, Bowen, et al.
Publicado: (2024)
por: Dong, Bowen, et al.
Publicado: (2024)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
por: Dong, Bowen, et al.
Publicado: (2024)
por: Dong, Bowen, et al.
Publicado: (2024)
Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
por: Li, Yuanze, et al.
Publicado: (2024)
por: Li, Yuanze, et al.
Publicado: (2024)
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
por: Dong, Bowen, et al.
Publicado: (2025)
por: Dong, Bowen, et al.
Publicado: (2025)
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
por: Liu, Chunpu, et al.
Publicado: (2024)
por: Liu, Chunpu, et al.
Publicado: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
por: Xu, Wan, et al.
Publicado: (2023)
por: Xu, Wan, et al.
Publicado: (2023)
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
por: Huang, Zitong, et al.
Publicado: (2024)
por: Huang, Zitong, et al.
Publicado: (2024)
Multi-Modality Driven LoRA for Adverse Condition Depth Estimation
por: Yang, Guanglei, et al.
Publicado: (2024)
por: Yang, Guanglei, et al.
Publicado: (2024)
S2AM3D: Scale-controllable Part Segmentation of 3D Point Clouds
por: Su, Han, et al.
Publicado: (2025)
por: Su, Han, et al.
Publicado: (2025)
FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation
por: Wang, Zehao, et al.
Publicado: (2026)
por: Wang, Zehao, et al.
Publicado: (2026)
All-in-One Video Restoration under Smoothly Evolving Unknown Weather Degradations
por: Li, Wenrui, et al.
Publicado: (2026)
por: Li, Wenrui, et al.
Publicado: (2026)
World4Drive: End-to-End Autonomous Driving via Intention-aware Physical Latent World Model
por: Zheng, Yupeng, et al.
Publicado: (2025)
por: Zheng, Yupeng, et al.
Publicado: (2025)
TwinLiteNet+: An Enhanced Multi-Task Segmentation Model for Autonomous Driving
por: Che, Quang-Huy, et al.
Publicado: (2024)
por: Che, Quang-Huy, et al.
Publicado: (2024)
MARS: An Instance-aware, Modular and Realistic Simulator for Autonomous Driving
por: Wu, Zirui, et al.
Publicado: (2023)
por: Wu, Zirui, et al.
Publicado: (2023)
DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning
por: Liu, Zhe, et al.
Publicado: (2025)
por: Liu, Zhe, et al.
Publicado: (2025)
Spatial-aware Vision Language Model for Autonomous Driving
por: Wei, Weijie, et al.
Publicado: (2025)
por: Wei, Weijie, et al.
Publicado: (2025)
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention
por: Lu, Hannan, et al.
Publicado: (2024)
por: Lu, Hannan, et al.
Publicado: (2024)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
por: Chu, Jisheng, et al.
Publicado: (2026)
por: Chu, Jisheng, et al.
Publicado: (2026)
Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models
por: Huang, Zitong, et al.
Publicado: (2026)
por: Huang, Zitong, et al.
Publicado: (2026)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
por: Qiao, Weidong, et al.
Publicado: (2026)
por: Qiao, Weidong, et al.
Publicado: (2026)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
por: Yang, Feng, et al.
Publicado: (2025)
por: Yang, Feng, et al.
Publicado: (2025)
Dual-Camera Smooth Zoom on Mobile Phones
por: Wu, Renlong, et al.
Publicado: (2024)
por: Wu, Renlong, et al.
Publicado: (2024)
CGL: Advancing Continual GUI Learning via Reinforcement Fine-Tuning
por: Yao, Zhenquan, et al.
Publicado: (2026)
por: Yao, Zhenquan, et al.
Publicado: (2026)
Topology-aware Mamba for Crack Segmentation in Structures
por: Zuo, Xin, et al.
Publicado: (2024)
por: Zuo, Xin, et al.
Publicado: (2024)
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
por: Diao, Muxi, et al.
Publicado: (2025)
por: Diao, Muxi, et al.
Publicado: (2025)
InVDriver: Intra-Instance Aware Vectorized Query-Based Autonomous Driving Transformer
por: Zhang, Bo, et al.
Publicado: (2025)
por: Zhang, Bo, et al.
Publicado: (2025)
Post-interactive Multimodal Trajectory Prediction for Autonomous Driving
por: Huang, Ziyi, et al.
Publicado: (2025)
por: Huang, Ziyi, et al.
Publicado: (2025)
Segmenting and Understanding: Region-aware Semantic Attention for Fine-grained Image Quality Assessment with Large Language Models
por: Song, Chenyue, et al.
Publicado: (2025)
por: Song, Chenyue, et al.
Publicado: (2025)
Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving
por: Yang, Lijin, et al.
Publicado: (2026)
por: Yang, Lijin, et al.
Publicado: (2026)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
por: Li, Xiaoming, et al.
Publicado: (2025)
por: Li, Xiaoming, et al.
Publicado: (2025)
DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
por: Xu, Zhenhua, et al.
Publicado: (2023)
por: Xu, Zhenhua, et al.
Publicado: (2023)
DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving
por: Su, Haisheng, et al.
Publicado: (2026)
por: Su, Haisheng, et al.
Publicado: (2026)
Domain-Incremental Semantic Segmentation for Autonomous Driving under Adverse Driving Conditions
por: Muralidhara, Shishir, et al.
Publicado: (2025)
por: Muralidhara, Shishir, et al.
Publicado: (2025)
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
por: Li, Wenrui, et al.
Publicado: (2024)
por: Li, Wenrui, et al.
Publicado: (2024)
An Overview about Emerging Technologies of Autonomous Driving
por: Huang, Yu, et al.
Publicado: (2023)
por: Huang, Yu, et al.
Publicado: (2023)
FlowAD: Ego-Scene Interactive Modeling for Autonomous Driving
por: Guo, Mingzhe, et al.
Publicado: (2026)
por: Guo, Mingzhe, et al.
Publicado: (2026)
GEMINUS: Dual-aware Global and Scene-Adaptive Mixture-of-Experts for End-to-End Autonomous Driving
por: Wan, Chi, et al.
Publicado: (2025)
por: Wan, Chi, et al.
Publicado: (2025)
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation
por: Dong, Zhe, et al.
Publicado: (2024)
por: Dong, Zhe, et al.
Publicado: (2024)
GenAD: Generative End-to-End Autonomous Driving
por: Zheng, Wenzhao, et al.
Publicado: (2024)
por: Zheng, Wenzhao, et al.
Publicado: (2024)
Ejemplares similares
-
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
por: Zou, Jian, et al.
Publicado: (2023) -
MR-GDINO: Efficient Open-World Continual Object Detection
por: Dong, Bowen, et al.
Publicado: (2024) -
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
por: Dong, Bowen, et al.
Publicado: (2024) -
Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
por: Li, Yuanze, et al.
Publicado: (2024) -
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
por: Dong, Bowen, et al.
Publicado: (2025)