UniPLV: Towards Label-Efficient Open-World 3D Scene Understanding by Regional Visual Language Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yuru, Liu, Pei, Wang, Songtao, Zhang, Zehan, Lu, Xinyan, Cai, Changwei, Li, Hao, Liu, Fu, Jia, Peng, Lang, Xianpeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes
von: Ni, Shuo, et al.
Veröffentlicht: (2025)
von: Ni, Shuo, et al.
Veröffentlicht: (2025)
Alpha PLV Professional Report
von: Shaub, Jarid Shaub
Veröffentlicht: (2026)
von: Shaub, Jarid Shaub
Veröffentlicht: (2026)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
von: Tang, Tao, et al.
Veröffentlicht: (2024)
von: Tang, Tao, et al.
Veröffentlicht: (2024)
CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding
von: Yang, Jihan, et al.
Veröffentlicht: (2023)
von: Yang, Jihan, et al.
Veröffentlicht: (2023)
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
TokenFLEX: Unified VLM Training for Flexible Visual Tokens Inference
von: Hu, Junshan, et al.
Veröffentlicht: (2025)
von: Hu, Junshan, et al.
Veröffentlicht: (2025)
UniM-OV3D: Uni-Modality Open-Vocabulary 3D Scene Understanding with Fine-Grained Feature Representation
von: He, Qingdong, et al.
Veröffentlicht: (2024)
von: He, Qingdong, et al.
Veröffentlicht: (2024)
Open-Vocabulary vs Supervised Learning Methods for Post-Disaster Visual Scene Understanding
von: Michailidou, Anna, et al.
Veröffentlicht: (2026)
von: Michailidou, Anna, et al.
Veröffentlicht: (2026)
RenderWorld: World Model with Self-Supervised 3D Label
von: Yan, Ziyang, et al.
Veröffentlicht: (2024)
von: Yan, Ziyang, et al.
Veröffentlicht: (2024)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Driving
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
Foundation Model-Driven Grasping of Unknown Objects via Center of Gravity Estimation
von: Xiangli, Kang, et al.
Veröffentlicht: (2025)
von: Xiangli, Kang, et al.
Veröffentlicht: (2025)
Foundation Model‐Driven Grasping of Unknown Objects via Center of Gravity Estimation
von: Kang Xiangli, et al.
Veröffentlicht: (2026)
von: Kang Xiangli, et al.
Veröffentlicht: (2026)
YOLO-UniOW: Efficient Universal Open-World Object Detection
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition
von: Chen, Shunpeng, et al.
Veröffentlicht: (2026)
von: Chen, Shunpeng, et al.
Veröffentlicht: (2026)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
OwMatch: Conditional Self-Labeling with Consistency for Open-World Semi-Supervised Learning
von: Niu, Shengjie, et al.
Veröffentlicht: (2024)
von: Niu, Shengjie, et al.
Veröffentlicht: (2024)
HiNeuS: High-fidelity Neural Surface Mitigating Low-texture and Reflective Ambiguity
von: Wang, Yida, et al.
Veröffentlicht: (2025)
von: Wang, Yida, et al.
Veröffentlicht: (2025)
OpenSU3D: Open World 3D Scene Understanding using Foundation Models
von: Mohiuddin, Rafay, et al.
Veröffentlicht: (2024)
von: Mohiuddin, Rafay, et al.
Veröffentlicht: (2024)
Towards 3D Objectness Learning in an Open World
von: Liu, Taichi, et al.
Veröffentlicht: (2025)
von: Liu, Taichi, et al.
Veröffentlicht: (2025)
Language-Driven Object-Oriented Two-Stage Method for Scene Graph Anticipation
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
UniVG: Towards UNIfied-modal Video Generation
von: Ruan, Ludan, et al.
Veröffentlicht: (2024)
von: Ruan, Ludan, et al.
Veröffentlicht: (2024)
Uni-Animator: Towards Unified Visual Colorization
von: Chen, Xinyuan, et al.
Veröffentlicht: (2026)
von: Chen, Xinyuan, et al.
Veröffentlicht: (2026)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
Towards Automatic Soccer Commentary Generation with Knowledge-Enhanced Visual Reasoning
von: Jin, Zeyu, et al.
Veröffentlicht: (2026)
von: Jin, Zeyu, et al.
Veröffentlicht: (2026)
UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
von: Lin, Bin, et al.
Veröffentlicht: (2025)
von: Lin, Bin, et al.
Veröffentlicht: (2025)
Masked Point-Entity Contrast for Open-Vocabulary 3D Scene Understanding
von: Wang, Yan, et al.
Veröffentlicht: (2025)
von: Wang, Yan, et al.
Veröffentlicht: (2025)
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
Exploiting Minority Pseudo-Labels for Semi-Supervised Fine-grained Road Scene Understanding
von: Hong, Yuting, et al.
Veröffentlicht: (2024)
von: Hong, Yuting, et al.
Veröffentlicht: (2024)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
GaussianGraph: 3D Gaussian-based Scene Graph Generation for Open-world Scene Understanding
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
von: Fu, Rao, et al.
Veröffentlicht: (2024)
von: Fu, Rao, et al.
Veröffentlicht: (2024)
UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation
von: Yue, Zhengrong, et al.
Veröffentlicht: (2025)
von: Yue, Zhengrong, et al.
Veröffentlicht: (2025)
SNOW: Spatio-Temporal Scene Understanding with World Knowledge for Open-World Embodied Reasoning
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
RoadSceneBench: A Lightweight Benchmark for Mid-Level Road Scene Understanding
von: Liu, Xiyan, et al.
Veröffentlicht: (2025)
von: Liu, Xiyan, et al.
Veröffentlicht: (2025)
Open-World Semi-Supervised Learning for Node Classification
von: Wang, Yanling, et al.
Veröffentlicht: (2024)
von: Wang, Yanling, et al.
Veröffentlicht: (2024)
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
von: Yan, Yunzhi, et al.
Veröffentlicht: (2024)
von: Yan, Yunzhi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes
von: Ni, Shuo, et al.
Veröffentlicht: (2025) -
Alpha PLV Professional Report
von: Shaub, Jarid Shaub
Veröffentlicht: (2026) -
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
von: Tang, Tao, et al.
Veröffentlicht: (2024) -
CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving
von: Liu, Pei, et al.
Veröffentlicht: (2025) -
RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding
von: Yang, Jihan, et al.
Veröffentlicht: (2023)