Towards Large-scale 3D Representation Learning with Multi-dataset Point Prompt Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Xiaoyang, Tian, Zhuotao, Wen, Xin, Peng, Bohao, Liu, Xihui, Yu, Kaicheng, Zhao, Hengshuang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GroupContrast: Semantic-aware Self-supervised Representation Learning for 3D Understanding
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations
von: Zhang, Yujia, et al.
Veröffentlicht: (2025)
von: Zhang, Yujia, et al.
Veröffentlicht: (2025)
Point Transformer V3: Simpler, Faster, Stronger
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2023)
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2023)
DreamComposer: Controllable 3D Object Generation via Multi-View Conditions
von: Yang, Yunhan, et al.
Veröffentlicht: (2023)
von: Yang, Yunhan, et al.
Veröffentlicht: (2023)
One for All: Multi-Domain Joint Training for Point Cloud Based 3D Object Detection
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Scalable Language Model with Generalized Continual Learning
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
DreamComposer++: Empowering Diffusion Models with Multi-View Conditions for 3D Content Generation
von: Yang, Yunhan, et al.
Veröffentlicht: (2025)
von: Yang, Yunhan, et al.
Veröffentlicht: (2025)
LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model
von: Yang, Senqiao, et al.
Veröffentlicht: (2023)
von: Yang, Senqiao, et al.
Veröffentlicht: (2023)
Tailor3D: Customized 3D Assets Editing and Generation with Dual-Side Images
von: Qi, Zhangyang, et al.
Veröffentlicht: (2024)
von: Qi, Zhangyang, et al.
Veröffentlicht: (2024)
Sonata: Self-Supervised Learning of Reliable Point Representations
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2025)
OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation
von: Huang, Zhening, et al.
Veröffentlicht: (2023)
von: Huang, Zhening, et al.
Veröffentlicht: (2023)
Utonia: Toward One Encoder for All Point Clouds
von: Zhang, Yujia, et al.
Veröffentlicht: (2026)
von: Zhang, Yujia, et al.
Veröffentlicht: (2026)
TMT-VIS: Taxonomy-aware Multi-dataset Joint Training for Video Instance Segmentation
von: Zheng, Rongkun, et al.
Veröffentlicht: (2023)
von: Zheng, Rongkun, et al.
Veröffentlicht: (2023)
Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception
von: Wang, Junjie, et al.
Veröffentlicht: (2025)
von: Wang, Junjie, et al.
Veröffentlicht: (2025)
LiteReality: Graphics-Ready 3D Scene Reconstruction from RGB-D Scans
von: Huang, Zhening, et al.
Veröffentlicht: (2025)
von: Huang, Zhening, et al.
Veröffentlicht: (2025)
GPT4Point: A Unified Framework for Point-Language Understanding and Generation
von: Qi, Zhangyang, et al.
Veröffentlicht: (2023)
von: Qi, Zhangyang, et al.
Veröffentlicht: (2023)
GeoAuxNet: Towards Universal 3D Representation Learning for Multi-sensor Point Clouds
von: Zhang, Shengjun, et al.
Veröffentlicht: (2024)
von: Zhang, Shengjun, et al.
Veröffentlicht: (2024)
Any3D-VLA: Enhancing VLA Robustness via Diverse Point Clouds
von: Fan, Xianzhe, et al.
Veröffentlicht: (2026)
von: Fan, Xianzhe, et al.
Veröffentlicht: (2026)
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
Mind the Interference: Retaining Pre-trained Knowledge in Parameter Efficient Continual Learning of Vision-Language Models
von: Tang, Longxiang, et al.
Veröffentlicht: (2024)
von: Tang, Longxiang, et al.
Veröffentlicht: (2024)
Multi-scale Feature Fusion with Point Pyramid for 3D Object Detection
von: Lu, Weihao, et al.
Veröffentlicht: (2024)
von: Lu, Weihao, et al.
Veröffentlicht: (2024)
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation
von: Shao, Tong, et al.
Veröffentlicht: (2024)
von: Shao, Tong, et al.
Veröffentlicht: (2024)
Point Transformer V3 Extreme: 1st Place Solution for 2024 Waymo Open Dataset Challenge in Semantic Segmentation
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2024)
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2024)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
Towards Unified 3D Object Detection via Algorithm and Data Unification
von: Li, Zhuoling, et al.
Veröffentlicht: (2024)
von: Li, Zhuoling, et al.
Veröffentlicht: (2024)
PDF: Point Diffusion Implicit Function for Large-scale Scene Neural Representation
von: Ding, Yuhan, et al.
Veröffentlicht: (2023)
von: Ding, Yuhan, et al.
Veröffentlicht: (2023)
Prompt Highlighter: Interactive Control for Multi-Modal LLMs
von: Zhang, Yuechen, et al.
Veröffentlicht: (2023)
von: Zhang, Yuechen, et al.
Veröffentlicht: (2023)
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
Multi-label Cluster Discrimination for Visual Representation Learning
von: An, Xiang, et al.
Veröffentlicht: (2024)
von: An, Xiang, et al.
Veröffentlicht: (2024)
HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation
von: Zhou, Xin, et al.
Veröffentlicht: (2026)
von: Zhou, Xin, et al.
Veröffentlicht: (2026)
OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
LION: Linear Group RNN for 3D Object Detection in Point Clouds
von: Liu, Zhe, et al.
Veröffentlicht: (2024)
von: Liu, Zhe, et al.
Veröffentlicht: (2024)
WeatherPrompt: Multi-modality Representation Learning for All-Weather Drone Visual Geo-Localization
von: Wen, Jiahao, et al.
Veröffentlicht: (2025)
von: Wen, Jiahao, et al.
Veröffentlicht: (2025)
Edit360: 2D Image Edits to 3D Assets from Any Angle
von: Huang, Junchao, et al.
Veröffentlicht: (2025)
von: Huang, Junchao, et al.
Veröffentlicht: (2025)
Towards Robust Multi-tab Website Fingerprinting
von: Deng, Xinhao, et al.
Veröffentlicht: (2025)
von: Deng, Xinhao, et al.
Veröffentlicht: (2025)
Towards Training A Chinese Large Language Model for Anesthesiology
von: Wang, Zhonghai, et al.
Veröffentlicht: (2024)
von: Wang, Zhonghai, et al.
Veröffentlicht: (2024)
MultiWorld: Scalable Multi-Agent Multi-View Video World Models
von: Wu, Haoyu, et al.
Veröffentlicht: (2026)
von: Wu, Haoyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GroupContrast: Semantic-aware Self-supervised Representation Learning for 3D Understanding
von: Wang, Chengyao, et al.
Veröffentlicht: (2024) -
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
von: Peng, Bohao, et al.
Veröffentlicht: (2024) -
Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations
von: Zhang, Yujia, et al.
Veröffentlicht: (2025) -
Point Transformer V3: Simpler, Faster, Stronger
von: Wu, Xiaoyang, et al.
Veröffentlicht: (2023) -
DreamComposer: Controllable 3D Object Generation via Multi-View Conditions
von: Yang, Yunhan, et al.
Veröffentlicht: (2023)