P3P: Pseudo-3D Pre-training for Scaling 3D Voxel-based Masked Autoencoders
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xuechao, Chen, Ying, Li, Jialin, Nie, Qiang, Deng, Hanqiu, Liu, Yong, Huang, Qixing, Li, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
von: Gu, Pengfei, et al.
Veröffentlicht: (2024)
von: Gu, Pengfei, et al.
Veröffentlicht: (2024)
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
von: Li, Wenyu, et al.
Veröffentlicht: (2025)
von: Li, Wenyu, et al.
Veröffentlicht: (2025)
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
von: Das, Badhan Kumar, et al.
Veröffentlicht: (2025)
von: Das, Badhan Kumar, et al.
Veröffentlicht: (2025)
Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data
von: Xu, Yizhao, et al.
Veröffentlicht: (2026)
von: Xu, Yizhao, et al.
Veröffentlicht: (2026)
Multimodal Masked Autoencoder Pre-training for 3D MRI-Based Brain Tumor Analysis with Missing Modalities
von: Robinet, Lucas, et al.
Veröffentlicht: (2025)
von: Robinet, Lucas, et al.
Veröffentlicht: (2025)
SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-training
von: Lin, Nie, et al.
Veröffentlicht: (2025)
von: Lin, Nie, et al.
Veröffentlicht: (2025)
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
von: Jiang, Hanwen, et al.
Veröffentlicht: (2024)
von: Jiang, Hanwen, et al.
Veröffentlicht: (2024)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
von: Lu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Lu, Yuxuan, et al.
Veröffentlicht: (2024)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
VPGS-SLAM: Voxel-based Progressive 3D Gaussian SLAM in Large-Scale Scenes
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
Sense Less, Generate More: Pre-training LiDAR Perception with Masked Autoencoders for Ultra-Efficient 3D Sensing
von: Tayebati, Sina, et al.
Veröffentlicht: (2024)
von: Tayebati, Sina, et al.
Veröffentlicht: (2024)
3D Feature Prediction for Masked-AutoEncoder-Based Point Cloud Pretraining
von: Yan, Siming, et al.
Veröffentlicht: (2023)
von: Yan, Siming, et al.
Veröffentlicht: (2023)
SelfMedHPM: Self Pre-training With Hard Patches Mining Masked Autoencoders For Medical Image Segmentation
von: Lv, Yunhao, et al.
Veröffentlicht: (2025)
von: Lv, Yunhao, et al.
Veröffentlicht: (2025)
Point Cloud Unsupervised Pre-training via 3D Gaussian Splatting
von: Liu, Hao, et al.
Veröffentlicht: (2024)
von: Liu, Hao, et al.
Veröffentlicht: (2024)
TutteNet: Injective 3D Deformations by Composition of 2D Mesh Deformations
von: Sun, Bo, et al.
Veröffentlicht: (2024)
von: Sun, Bo, et al.
Veröffentlicht: (2024)
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors
von: Wu, Linshan, et al.
Veröffentlicht: (2024)
von: Wu, Linshan, et al.
Veröffentlicht: (2024)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
von: Sun, Haowen, et al.
Veröffentlicht: (2026)
von: Sun, Haowen, et al.
Veröffentlicht: (2026)
P3-SAM: Native 3D Part Segmentation
von: Ma, Changfeng, et al.
Veröffentlicht: (2025)
von: Ma, Changfeng, et al.
Veröffentlicht: (2025)
$\mathsf{CSMAE~}$:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
MAPSeg: Unified Unsupervised Domain Adaptation for Heterogeneous Medical Image Segmentation Based on 3D Masked Autoencoding and Pseudo-Labeling
von: Zhang, Xuzhe, et al.
Veröffentlicht: (2023)
von: Zhang, Xuzhe, et al.
Veröffentlicht: (2023)
SUGAR: Pre-training 3D Visual Representations for Robotics
von: Chen, Shizhe, et al.
Veröffentlicht: (2024)
von: Chen, Shizhe, et al.
Veröffentlicht: (2024)
PI3D: Efficient Text-to-3D Generation with Pseudo-Image Diffusion
von: Liu, Ying-Tian, et al.
Veröffentlicht: (2023)
von: Liu, Ying-Tian, et al.
Veröffentlicht: (2023)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
von: Zou, Jian, et al.
Veröffentlicht: (2023)
von: Zou, Jian, et al.
Veröffentlicht: (2023)
Pre-training a Density-Aware Pose Transformer for Robust LiDAR-based 3D Human Pose Estimation
von: An, Xiaoqi, et al.
Veröffentlicht: (2024)
von: An, Xiaoqi, et al.
Veröffentlicht: (2024)
VPIT: Real-time Embedded Single Object 3D Tracking Using Voxel Pseudo Images
von: Oleksiienko, Illia, et al.
Veröffentlicht: (2022)
von: Oleksiienko, Illia, et al.
Veröffentlicht: (2022)
SAM-Guided Masked Token Prediction for 3D Scene Understanding
von: Chen, Zhimin, et al.
Veröffentlicht: (2024)
von: Chen, Zhimin, et al.
Veröffentlicht: (2024)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2023)
von: Min, Chen, et al.
Veröffentlicht: (2023)
Structural Teacher-Student Normality Learning for Multi-Class Anomaly Detection and Localization
von: Deng, Hanqiu, et al.
Veröffentlicht: (2024)
von: Deng, Hanqiu, et al.
Veröffentlicht: (2024)
FastLogAD: Log Anomaly Detection with Mask-Guided Pseudo Anomaly Generation and Discrimination
von: Lin, Yifei, et al.
Veröffentlicht: (2024)
von: Lin, Yifei, et al.
Veröffentlicht: (2024)
Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors
von: Dong, Xin, et al.
Veröffentlicht: (2026)
von: Dong, Xin, et al.
Veröffentlicht: (2026)
3D Scene Graph Guided Vision-Language Pre-training
von: Liu, Hao, et al.
Veröffentlicht: (2024)
von: Liu, Hao, et al.
Veröffentlicht: (2024)
CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
von: Yoo, Youngju, et al.
Veröffentlicht: (2025)
von: Yoo, Youngju, et al.
Veröffentlicht: (2025)
Robust 3D Brain MRI Inpainting with Random Masking Augmentation
von: Zhang, Juexin, et al.
Veröffentlicht: (2025)
von: Zhang, Juexin, et al.
Veröffentlicht: (2025)
Atlas Gaussians Diffusion for 3D Generation
von: Yang, Haitao, et al.
Veröffentlicht: (2024)
von: Yang, Haitao, et al.
Veröffentlicht: (2024)
Pseudo Labelling for Enhanced Masked Autoencoders
von: Nandam, Srinivasa Rao, et al.
Veröffentlicht: (2024)
von: Nandam, Srinivasa Rao, et al.
Veröffentlicht: (2024)
Structure-Adaptive Sparse Diffusion in Voxel Space for 3D Medical Image Enhancement
von: Jiang, Hongxu, et al.
Veröffentlicht: (2026)
von: Jiang, Hongxu, et al.
Veröffentlicht: (2026)
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
von: Ren, Xuanchi, et al.
Veröffentlicht: (2023)
von: Ren, Xuanchi, et al.
Veröffentlicht: (2023)
SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
von: Gu, Pengfei, et al.
Veröffentlicht: (2024) -
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
von: Li, Wenyu, et al.
Veröffentlicht: (2025) -
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
von: Das, Badhan Kumar, et al.
Veröffentlicht: (2025) -
Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data
von: Xu, Yizhao, et al.
Veröffentlicht: (2026) -
Multimodal Masked Autoencoder Pre-training for 3D MRI-Based Brain Tumor Analysis with Missing Modalities
von: Robinet, Lucas, et al.
Veröffentlicht: (2025)