P3P: Pseudo-3D Pre-training for Scaling 3D Voxel-based Masked Autoencoders
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Xuechao, Chen, Ying, Li, Jialin, Nie, Qiang, Deng, Hanqiu, Liu, Yong, Huang, Qixing, Li, Yang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
por: Gu, Pengfei, et al.
Publicado: (2024)
por: Gu, Pengfei, et al.
Publicado: (2024)
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
por: Li, Wenyu, et al.
Publicado: (2025)
por: Li, Wenyu, et al.
Publicado: (2025)
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
por: Das, Badhan Kumar, et al.
Publicado: (2025)
por: Das, Badhan Kumar, et al.
Publicado: (2025)
Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data
por: Xu, Yizhao, et al.
Publicado: (2026)
por: Xu, Yizhao, et al.
Publicado: (2026)
Multimodal Masked Autoencoder Pre-training for 3D MRI-Based Brain Tumor Analysis with Missing Modalities
por: Robinet, Lucas, et al.
Publicado: (2025)
por: Robinet, Lucas, et al.
Publicado: (2025)
SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-training
por: Lin, Nie, et al.
Publicado: (2025)
por: Lin, Nie, et al.
Publicado: (2025)
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
por: Jiang, Hanwen, et al.
Publicado: (2024)
por: Jiang, Hanwen, et al.
Publicado: (2024)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
por: Lu, Yuxuan, et al.
Publicado: (2024)
por: Lu, Yuxuan, et al.
Publicado: (2024)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
por: Xie, Yuechen, et al.
Publicado: (2025)
por: Xie, Yuechen, et al.
Publicado: (2025)
VPGS-SLAM: Voxel-based Progressive 3D Gaussian SLAM in Large-Scale Scenes
por: Deng, Tianchen, et al.
Publicado: (2025)
por: Deng, Tianchen, et al.
Publicado: (2025)
Sense Less, Generate More: Pre-training LiDAR Perception with Masked Autoencoders for Ultra-Efficient 3D Sensing
por: Tayebati, Sina, et al.
Publicado: (2024)
por: Tayebati, Sina, et al.
Publicado: (2024)
3D Feature Prediction for Masked-AutoEncoder-Based Point Cloud Pretraining
por: Yan, Siming, et al.
Publicado: (2023)
por: Yan, Siming, et al.
Publicado: (2023)
SelfMedHPM: Self Pre-training With Hard Patches Mining Masked Autoencoders For Medical Image Segmentation
por: Lv, Yunhao, et al.
Publicado: (2025)
por: Lv, Yunhao, et al.
Publicado: (2025)
Point Cloud Unsupervised Pre-training via 3D Gaussian Splatting
por: Liu, Hao, et al.
Publicado: (2024)
por: Liu, Hao, et al.
Publicado: (2024)
TutteNet: Injective 3D Deformations by Composition of 2D Mesh Deformations
por: Sun, Bo, et al.
Publicado: (2024)
por: Sun, Bo, et al.
Publicado: (2024)
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors
por: Wu, Linshan, et al.
Publicado: (2024)
por: Wu, Linshan, et al.
Publicado: (2024)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
por: Sun, Haowen, et al.
Publicado: (2026)
por: Sun, Haowen, et al.
Publicado: (2026)
P3-SAM: Native 3D Part Segmentation
por: Ma, Changfeng, et al.
Publicado: (2025)
por: Ma, Changfeng, et al.
Publicado: (2025)
$\mathsf{CSMAE~}$:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
por: Shah, Nisarg A., et al.
Publicado: (2025)
por: Shah, Nisarg A., et al.
Publicado: (2025)
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
por: Huang, Zhenpeng, et al.
Publicado: (2024)
por: Huang, Zhenpeng, et al.
Publicado: (2024)
MAPSeg: Unified Unsupervised Domain Adaptation for Heterogeneous Medical Image Segmentation Based on 3D Masked Autoencoding and Pseudo-Labeling
por: Zhang, Xuzhe, et al.
Publicado: (2023)
por: Zhang, Xuzhe, et al.
Publicado: (2023)
SUGAR: Pre-training 3D Visual Representations for Robotics
por: Chen, Shizhe, et al.
Publicado: (2024)
por: Chen, Shizhe, et al.
Publicado: (2024)
PI3D: Efficient Text-to-3D Generation with Pseudo-Image Diffusion
por: Liu, Ying-Tian, et al.
Publicado: (2023)
por: Liu, Ying-Tian, et al.
Publicado: (2023)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
por: Zou, Jian, et al.
Publicado: (2023)
por: Zou, Jian, et al.
Publicado: (2023)
Pre-training a Density-Aware Pose Transformer for Robust LiDAR-based 3D Human Pose Estimation
por: An, Xiaoqi, et al.
Publicado: (2024)
por: An, Xiaoqi, et al.
Publicado: (2024)
VPIT: Real-time Embedded Single Object 3D Tracking Using Voxel Pseudo Images
por: Oleksiienko, Illia, et al.
Publicado: (2022)
por: Oleksiienko, Illia, et al.
Publicado: (2022)
SAM-Guided Masked Token Prediction for 3D Scene Understanding
por: Chen, Zhimin, et al.
Publicado: (2024)
por: Chen, Zhimin, et al.
Publicado: (2024)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
por: Min, Chen, et al.
Publicado: (2023)
por: Min, Chen, et al.
Publicado: (2023)
Structural Teacher-Student Normality Learning for Multi-Class Anomaly Detection and Localization
por: Deng, Hanqiu, et al.
Publicado: (2024)
por: Deng, Hanqiu, et al.
Publicado: (2024)
FastLogAD: Log Anomaly Detection with Mask-Guided Pseudo Anomaly Generation and Discrimination
por: Lin, Yifei, et al.
Publicado: (2024)
por: Lin, Yifei, et al.
Publicado: (2024)
Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors
por: Dong, Xin, et al.
Publicado: (2026)
por: Dong, Xin, et al.
Publicado: (2026)
3D Scene Graph Guided Vision-Language Pre-training
por: Liu, Hao, et al.
Publicado: (2024)
por: Liu, Hao, et al.
Publicado: (2024)
CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework
por: Wu, Wentao, et al.
Publicado: (2025)
por: Wu, Wentao, et al.
Publicado: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
por: Yoo, Youngju, et al.
Publicado: (2025)
por: Yoo, Youngju, et al.
Publicado: (2025)
Robust 3D Brain MRI Inpainting with Random Masking Augmentation
por: Zhang, Juexin, et al.
Publicado: (2025)
por: Zhang, Juexin, et al.
Publicado: (2025)
Atlas Gaussians Diffusion for 3D Generation
por: Yang, Haitao, et al.
Publicado: (2024)
por: Yang, Haitao, et al.
Publicado: (2024)
Pseudo Labelling for Enhanced Masked Autoencoders
por: Nandam, Srinivasa Rao, et al.
Publicado: (2024)
por: Nandam, Srinivasa Rao, et al.
Publicado: (2024)
Structure-Adaptive Sparse Diffusion in Voxel Space for 3D Medical Image Enhancement
por: Jiang, Hongxu, et al.
Publicado: (2026)
por: Jiang, Hongxu, et al.
Publicado: (2026)
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
por: Ren, Xuanchi, et al.
Publicado: (2023)
por: Ren, Xuanchi, et al.
Publicado: (2023)
SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations
por: Yan, Xiangchao, et al.
Publicado: (2023)
por: Yan, Xiangchao, et al.
Publicado: (2023)
Ejemplares similares
-
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
por: Gu, Pengfei, et al.
Publicado: (2024) -
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
por: Li, Wenyu, et al.
Publicado: (2025) -
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
por: Das, Badhan Kumar, et al.
Publicado: (2025) -
Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data
por: Xu, Yizhao, et al.
Publicado: (2026) -
Multimodal Masked Autoencoder Pre-training for 3D MRI-Based Brain Tumor Analysis with Missing Modalities
por: Robinet, Lucas, et al.
Publicado: (2025)