3D sans 3D Scans: Scalable Pre-training from Video-Generated Point Clouds
Fuente:
arXiv
Saved in:
| Main Authors: | Yamada, Ryousuke, Ide, Kohsuke, Fukuhara, Yoshihiro, Kataoka, Hirokatsu, Puy, Gilles, Bursuc, Andrei, Asano, Yuki M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Industrial Synthetic Segment Pre-training
by: Mae, Shinichi, et al.
Published: (2025)
by: Mae, Shinichi, et al.
Published: (2025)
Pre-training with 3D Synthetic Data: Learning 3D Point Cloud Instance Segmentation from 3D Synthetic Scenes
by: Otsuka, Daichi, et al.
Published: (2025)
by: Otsuka, Daichi, et al.
Published: (2025)
Primitive Geometry Segment Pre-training for 3D Medical Image Segmentation
by: Tadokoro, Ryu, et al.
Published: (2024)
by: Tadokoro, Ryu, et al.
Published: (2024)
MoireMix: A Formula-Based Data Augmentation for Improving Image Classification Robustness
by: Matsuo, Yuto, et al.
Published: (2026)
by: Matsuo, Yuto, et al.
Published: (2026)
Text-guided Synthetic Geometric Augmentation for Zero-shot 3D Understanding
by: Torimi, Kohei, et al.
Published: (2025)
by: Torimi, Kohei, et al.
Published: (2025)
Scaling Backwards: Minimal Synthetic Pre-training?
by: Nakamura, Ryo, et al.
Published: (2024)
by: Nakamura, Ryo, et al.
Published: (2024)
Simple Visual Artifact Detection in Sora-Generated Videos
by: Sugiyama, Misora, et al.
Published: (2025)
by: Sugiyama, Misora, et al.
Published: (2025)
Formula-Supervised Visual-Geometric Pre-training
by: Yamada, Ryosuke, et al.
Published: (2024)
by: Yamada, Ryosuke, et al.
Published: (2024)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
by: Simoncini, Walter, et al.
Published: (2024)
by: Simoncini, Walter, et al.
Published: (2024)
Robust Fine-tuning for Pre-trained 3D Point Cloud Models
by: Zhang, Zhibo, et al.
Published: (2024)
by: Zhang, Zhibo, et al.
Published: (2024)
Point Cloud Unsupervised Pre-training via 3D Gaussian Splatting
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
by: Kupyn, Orest, et al.
Published: (2025)
by: Kupyn, Orest, et al.
Published: (2025)
Pre-training Vision Transformers with Formula-driven Supervised Learning
by: Kataoka, Hirokatsu, et al.
Published: (2022)
by: Kataoka, Hirokatsu, et al.
Published: (2022)
UniPre3D: Unified Pre-training of 3D Point Cloud Models with Cross-Modal Gaussian Splatting
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Indoor Space Authentication by ISS-based Keypoint Extraction from 3D Point Clouds
by: Yamada, Yuki, et al.
Published: (2026)
by: Yamada, Yuki, et al.
Published: (2026)
Performance monitoring
by: Andrei, Bursuc
Published: (2025)
by: Andrei, Bursuc
Published: (2025)
Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning
by: Venkataramanan, Shashanka, et al.
Published: (2025)
by: Venkataramanan, Shashanka, et al.
Published: (2025)
FDIF: Formula-Driven supervised Learning with Implicit Functions for 3D Medical Image Segmentation
by: Yamamoto, Yukinori, et al.
Published: (2026)
by: Yamamoto, Yukinori, et al.
Published: (2026)
Three Pillars improving Vision Foundation Model Distillation for Lidar
by: Puy, Gilles, et al.
Published: (2023)
by: Puy, Gilles, et al.
Published: (2023)
RESSCAL3D: Resolution Scalable 3D Semantic Segmentation of Point Clouds
by: Royen, Remco, et al.
Published: (2024)
by: Royen, Remco, et al.
Published: (2024)
Leveraging LLMs with Iterative Loop Structure for Enhanced Social Intelligence in Video Question Answering
by: Mori, Erika, et al.
Published: (2025)
by: Mori, Erika, et al.
Published: (2025)
PowerCLIP: Powerset Alignment for Contrastive Pre-Training
by: Kawamura, Masaki, et al.
Published: (2025)
by: Kawamura, Masaki, et al.
Published: (2025)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
by: Puy, Gilles, et al.
Published: (2026)
by: Puy, Gilles, et al.
Published: (2026)
POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images
by: Vobecky, Antonin, et al.
Published: (2024)
by: Vobecky, Antonin, et al.
Published: (2024)
ScanDP: Generalizable 3D Scanning with Diffusion Policy
by: Hirako, Itsuki, et al.
Published: (2026)
by: Hirako, Itsuki, et al.
Published: (2026)
Dynamic 3D Point Cloud Sequences as 2D Videos
by: Zeng, Yiming, et al.
Published: (2024)
by: Zeng, Yiming, et al.
Published: (2024)
Rethinking Image Super-Resolution from Training Data Perspectives
by: Ohtani, Go, et al.
Published: (2024)
by: Ohtani, Go, et al.
Published: (2024)
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
by: Xia, Jiatong, et al.
Published: (2026)
by: Xia, Jiatong, et al.
Published: (2026)
Human Action Recognition without Human
by: Kataoka, Hirokatsu, et al.
Published: (2016)
by: Kataoka, Hirokatsu, et al.
Published: (2016)
Watermark-embedded Adversarial Examples for Copyright Protection against Diffusion Models
by: Zhu, Peifei, et al.
Published: (2024)
by: Zhu, Peifei, et al.
Published: (2024)
Graph-based Scalable Sampling of 3D Point Cloud Attributes
by: Sridhara, Shashank N., et al.
Published: (2024)
by: Sridhara, Shashank N., et al.
Published: (2024)
3D-BBS: Global Localization for 3D Point Cloud Scan Matching Using Branch-and-Bound Algorithm
by: Aoki, Koki, et al.
Published: (2023)
by: Aoki, Koki, et al.
Published: (2023)
Towards Scalable Language-Image Pre-training for 3D Medical Imaging
by: Zhao, Chenhui, et al.
Published: (2025)
by: Zhao, Chenhui, et al.
Published: (2025)
ULIP-2: Towards Scalable Multimodal Pre-training for 3D Understanding
by: Xue, Le, et al.
Published: (2023)
by: Xue, Le, et al.
Published: (2023)
SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations
by: Yan, Xiangchao, et al.
Published: (2023)
by: Yan, Xiangchao, et al.
Published: (2023)
3D View Optimization for Improving Image Aesthetics
by: Uchida, Taichi, et al.
Published: (2024)
by: Uchida, Taichi, et al.
Published: (2024)
HanDyVQA: A Video QA Benchmark for Fine-Grained Hand-Object Interaction Dynamics
by: Tateno, Masatoshi, et al.
Published: (2025)
by: Tateno, Masatoshi, et al.
Published: (2025)
Formula-Supervised Sound Event Detection: Pre-Training Without Real Data
by: Shibata, Yuto, et al.
Published: (2025)
by: Shibata, Yuto, et al.
Published: (2025)
IGLOSS: Image Generation for Lidar Open-vocabulary Semantic Segmentation
by: Samet, Nermin, et al.
Published: (2026)
by: Samet, Nermin, et al.
Published: (2026)
Improving Multimodal Distillation for 3D Semantic Segmentation under Domain Shift
by: Michele, Björn, et al.
Published: (2025)
by: Michele, Björn, et al.
Published: (2025)
Similar Items
-
Industrial Synthetic Segment Pre-training
by: Mae, Shinichi, et al.
Published: (2025) -
Pre-training with 3D Synthetic Data: Learning 3D Point Cloud Instance Segmentation from 3D Synthetic Scenes
by: Otsuka, Daichi, et al.
Published: (2025) -
Primitive Geometry Segment Pre-training for 3D Medical Image Segmentation
by: Tadokoro, Ryu, et al.
Published: (2024) -
MoireMix: A Formula-Based Data Augmentation for Improving Image Classification Robustness
by: Matsuo, Yuto, et al.
Published: (2026) -
Text-guided Synthetic Geometric Augmentation for Zero-shot 3D Understanding
by: Torimi, Kohei, et al.
Published: (2025)