POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vobecky, Antonin, Siméoni, Oriane, Hurych, David, Gidaris, Spyros, Bursuc, Andrei, Pérez, Patrick, Sivic, Josef |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
von: Vobecky, Antonin, et al.
Veröffentlicht: (2022)
von: Vobecky, Antonin, et al.
Veröffentlicht: (2022)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2024)
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2024)
DIP: Unsupervised Dense In-Context Post-training of Visual Representations
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2025)
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2025)
Test-time Contrastive Concepts for Open-world Semantic Segmentation with Vision-Language Models
von: Wysoczańska, Monika, et al.
Veröffentlicht: (2024)
von: Wysoczańska, Monika, et al.
Veröffentlicht: (2024)
Three Pillars improving Vision Foundation Model Distillation for Lidar
von: Puy, Gilles, et al.
Veröffentlicht: (2023)
von: Puy, Gilles, et al.
Veröffentlicht: (2023)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
CLIP-DINOiser: Teaching CLIP a few DINO tricks for open-vocabulary semantic segmentation
von: Wysoczańska, Monika, et al.
Veröffentlicht: (2023)
von: Wysoczańska, Monika, et al.
Veröffentlicht: (2023)
BIGFix: Bidirectional Image Generation with Token Fixing
von: Besnier, Victor, et al.
Veröffentlicht: (2025)
von: Besnier, Victor, et al.
Veröffentlicht: (2025)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
von: Simoncini, Walter, et al.
Veröffentlicht: (2024)
von: Simoncini, Walter, et al.
Veröffentlicht: (2024)
Boosting Visual Instruction Tuning with Self-Supervised Guidance
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2026)
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2026)
Valeo4Cast: A Modular Approach to End-to-End Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2026)
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2026)
Coevolving Representations in Joint Image-Feature Diffusion
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2026)
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2026)
Regularizing Self-supervised 3D Scene Flows with Surface Awareness and Cyclic Consistency
von: Vacek, Patrik, et al.
Veröffentlicht: (2023)
von: Vacek, Patrik, et al.
Veröffentlicht: (2023)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning
von: Venkataramanan, Shashanka, et al.
Veröffentlicht: (2025)
von: Venkataramanan, Shashanka, et al.
Veröffentlicht: (2025)
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2025)
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2025)
O3N: Omnidirectional Open-Vocabulary Occupancy Prediction
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
AGO: Adaptive Grounding for Open World 3D Occupancy Prediction
von: Li, Peizheng, et al.
Veröffentlicht: (2025)
von: Li, Peizheng, et al.
Veröffentlicht: (2025)
Let-It-Flow: Simultaneous Optimization of 3D Flow and Object Clustering
von: Vacek, Patrik, et al.
Veröffentlicht: (2024)
von: Vacek, Patrik, et al.
Veröffentlicht: (2024)
Boosting Generative Image Modeling via Joint Image-Feature Synthesis
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2025)
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2025)
Multi-Token Prediction Needs Registers
von: Gerontopoulos, Anastasios, et al.
Veröffentlicht: (2025)
von: Gerontopoulos, Anastasios, et al.
Veröffentlicht: (2025)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
MILAN: Milli-Annotations for Lidar Semantic Segmentation
von: Samet, Nermin, et al.
Veröffentlicht: (2024)
von: Samet, Nermin, et al.
Veröffentlicht: (2024)
Test-Time 3D Occupancy Prediction
von: Zhang, Fengyi, et al.
Veröffentlicht: (2025)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2025)
Fully Sparse 3D Occupancy Prediction
von: Liu, Haisong, et al.
Veröffentlicht: (2023)
von: Liu, Haisong, et al.
Veröffentlicht: (2023)
VEON: Vocabulary-Enhanced Occupancy Prediction
von: Zheng, Jilai, et al.
Veröffentlicht: (2024)
von: Zheng, Jilai, et al.
Veröffentlicht: (2024)
COS3D: Collaborative Open-Vocabulary 3D Segmentation
von: Zhu, Runsong, et al.
Veröffentlicht: (2025)
von: Zhu, Runsong, et al.
Veröffentlicht: (2025)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
von: Takmaz, Ayca, et al.
Veröffentlicht: (2024)
von: Takmaz, Ayca, et al.
Veröffentlicht: (2024)
3D sans 3D Scans: Scalable Pre-training from Video-Generated Point Clouds
von: Yamada, Ryousuke, et al.
Veröffentlicht: (2025)
von: Yamada, Ryousuke, et al.
Veröffentlicht: (2025)
FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction
von: Jiang, Zeyu, et al.
Veröffentlicht: (2026)
von: Jiang, Zeyu, et al.
Veröffentlicht: (2026)
COTR: Compact Occupancy TRansformer for Vision-based 3D Occupancy Prediction
von: Ma, Qihang, et al.
Veröffentlicht: (2023)
von: Ma, Qihang, et al.
Veröffentlicht: (2023)
SPOT: Self-Training with Patch-Order Permutation for Object-Centric Learning with Autoregressive Transformers
von: Kakogeorgiou, Ioannis, et al.
Veröffentlicht: (2023)
von: Kakogeorgiou, Ioannis, et al.
Veröffentlicht: (2023)
DINO-Foresight: Looking into the Future with DINO
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2024)
von: Karypidis, Efstathios, et al.
Veröffentlicht: (2024)
Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction
von: Guo, Zizhan, et al.
Veröffentlicht: (2026)
von: Guo, Zizhan, et al.
Veröffentlicht: (2026)
Open-Vocabulary 3D Semantic Segmentation with Text-to-Image Diffusion Models
von: Zhu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaoyu, et al.
Veröffentlicht: (2024)
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
Sparse Multiview Open-Vocabulary 3D Detection
von: Moliner, Olivier, et al.
Veröffentlicht: (2025)
von: Moliner, Olivier, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
von: Vobecky, Antonin, et al.
Veröffentlicht: (2022) -
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023) -
OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2024) -
DIP: Unsupervised Dense In-Context Post-training of Visual Representations
von: Sirko-Galouchenko, Sophia, et al.
Veröffentlicht: (2025) -
Test-time Contrastive Concepts for Open-world Semantic Segmentation with Vision-Language Models
von: Wysoczańska, Monika, et al.
Veröffentlicht: (2024)