OmniNOCS: A unified NOCS dataset and model for 3D lifting of 2D objects
Fuente:
arXiv
Saved in:
| Main Authors: | Krishnan, Akshay, Kundu, Abhijit, Maninis, Kevis-Kokitsi, Hays, James, Brown, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffS-NOCS: 3D Point Cloud Reconstruction through Coloring Sketches to NOCS Maps Using Diffusion Models
by: Kong, Di, et al.
Published: (2025)
by: Kong, Di, et al.
Published: (2025)
DiffusionNOCS: Managing Symmetry and Uncertainty in Sim2Real Multi-Modal Category-level Pose Estimation
by: Ikeda, Takuya, et al.
Published: (2024)
by: Ikeda, Takuya, et al.
Published: (2024)
EgoCast: Forecasting Egocentric Human Pose in the Wild
by: Escobar, Maria, et al.
Published: (2024)
by: Escobar, Maria, et al.
Published: (2024)
Probing the 3D Awareness of Visual Foundation Models
by: Banani, Mohamed El, et al.
Published: (2024)
by: Banani, Mohamed El, et al.
Published: (2024)
OmniIndoor3D: Comprehensive Indoor 3D Reconstruction
by: Wei, Xiaobao, et al.
Published: (2025)
by: Wei, Xiaobao, et al.
Published: (2025)
Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection
by: Khurana, Mehar, et al.
Published: (2024)
by: Khurana, Mehar, et al.
Published: (2024)
OmniLiDAR: A Unified Diffusion Framework for Multi-Domain 3D LiDAR Generation
by: Liu, Youquan, et al.
Published: (2026)
by: Liu, Youquan, et al.
Published: (2026)
PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects
by: Cao, Ziang, et al.
Published: (2026)
by: Cao, Ziang, et al.
Published: (2026)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
by: Jeong, David C., et al.
Published: (2025)
by: Jeong, David C., et al.
Published: (2025)
IndustryShapes: An RGB-D Benchmark dataset for 6D object pose estimation of industrial assembly components and tools
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
by: Krishnan, Akshay, et al.
Published: (2025)
by: Krishnan, Akshay, et al.
Published: (2025)
Multi-modal panoramic 3D outdoor datasets for place categorization
by: Jung, Hojung, et al.
Published: (2026)
by: Jung, Hojung, et al.
Published: (2026)
OmniPose6D: Towards Short-Term Object Pose Tracking in Dynamic Scenes from Monocular RGB
by: Lin, Yunzhi, et al.
Published: (2024)
by: Lin, Yunzhi, et al.
Published: (2024)
ParisLuco3D: A high-quality target dataset for domain generalization of LiDAR perception
by: Sanchez, Jules, et al.
Published: (2023)
by: Sanchez, Jules, et al.
Published: (2023)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
by: Yang, Anqi Joyce, et al.
Published: (2026)
by: Yang, Anqi Joyce, et al.
Published: (2026)
TIPS: Text-Image Pretraining with Spatial awareness
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
Unifying Scene Representation and Hand-Eye Calibration with 3D Foundation Models
by: Zhi, Weiming, et al.
Published: (2024)
by: Zhi, Weiming, et al.
Published: (2024)
R3D: Revisiting 3D Policy Learning
by: Hong, Zhengdong, et al.
Published: (2026)
by: Hong, Zhengdong, et al.
Published: (2026)
OmniLRS: A Photorealistic Simulator for Lunar Robotics
by: Richard, Antoine, et al.
Published: (2023)
by: Richard, Antoine, et al.
Published: (2023)
Touch-GS: Visual-Tactile Supervised 3D Gaussian Splatting
by: Swann, Aiden, et al.
Published: (2024)
by: Swann, Aiden, et al.
Published: (2024)
Configurable Embodied Data Generation for Class-Agnostic RGB-D Video Segmentation
by: Opipari, Anthony, et al.
Published: (2024)
by: Opipari, Anthony, et al.
Published: (2024)
DarkGS: Learning Neural Illumination and 3D Gaussians Relighting for Robotic Exploration in the Dark
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
3D-MVP: 3D Multiview Pretraining for Robotic Manipulation
by: Qian, Shengyi, et al.
Published: (2024)
by: Qian, Shengyi, et al.
Published: (2024)
D3D-VLP: Dynamic 3D Vision-Language-Planning Model for Embodied Grounding and Navigation
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Augmenting cobots for sheet-metal SMEs with 3D object recognition and localisation
by: Cramer, Martijn, et al.
Published: (2025)
by: Cramer, Martijn, et al.
Published: (2025)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
by: Wang, Siyin, et al.
Published: (2025)
by: Wang, Siyin, et al.
Published: (2025)
Viser: Imperative, Web-based 3D Visualization in Python
by: Yi, Brent, et al.
Published: (2025)
by: Yi, Brent, et al.
Published: (2025)
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning
by: Yang, Yuncong, et al.
Published: (2024)
by: Yang, Yuncong, et al.
Published: (2024)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
by: Rim, Patrick, et al.
Published: (2026)
by: Rim, Patrick, et al.
Published: (2026)
REACT3D: Recovering Articulations for Interactive Physical 3D Scenes
by: Huang, Zhao, et al.
Published: (2025)
by: Huang, Zhao, et al.
Published: (2025)
OmniSAT: Compact Action Token, Faster Auto Regression
by: Lyu, Huaihai, et al.
Published: (2025)
by: Lyu, Huaihai, et al.
Published: (2025)
3D and 4D World Modeling: A Survey
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Aug3D: Augmenting large scale outdoor datasets for Generalizable Novel View Synthesis
by: Rauniyar, Aditya, et al.
Published: (2025)
by: Rauniyar, Aditya, et al.
Published: (2025)
Performance Evaluation of 3D Keypoint Detectors and Descriptors on Coloured Point Clouds in Subsea Environments
by: Jung, Kyungmin, et al.
Published: (2022)
by: Jung, Kyungmin, et al.
Published: (2022)
Articulate3D: Holistic Understanding of 3D Scenes as Universal Scene Description
by: Halacheva, Anna-Maria, et al.
Published: (2024)
by: Halacheva, Anna-Maria, et al.
Published: (2024)
3D Dynamics-Aware Manipulation: Endowing Manipulation Policies with 3D Foresight
by: He, Yuxin, et al.
Published: (2025)
by: He, Yuxin, et al.
Published: (2025)
Next Best Sense: Guiding Vision and Touch with FisherRF for 3D Gaussian Splatting
by: Strong, Matthew, et al.
Published: (2024)
by: Strong, Matthew, et al.
Published: (2024)
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
Analyzing the impact of semantic LoD3 building models on image-based vehicle localization
by: Bieringer, Antonia, et al.
Published: (2024)
by: Bieringer, Antonia, et al.
Published: (2024)
Similar Items
-
DiffS-NOCS: 3D Point Cloud Reconstruction through Coloring Sketches to NOCS Maps Using Diffusion Models
by: Kong, Di, et al.
Published: (2025) -
DiffusionNOCS: Managing Symmetry and Uncertainty in Sim2Real Multi-Modal Category-level Pose Estimation
by: Ikeda, Takuya, et al.
Published: (2024) -
EgoCast: Forecasting Egocentric Human Pose in the Wild
by: Escobar, Maria, et al.
Published: (2024) -
Probing the 3D Awareness of Visual Foundation Models
by: Banani, Mohamed El, et al.
Published: (2024) -
OmniIndoor3D: Comprehensive Indoor 3D Reconstruction
by: Wei, Xiaobao, et al.
Published: (2025)