Generating Images with 3D Annotations Using Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Wufei, Liu, Qihao, Wang, Jiahao, Wang, Angtian, Yuan, Xiaoding, Zhang, Yi, Xiao, Zihao, Zhang, Guofeng, Lu, Beijia, Duan, Ruxiao, Qi, Yongrui, Kortylewski, Adam, Liu, Yaoyao, Yuille, Alan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NOVUM: Neural Object Volumes for Robust Object Classification
by: Jesslen, Artur, et al.
Published: (2023)
by: Jesslen, Artur, et al.
Published: (2023)
Prompt-Based Exemplar Super-Compression and Regeneration for Class-Incremental Learning
by: Duan, Ruxiao, et al.
Published: (2023)
by: Duan, Ruxiao, et al.
Published: (2023)
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
Learning a Category-level Object Pose Estimator without Pose Annotations
by: Tian, Fengrui, et al.
Published: (2024)
by: Tian, Fengrui, et al.
Published: (2024)
Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
PASR: Pose-Aware 3D Shape Retrieval from Occluded Single Views
by: Shi, Jiaxin, et al.
Published: (2026)
by: Shi, Jiaxin, et al.
Published: (2026)
DINeMo: Learning Neural Mesh Models with no 3D Annotations
by: Guo, Weijie, et al.
Published: (2025)
by: Guo, Weijie, et al.
Published: (2025)
VoGE: A Differentiable Volume Renderer using Gaussian Ellipsoids for Analysis-by-Synthesis
by: Wang, Angtian, et al.
Published: (2022)
by: Wang, Angtian, et al.
Published: (2022)
DIRECT-3D: Learning Direct Text-to-3D Generation on Massive Noisy 3D Data
by: Liu, Qihao, et al.
Published: (2024)
by: Liu, Qihao, et al.
Published: (2024)
iNeMo: Incremental Neural Mesh Models for Robust Class-Incremental Learning
by: Fischer, Tom, et al.
Published: (2024)
by: Fischer, Tom, et al.
Published: (2024)
Animal3D: A Comprehensive Dataset of 3D Animal Pose and Shape
by: Xu, Jiacong, et al.
Published: (2023)
by: Xu, Jiacong, et al.
Published: (2023)
A Bayesian Approach to OOD Robustness in Image Classification
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
TriDiff-4D: Fast 4D Generation through Diffusion-based Triplane Re-posing
by: Sheung, Eddie Pokming, et al.
Published: (2025)
by: Sheung, Eddie Pokming, et al.
Published: (2025)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
by: Liu, Qihao, et al.
Published: (2025)
by: Liu, Qihao, et al.
Published: (2025)
Structure-Aware Sparse-View X-ray 3D Reconstruction
by: Cai, Yuanhao, et al.
Published: (2023)
by: Cai, Yuanhao, et al.
Published: (2023)
4D-Animal: Freely Reconstructing Animatable 3D Animals from Videos
by: Zhong, Shanshan, et al.
Published: (2025)
by: Zhong, Shanshan, et al.
Published: (2025)
CamFreeDiff: Camera-free Image to Panorama Generation with Diffusion Model
by: Yuan, Xiaoding, et al.
Published: (2024)
by: Yuan, Xiaoding, et al.
Published: (2024)
Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose Estimation
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
HECTOR: Hybrid Editable Compositional Object References for Video Generation
by: Zhang, Guofeng, et al.
Published: (2026)
by: Zhang, Guofeng, et al.
Published: (2026)
Rethinking Video-Text Understanding: Retrieval from Counterfactually Augmented Data
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
SpatialReasoner: Towards Explicit and Generalizable 3D Spatial Reasoning
by: Ma, Wufei, et al.
Published: (2025)
by: Ma, Wufei, et al.
Published: (2025)
Radiative Gaussian Splatting for Efficient X-ray Novel View Synthesis
by: Cai, Yuanhao, et al.
Published: (2024)
by: Cai, Yuanhao, et al.
Published: (2024)
From Pixel to Cancer: Cellular Automata in Computed Tomography
by: Lai, Yuxiang, et al.
Published: (2024)
by: Lai, Yuxiang, et al.
Published: (2024)
3DSRBench: A Comprehensive 3D Spatial Reasoning Benchmark
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
ReVision: Refining Video Diffusion with Explicit 3D Motion Modeling
by: Liu, Qihao, et al.
Published: (2025)
by: Liu, Qihao, et al.
Published: (2025)
HDR-GS: Efficient High Dynamic Range Novel View Synthesis at 1000x Speed via Gaussian Splatting
by: Cai, Yuanhao, et al.
Published: (2024)
by: Cai, Yuanhao, et al.
Published: (2024)
LychSim: A Controllable and Interactive Simulation Framework for Vision Research
by: Ma, Wufei, et al.
Published: (2026)
by: Ma, Wufei, et al.
Published: (2026)
TGT: Text-Grounded Trajectories for Locally Controlled Video Generation
by: Zhang, Guofeng, et al.
Published: (2025)
by: Zhang, Guofeng, et al.
Published: (2025)
Spatial457: A Diagnostic Benchmark for 6D Spatial Reasoning of Large Multimodal Models
by: Wang, Xingrui, et al.
Published: (2025)
by: Wang, Xingrui, et al.
Published: (2025)
Evidential Neural Radiance Fields
by: Duan, Ruxiao, et al.
Published: (2026)
by: Duan, Ruxiao, et al.
Published: (2026)
Differences That Matter: Auditing Models for Capability Gap Discovery and Rectification
by: Liu, Qihao, et al.
Published: (2025)
by: Liu, Qihao, et al.
Published: (2025)
Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality Evolution
by: Liu, Qihao, et al.
Published: (2024)
by: Liu, Qihao, et al.
Published: (2024)
SpatialLLM: A Compound 3D-Informed Design towards Spatially-Intelligent Large Multimodal Models
by: Ma, Wufei, et al.
Published: (2025)
by: Ma, Wufei, et al.
Published: (2025)
Intermittency for the stochastic heat and wave equations with generalized fractional noise
by: Qian, Ruxiao
Published: (2025)
by: Qian, Ruxiao
Published: (2025)
PartInstruct: Part-level Instruction Following for Fine-grained Robot Manipulation
by: Yin, Yifan, et al.
Published: (2025)
by: Yin, Yifan, et al.
Published: (2025)
Leveraging AI Predicted and Expert Revised Annotations in Interactive Segmentation: Continual Tuning or Full Training?
by: Zhang, Tiezheng, et al.
Published: (2024)
by: Zhang, Tiezheng, et al.
Published: (2024)
Rapid Hydrolysis Precipitation of BiOBr Nanosheets for Boosting Photocatalytic N 2 Fixation
by: Liang Bao, et al.
Published: (2024)
by: Liang Bao, et al.
Published: (2024)
AbdomenAtlas-8K: Annotating 8,000 CT Volumes for Multi-Organ Segmentation in Three Weeks
by: Qu, Chongyu, et al.
Published: (2023)
by: Qu, Chongyu, et al.
Published: (2023)
Primordial Black Hole Formation from the Upward Step Model: Avoiding Overproduction
by: Wang, Xiaoding, et al.
Published: (2024)
by: Wang, Xiaoding, et al.
Published: (2024)
Captain Safari: A World Engine with Pose-Aligned 3D Memory
by: Chou, Yu-Cheng, et al.
Published: (2025)
by: Chou, Yu-Cheng, et al.
Published: (2025)
Similar Items
-
NOVUM: Neural Object Volumes for Robust Object Classification
by: Jesslen, Artur, et al.
Published: (2023) -
Prompt-Based Exemplar Super-Compression and Regeneration for Class-Incremental Learning
by: Duan, Ruxiao, et al.
Published: (2023) -
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding
by: Ma, Wufei, et al.
Published: (2024) -
Learning a Category-level Object Pose Estimator without Pose Annotations
by: Tian, Fengrui, et al.
Published: (2024) -
Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering
by: Wang, Xingrui, et al.
Published: (2024)