Pomo3D: 3D-Aware Portrait Accessorizing and More
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Tzu-Chieh, Liu, Chih-Ting, Chien, Shao-Yi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SPLite Hand: Sparsity-Aware Lightweight 3D Hand Pose Estimation
por: Hao, Yeh Keng, et al.
Publicado: (2025)
por: Hao, Yeh Keng, et al.
Publicado: (2025)
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
por: Wei, Haoran, et al.
Publicado: (2024)
por: Wei, Haoran, et al.
Publicado: (2024)
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
por: Xia, Jiatong, et al.
Publicado: (2026)
por: Xia, Jiatong, et al.
Publicado: (2026)
CA-W3D: Leveraging Context-Aware Knowledge for Weakly Supervised Monocular 3D Detection
por: Liu, Chupeng, et al.
Publicado: (2025)
por: Liu, Chupeng, et al.
Publicado: (2025)
Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning
por: Chien, Tzu-Chun, et al.
Publicado: (2025)
por: Chien, Tzu-Chun, et al.
Publicado: (2025)
Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models
por: Wang, Xiaoyan, et al.
Publicado: (2025)
por: Wang, Xiaoyan, et al.
Publicado: (2025)
Before the Shutter: Aesthetic and Actionable Portrait Photography Planning in 3D Scenes
por: Jiang, Ruixiang, et al.
Publicado: (2026)
por: Jiang, Ruixiang, et al.
Publicado: (2026)
SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images
por: Liu, Zishan, et al.
Publicado: (2026)
por: Liu, Zishan, et al.
Publicado: (2026)
3DID: Direct 3D Inverse Design for Aerodynamics with Physics-Aware Optimization
por: Hao, Yuze, et al.
Publicado: (2025)
por: Hao, Yuze, et al.
Publicado: (2025)
Learning to Generate Conditional Tri-plane for 3D-aware Expression Controllable Portrait Animation
por: Ki, Taekyung, et al.
Publicado: (2024)
por: Ki, Taekyung, et al.
Publicado: (2024)
EgoDTM: Towards 3D-Aware Egocentric Video-Language Pretraining
por: Xu, Boshen, et al.
Publicado: (2025)
por: Xu, Boshen, et al.
Publicado: (2025)
MetaFind: Scene-Aware 3D Asset Retrieval for Coherent Metaverse Scene Generation
por: Pan, Zhenyu, et al.
Publicado: (2025)
por: Pan, Zhenyu, et al.
Publicado: (2025)
Dual-Domain Representation Alignment: Bridging 2D and 3D Vision via Geometry-Aware Architecture Search
por: Zhang, Haoyu, et al.
Publicado: (2026)
por: Zhang, Haoyu, et al.
Publicado: (2026)
SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation
por: Agrawal, Vaibhav, et al.
Publicado: (2026)
por: Agrawal, Vaibhav, et al.
Publicado: (2026)
SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild
por: Hu, Xuyi, et al.
Publicado: (2026)
por: Hu, Xuyi, et al.
Publicado: (2026)
How Far are AI-generated Videos from Simulating the 3D Visual World: A Learned 3D Evaluation Approach
por: Chang, Chirui, et al.
Publicado: (2024)
por: Chang, Chirui, et al.
Publicado: (2024)
Hunyuan3D 1.0: A Unified Framework for Text-to-3D and Image-to-3D Generation
por: Yang, Xianghui, et al.
Publicado: (2024)
por: Yang, Xianghui, et al.
Publicado: (2024)
Audio-visual Event Localization on Portrait Mode Short Videos
por: Liu, Wuyang, et al.
Publicado: (2025)
por: Liu, Wuyang, et al.
Publicado: (2025)
CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images
por: Liu, Jian, et al.
Publicado: (2024)
por: Liu, Jian, et al.
Publicado: (2024)
RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing
por: Zhao, Longjie, et al.
Publicado: (2025)
por: Zhao, Longjie, et al.
Publicado: (2025)
RenderWorld: World Model with Self-Supervised 3D Label
por: Yan, Ziyang, et al.
Publicado: (2024)
por: Yan, Ziyang, et al.
Publicado: (2024)
TGP: Two-modal occupancy prediction with 3D Gaussian and sparse points for 3D Environment Awareness
por: Chen, Mu, et al.
Publicado: (2025)
por: Chen, Mu, et al.
Publicado: (2025)
A Framework for Portrait Stylization with Skin-Tone Awareness and Nudity Identification
por: Kim, Seungkwon, et al.
Publicado: (2024)
por: Kim, Seungkwon, et al.
Publicado: (2024)
Dynamics-Aware Gaussian Splatting Streaming Towards Fast On-the-Fly 4D Reconstruction
por: Liu, Zhening, et al.
Publicado: (2024)
por: Liu, Zhening, et al.
Publicado: (2024)
EndoVGGT: GNN-Enhanced Depth Estimation for Surgical 3D Reconstruction
por: Fan, Falong, et al.
Publicado: (2026)
por: Fan, Falong, et al.
Publicado: (2026)
NavCrafter: Exploring 3D Scenes from a Single Image
por: Duan, Hongbo, et al.
Publicado: (2026)
por: Duan, Hongbo, et al.
Publicado: (2026)
3DGS-Enhancer: Enhancing Unbounded 3D Gaussian Splatting with View-consistent 2D Diffusion Priors
por: Liu, Xi, et al.
Publicado: (2024)
por: Liu, Xi, et al.
Publicado: (2024)
MVP4D: Multi-View Portrait Video Diffusion for Animatable 4D Avatars
por: Taubner, Felix, et al.
Publicado: (2025)
por: Taubner, Felix, et al.
Publicado: (2025)
Semantic Aware Feature Extraction for Enhanced 3D Reconstruction
por: Nap, Ronald, et al.
Publicado: (2026)
por: Nap, Ronald, et al.
Publicado: (2026)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
por: Xie, You, et al.
Publicado: (2024)
por: Xie, You, et al.
Publicado: (2024)
UW-3DGS: Underwater 3D Reconstruction with Physics-Aware Gaussian Splatting
por: Xing, Wenpeng, et al.
Publicado: (2025)
por: Xing, Wenpeng, et al.
Publicado: (2025)
Is 3D Convolution with 5D Tensors Really Necessary for Video Analysis?
por: Hajimolahoseini, Habib, et al.
Publicado: (2024)
por: Hajimolahoseini, Habib, et al.
Publicado: (2024)
Incorporating Eye-Tracking Signals Into Multimodal Deep Visual Models For Predicting User Aesthetic Experience In Residential Interiors
por: Chien, Chen-Ying, et al.
Publicado: (2026)
por: Chien, Chen-Ying, et al.
Publicado: (2026)
HY3D-Bench: Generation of 3D Assets
por: Hunyuan3D, Team, et al.
Publicado: (2026)
por: Hunyuan3D, Team, et al.
Publicado: (2026)
DOR3D-Net: Dense Ordinal Regression Network for 3D Hand Pose Estimation
por: Mao, Yamin, et al.
Publicado: (2024)
por: Mao, Yamin, et al.
Publicado: (2024)
Hunyuan3D 2.5: Towards High-Fidelity 3D Assets Generation with Ultimate Details
por: Lai, Zeqiang, et al.
Publicado: (2025)
por: Lai, Zeqiang, et al.
Publicado: (2025)
MRN: Harnessing 2D Vision Foundation Models for Diagnosing Parkinson's Disease with Limited 3D MR Data
por: Shaodong, Ding, et al.
Publicado: (2025)
por: Shaodong, Ding, et al.
Publicado: (2025)
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
por: Burgess, James, et al.
Publicado: (2023)
por: Burgess, James, et al.
Publicado: (2023)
A Comprehensive Survey on 3D Content Generation
por: Liu, Jian, et al.
Publicado: (2024)
por: Liu, Jian, et al.
Publicado: (2024)
MoViD: View-Invariant 3D Human Pose Estimation via Motion-View Disentanglement
por: Liu, Yejia, et al.
Publicado: (2026)
por: Liu, Yejia, et al.
Publicado: (2026)
Ejemplares similares
-
SPLite Hand: Sparsity-Aware Lightweight 3D Hand Pose Estimation
por: Hao, Yeh Keng, et al.
Publicado: (2025) -
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
por: Wei, Haoran, et al.
Publicado: (2024) -
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
por: Xia, Jiatong, et al.
Publicado: (2026) -
CA-W3D: Leveraging Context-Aware Knowledge for Weakly Supervised Monocular 3D Detection
por: Liu, Chupeng, et al.
Publicado: (2025) -
Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning
por: Chien, Tzu-Chun, et al.
Publicado: (2025)