Direct and Explicit 3D Generation from a Single Image
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Haoyu, Karumuri, Meher Gitika, Zou, Chuhang, Bang, Seungbae, Li, Yuelong, Samaras, Dimitris, Hadap, Sunil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MRC-Net: 6-DoF Pose Estimation with MultiScale Residual Correlation
di: Li, Yuelong, et al.
Pubblicazione: (2024)
di: Li, Yuelong, et al.
Pubblicazione: (2024)
Importance-Based Token Merging for Efficient Image and Video Generation
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
Learning 3D Reconstruction with Priors in Test Time
di: Zhou, Lei, et al.
Pubblicazione: (2026)
di: Zhou, Lei, et al.
Pubblicazione: (2026)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
di: Chakraborty, Souradeep, et al.
Pubblicazione: (2024)
di: Chakraborty, Souradeep, et al.
Pubblicazione: (2024)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
di: Chatziagapi, Aggelina, et al.
Pubblicazione: (2024)
di: Chatziagapi, Aggelina, et al.
Pubblicazione: (2024)
Assessing Sample Quality via the Latent Space of Generative Models
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
Learning Relighting and Intrinsic Decomposition in Neural Radiance Fields
di: Yang, Yixiong, et al.
Pubblicazione: (2024)
di: Yang, Yixiong, et al.
Pubblicazione: (2024)
MLI-NeRF: Multi-Light Intrinsic-Aware Neural Radiance Fields
di: Yang, Yixiong, et al.
Pubblicazione: (2024)
di: Yang, Yixiong, et al.
Pubblicazione: (2024)
GriDiT: Factorized Grid-Based Diffusion for Efficient Long Image Sequence Generation
di: Tomar, Snehal Singh, et al.
Pubblicazione: (2025)
di: Tomar, Snehal Singh, et al.
Pubblicazione: (2025)
Direct May Not Be the Best: An Incremental Evolution View of Pose Generation
di: Li, Yuelong, et al.
Pubblicazione: (2024)
di: Li, Yuelong, et al.
Pubblicazione: (2024)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
di: Chakkera, Sai Tanmay Reddy, et al.
Pubblicazione: (2024)
di: Chakkera, Sai Tanmay Reddy, et al.
Pubblicazione: (2024)
MVGBench: Comprehensive Benchmark for Multi-view Generation Models
di: Xie, Xianghui, et al.
Pubblicazione: (2025)
di: Xie, Xianghui, et al.
Pubblicazione: (2025)
TopoDiffusionNet: A Topology-aware Diffusion Model
di: Gupta, Saumya, et al.
Pubblicazione: (2024)
di: Gupta, Saumya, et al.
Pubblicazione: (2024)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
di: Howlader, Prantik, et al.
Pubblicazione: (2024)
di: Howlader, Prantik, et al.
Pubblicazione: (2024)
PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization
di: Wu, Dongli, et al.
Pubblicazione: (2026)
di: Wu, Dongli, et al.
Pubblicazione: (2026)
Talking Head Generation via AU-Guided Landmark Prediction
di: Chang, Shao-Yu, et al.
Pubblicazione: (2025)
di: Chang, Shao-Yu, et al.
Pubblicazione: (2025)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
di: Chatziagapi, Aggelina, et al.
Pubblicazione: (2024)
di: Chatziagapi, Aggelina, et al.
Pubblicazione: (2024)
Rig3DGS: Creating Controllable Portraits from Casual Monocular Videos
di: Rivero, Alfredo, et al.
Pubblicazione: (2024)
di: Rivero, Alfredo, et al.
Pubblicazione: (2024)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
di: Hu, Shilin, et al.
Pubblicazione: (2025)
di: Hu, Shilin, et al.
Pubblicazione: (2025)
Language-driven Description Generation and Common Sense Reasoning for Video Action Recognition
di: Hu, Xiaodan, et al.
Pubblicazione: (2025)
di: Hu, Xiaodan, et al.
Pubblicazione: (2025)
Fast SAM 3D Body: Accelerating SAM 3D Body for Real-Time Full-Body Human Mesh Recovery
di: Yang, Timing, et al.
Pubblicazione: (2026)
di: Yang, Timing, et al.
Pubblicazione: (2026)
Fast constrained sampling in pre-trained diffusion models
di: Graikos, Alexandros, et al.
Pubblicazione: (2024)
di: Graikos, Alexandros, et al.
Pubblicazione: (2024)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
di: Howlader, Prantik, et al.
Pubblicazione: (2024)
di: Howlader, Prantik, et al.
Pubblicazione: (2024)
Scalable and Realistic Virtual Try-on Application for Foundation Makeup with Kubelka-Munk Theory
di: Pang, Hui, et al.
Pubblicazione: (2025)
di: Pang, Hui, et al.
Pubblicazione: (2025)
Hummingbird: High Fidelity Image Generation via Multimodal Context Alignment
di: Le, Minh-Quan, et al.
Pubblicazione: (2025)
di: Le, Minh-Quan, et al.
Pubblicazione: (2025)
$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
di: Le, Minh-Quan, et al.
Pubblicazione: (2024)
di: Le, Minh-Quan, et al.
Pubblicazione: (2024)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
di: Le, Minh-Quan, et al.
Pubblicazione: (2025)
di: Le, Minh-Quan, et al.
Pubblicazione: (2025)
PathSegDiff: Pathology Segmentation using Diffusion model representations
di: Danisetty, Sachin Kumar, et al.
Pubblicazione: (2025)
di: Danisetty, Sachin Kumar, et al.
Pubblicazione: (2025)
MonoPatchNeRF: Improving Neural Radiance Fields with Patch-based Monocular Guidance
di: Wu, Yuqun, et al.
Pubblicazione: (2024)
di: Wu, Yuqun, et al.
Pubblicazione: (2024)
Plenoptic PNG: Real-Time Neural Radiance Fields in 150 KB
di: Lee, Jae Yong, et al.
Pubblicazione: (2024)
di: Lee, Jae Yong, et al.
Pubblicazione: (2024)
Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars
di: Gong, Yicheng, et al.
Pubblicazione: (2026)
di: Gong, Yicheng, et al.
Pubblicazione: (2026)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
Cue3D: Quantifying the Role of Image Cues in Single-Image 3D Generation
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
LBMamba: Locally Bi-directional Mamba
di: Zhang, Jingwei, et al.
Pubblicazione: (2025)
di: Zhang, Jingwei, et al.
Pubblicazione: (2025)
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
di: Hu, Shilin, et al.
Pubblicazione: (2025)
di: Hu, Shilin, et al.
Pubblicazione: (2025)
ESGaussianFace: Emotional and Stylized Audio-Driven Facial Animation via 3D Gaussian Splatting
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
Phrase-Instance Alignment for Generalized Referring Segmentation
di: Nguyen, E-Ro, et al.
Pubblicazione: (2024)
di: Nguyen, E-Ro, et al.
Pubblicazione: (2024)
Personalized Image Descriptions from Attention Sequences
di: Xue, Ruoyu, et al.
Pubblicazione: (2025)
di: Xue, Ruoyu, et al.
Pubblicazione: (2025)
Human as Points: Explicit Point-based 3D Human Reconstruction from Single-view RGB Images
di: Tang, Yingzhi, et al.
Pubblicazione: (2023)
di: Tang, Yingzhi, et al.
Pubblicazione: (2023)
Documenti analoghi
-
MRC-Net: 6-DoF Pose Estimation with MultiScale Residual Correlation
di: Li, Yuelong, et al.
Pubblicazione: (2024) -
Importance-Based Token Merging for Efficient Image and Video Generation
di: Wu, Haoyu, et al.
Pubblicazione: (2024) -
Learning 3D Reconstruction with Priors in Test Time
di: Zhou, Lei, et al.
Pubblicazione: (2026) -
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
di: Wu, Haoyu, et al.
Pubblicazione: (2025) -
Self-supervised co-salient object detection via feature correspondence at multiple scales
di: Chakraborty, Souradeep, et al.
Pubblicazione: (2024)