SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Xiangyu, Liu, Lijuan, Yan, Shuicheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Instant3D: Instant Text-to-3D Generation
von: Li, Ming, et al.
Veröffentlicht: (2023)
von: Li, Ming, et al.
Veröffentlicht: (2023)
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
von: Zhang, Zewei, et al.
Veröffentlicht: (2024)
von: Zhang, Zewei, et al.
Veröffentlicht: (2024)
Instruction-Driven 3D Facial Expression Generation and Transition
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
von: Bilecen, Bahri Batuhan, et al.
Veröffentlicht: (2024)
von: Bilecen, Bahri Batuhan, et al.
Veröffentlicht: (2024)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
Seeing World Dynamics in a Nutshell
von: Shen, Qiuhong, et al.
Veröffentlicht: (2025)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
von: Liu, Ziyuan, et al.
Veröffentlicht: (2026)
von: Liu, Ziyuan, et al.
Veröffentlicht: (2026)
ToonAging: Face Re-Aging upon Artistic Portrait Style Transfer
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
von: Singer, Assaf, et al.
Veröffentlicht: (2025)
von: Singer, Assaf, et al.
Veröffentlicht: (2025)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
von: Chiu, Pin-Yen, et al.
Veröffentlicht: (2025)
von: Chiu, Pin-Yen, et al.
Veröffentlicht: (2025)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
von: Girdhar, Rohit, et al.
Veröffentlicht: (2023)
von: Girdhar, Rohit, et al.
Veröffentlicht: (2023)
Cross-Scenario Deraining Adaptation with Unpaired Data: Superpixel Structural Priors and Multi-Stage Pseudo-Rain Synthesis
von: Zhao, Kangbo, et al.
Veröffentlicht: (2026)
von: Zhao, Kangbo, et al.
Veröffentlicht: (2026)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
von: Dai, Yudi, et al.
Veröffentlicht: (2024)
von: Dai, Yudi, et al.
Veröffentlicht: (2024)
A Survey on 3D Gaussian Splatting
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
von: Lin, Shengjie, et al.
Veröffentlicht: (2025)
von: Lin, Shengjie, et al.
Veröffentlicht: (2025)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
von: Flynn, John, et al.
Veröffentlicht: (2026)
von: Flynn, John, et al.
Veröffentlicht: (2026)
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
Reinforcement Learning-Driven Edge Management for Reliable Multi-view 3D Reconstruction
von: Mounesan, Motahare, et al.
Veröffentlicht: (2025)
von: Mounesan, Motahare, et al.
Veröffentlicht: (2025)
Neural Isometries: Taming Transformations for Equivariant ML
von: Mitchel, Thomas W., et al.
Veröffentlicht: (2024)
von: Mitchel, Thomas W., et al.
Veröffentlicht: (2024)
SCULPT: Shape-Conditioned Unpaired Learning of Pose-dependent Clothed and Textured Human Meshes
von: Sanyal, Soubhik, et al.
Veröffentlicht: (2023)
von: Sanyal, Soubhik, et al.
Veröffentlicht: (2023)
Size Matters: Reconstructing Real-Scale 3D Models from Monocular Images for Food Portion Estimation
von: Vinod, Gautham, et al.
Veröffentlicht: (2026)
von: Vinod, Gautham, et al.
Veröffentlicht: (2026)
SMPLest-X: Ultimate Scaling for Expressive Human Pose and Shape Estimation
von: Yin, Wanqi, et al.
Veröffentlicht: (2025)
von: Yin, Wanqi, et al.
Veröffentlicht: (2025)
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
von: Zakharov, Sergey, et al.
Veröffentlicht: (2024)
von: Zakharov, Sergey, et al.
Veröffentlicht: (2024)
Lester: rotoscope animation through video object segmentation and tracking
von: Tous, Ruben
Veröffentlicht: (2024)
von: Tous, Ruben
Veröffentlicht: (2024)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
von: Sar, Ayan, et al.
Veröffentlicht: (2025)
von: Sar, Ayan, et al.
Veröffentlicht: (2025)
Extreme Compression of Adaptive Neural Images
von: Hoshikawa, Leo, et al.
Veröffentlicht: (2024)
von: Hoshikawa, Leo, et al.
Veröffentlicht: (2024)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
Freehand Sketch Generation from Mechanical Components
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
DGS-LRM: Real-Time Deformable 3D Gaussian Reconstruction From Monocular Videos
von: Lin, Chieh Hubert, et al.
Veröffentlicht: (2025)
von: Lin, Chieh Hubert, et al.
Veröffentlicht: (2025)
AudCast: Audio-Driven Human Video Generation by Cascaded Diffusion Transformers
von: Guan, Jiazhi, et al.
Veröffentlicht: (2025)
von: Guan, Jiazhi, et al.
Veröffentlicht: (2025)
Improving Generative Adversarial Network Generalization for Facial Expression Synthesis
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
von: Lin, Jiantao, et al.
Veröffentlicht: (2025)
von: Lin, Jiantao, et al.
Veröffentlicht: (2025)
GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
von: Wang, Xuanchen, et al.
Veröffentlicht: (2025)
von: Wang, Xuanchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Instant3D: Instant Text-to-3D Generation
von: Li, Ming, et al.
Veröffentlicht: (2023) -
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
von: Zhang, Zewei, et al.
Veröffentlicht: (2024) -
Instruction-Driven 3D Facial Expression Generation and Transition
von: Vo, Anh H., et al.
Veröffentlicht: (2026) -
Identity Preserving 3D Head Stylization with Multiview Score Distillation
von: Bilecen, Bahri Batuhan, et al.
Veröffentlicht: (2024) -
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)