Deep learning for 3D human pose estimation and mesh recovery: A survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yang, Qiu, Changzhen, Zhang, Zhiyong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Textured mesh Quality Assessment using Geometry and Color Field Similarity
von: Yang, Kaifa, et al.
Veröffentlicht: (2025)
von: Yang, Kaifa, et al.
Veröffentlicht: (2025)
BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion
von: Jia, Tianzhi, et al.
Veröffentlicht: (2026)
von: Jia, Tianzhi, et al.
Veröffentlicht: (2026)
MOC-3D: Manifold-Order Consistency for Text-to-3D Generation
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
3D Gaussian Editing with A Single Image
von: Luo, Guan, et al.
Veröffentlicht: (2024)
von: Luo, Guan, et al.
Veröffentlicht: (2024)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
Generalized Video Anomaly Event Detection: Systematic Taxonomy and Comparison of Deep Models
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
von: Li, Zeju, et al.
Veröffentlicht: (2024)
von: Li, Zeju, et al.
Veröffentlicht: (2024)
GeoLink: A 3D-Aware Framework Towards Better Generalization in Cross-View Geo-Localization
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
RealX3D: A Physically-Degraded 3D Benchmark for Multi-view Visual Restoration and Reconstruction
von: Liu, Shuhong, et al.
Veröffentlicht: (2025)
von: Liu, Shuhong, et al.
Veröffentlicht: (2025)
Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval
von: Ding, Yiming, et al.
Veröffentlicht: (2026)
von: Ding, Yiming, et al.
Veröffentlicht: (2026)
SkyLink: Unifying Street-Satellite Geo-Localization via UAV-Mediated 3D Scene Alignment
von: Zhang, Hongyang, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyang, et al.
Veröffentlicht: (2025)
A Hierarchical Compression Technique for 3D Gaussian Splatting Compression
von: Huang, He, et al.
Veröffentlicht: (2024)
von: Huang, He, et al.
Veröffentlicht: (2024)
Human Motion Video Generation: A Survey
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
Querying Autonomous Vehicle Point Clouds: Enhanced by 3D Object Counting with CounterNet
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
InstructHumans: Editing Animated 3D Human Textures with Instructions
von: Zhu, Jiayin, et al.
Veröffentlicht: (2024)
von: Zhu, Jiayin, et al.
Veröffentlicht: (2024)
Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset
von: Mo, Wentao, et al.
Veröffentlicht: (2025)
von: Mo, Wentao, et al.
Veröffentlicht: (2025)
Magic3DSketch: Create Colorful 3D Models From Sketch-Based 3D Modeling Guided by Text and Language-Image Pre-Training
von: Zang, Ying, et al.
Veröffentlicht: (2024)
von: Zang, Ying, et al.
Veröffentlicht: (2024)
Enhancing 3D Gaussian Splatting Compression via Spatial Condition-based Prediction
von: Ma, Jingui, et al.
Veröffentlicht: (2025)
von: Ma, Jingui, et al.
Veröffentlicht: (2025)
Dual Attribute-Spatial Relation Alignment for 3D Visual Grounding
von: Xu, Yue, et al.
Veröffentlicht: (2024)
von: Xu, Yue, et al.
Veröffentlicht: (2024)
GaussianForest: Hierarchical-Hybrid 3D Gaussian Splatting for Compressed Scene Modeling
von: Zhang, Fengyi, et al.
Veröffentlicht: (2024)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2024)
SizeGS: Size-aware Compression of 3D Gaussian Splatting via Mixed Integer Programming
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
Radio Frequency Signal based Human Silhouette Segmentation: A Sequential Diffusion Approach
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
Adaptive 3D Gaussian Splatting Video Streaming
von: Gong, Han, et al.
Veröffentlicht: (2025)
von: Gong, Han, et al.
Veröffentlicht: (2025)
DeepSPG: Exploring Deep Semantic Prior Guidance for Low-light Image Enhancement with Multimodal Learning
von: Lu, Jialang, et al.
Veröffentlicht: (2025)
von: Lu, Jialang, et al.
Veröffentlicht: (2025)
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
StableDub: Taming Diffusion Prior for Generalized and Efficient Visual Dubbing
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
MotionPro: A Precise Motion Controller for Image-to-Video Generation
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
Sketch and Patch: Efficient 3D Gaussian Representation for Man-Made Scenes
von: Shi, Yuang, et al.
Veröffentlicht: (2025)
von: Shi, Yuang, et al.
Veröffentlicht: (2025)
DanceCamera3D: 3D Camera Movement Synthesis with Music and Dance
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
VIoTGPT: Learning to Schedule Vision Tools in LLMs towards Intelligent Video Internet of Things
von: Zhong, Yaoyao, et al.
Veröffentlicht: (2023)
von: Zhong, Yaoyao, et al.
Veröffentlicht: (2023)
DIP: Diffusion Learning of Inconsistency Pattern for General DeepFake Detection
von: Nie, Fan, et al.
Veröffentlicht: (2024)
von: Nie, Fan, et al.
Veröffentlicht: (2024)
Deep-JGAC: End-to-End Deep Joint Geometry and Attribute Compression for Dense Colored Point Clouds
von: Zhang, Yun, et al.
Veröffentlicht: (2025)
von: Zhang, Yun, et al.
Veröffentlicht: (2025)
3D2M Dataset: A 3-Dimension diverse Mesh Dataset
von: Dasgupta, Sankarshan
Veröffentlicht: (2024)
von: Dasgupta, Sankarshan
Veröffentlicht: (2024)
Rendering-Oriented 3D Point Cloud Attribute Compression using Sparse Tensor-based Transformer
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
HCNQA: Enhancing 3D VQA with Hierarchical Concentration Narrowing Supervision
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
T$^\text{3}$SVFND: Towards an Evolving Fake News Detector for Emergencies with Test-time Training on Short Video Platforms
von: Zhang, Liyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Liyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery
von: Liu, Yang, et al.
Veröffentlicht: (2026) -
Textured mesh Quality Assessment using Geometry and Color Field Similarity
von: Yang, Kaifa, et al.
Veröffentlicht: (2025) -
BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion
von: Jia, Tianzhi, et al.
Veröffentlicht: (2026) -
MOC-3D: Manifold-Order Consistency for Text-to-3D Generation
von: Fan, Chenyang, et al.
Veröffentlicht: (2026) -
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)