Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Choi, Seonghwa, Choi, Moonkyeong, Jang, Mingyu, Kim, Jaekyung, Cai, Jianfei, Cheng, Wen-Huang, Lee, Sanghoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
von: Zafeiri, Maria, et al.
Veröffentlicht: (2024)
von: Zafeiri, Maria, et al.
Veröffentlicht: (2024)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
von: Gerats, Beerend G. A., et al.
Veröffentlicht: (2024)
von: Gerats, Beerend G. A., et al.
Veröffentlicht: (2024)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
von: Lu, Wanglong, et al.
Veröffentlicht: (2025)
von: Lu, Wanglong, et al.
Veröffentlicht: (2025)
ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection
von: Samson, Hema Hariharan
Veröffentlicht: (2026)
von: Samson, Hema Hariharan
Veröffentlicht: (2026)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
von: Sartinas, Evangelos G., et al.
Veröffentlicht: (2023)
von: Sartinas, Evangelos G., et al.
Veröffentlicht: (2023)
Eleven Primitives and Three Gates: The Universal Structure of Computational Imaging
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
von: Lu, Wanglong, et al.
Veröffentlicht: (2022)
von: Lu, Wanglong, et al.
Veröffentlicht: (2022)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
Saliency-Aware Diffusion Reconstruction for Effective Invisible Watermark Removal
von: Alam, Inzamamul, et al.
Veröffentlicht: (2025)
von: Alam, Inzamamul, et al.
Veröffentlicht: (2025)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
von: Ömercikoğlu, Ahmet Can, et al.
Veröffentlicht: (2025)
von: Ömercikoğlu, Ahmet Can, et al.
Veröffentlicht: (2025)
Digital analysis of early color photographs taken using regular color screen processes
von: Hubička, Jan, et al.
Veröffentlicht: (2023)
von: Hubička, Jan, et al.
Veröffentlicht: (2023)
Revealing an Unattractivity Bias in Mental Reconstruction of Occluded Faces using Generative Image Models
von: Riedmann, Frederik, et al.
Veröffentlicht: (2024)
von: Riedmann, Frederik, et al.
Veröffentlicht: (2024)
Skullptor: High Fidelity 3D Head Reconstruction in Seconds with Multi-View Normal Prediction
von: Artru, Noé, et al.
Veröffentlicht: (2026)
von: Artru, Noé, et al.
Veröffentlicht: (2026)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
von: Long, Yuchong, et al.
Veröffentlicht: (2025)
von: Long, Yuchong, et al.
Veröffentlicht: (2025)
SelectiveKD: A semi-supervised framework for cancer detection in DBT through Knowledge Distillation and Pseudo-labeling
von: Dillard, Laurent, et al.
Veröffentlicht: (2024)
von: Dillard, Laurent, et al.
Veröffentlicht: (2024)
Mono-Modalizing Extremely Heterogeneous Multi-Modal Medical Image Registration
von: Choo, Kyobin, et al.
Veröffentlicht: (2025)
von: Choo, Kyobin, et al.
Veröffentlicht: (2025)
Deep Spectral Meshes: Multi-Frequency Facial Mesh Processing with Graph Neural Networks
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting
von: Maheshkar, Vaishali, et al.
Veröffentlicht: (2025)
von: Maheshkar, Vaishali, et al.
Veröffentlicht: (2025)
Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives
von: He, Shijing, et al.
Veröffentlicht: (2025)
von: He, Shijing, et al.
Veröffentlicht: (2025)
Deep Learning Approaches for Medical Imaging Under Varying Degrees of Label Availability: A Comprehensive Survey
von: Ma, Siteng, et al.
Veröffentlicht: (2025)
von: Ma, Siteng, et al.
Veröffentlicht: (2025)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Rethinking VLMs for Image Forgery Detection and Localization
von: Guo, Shaofeng, et al.
Veröffentlicht: (2026)
von: Guo, Shaofeng, et al.
Veröffentlicht: (2026)
Med-IC: Fusing a Single Layer Involution with Convolutions for Enhanced Medical Image Classification and Segmentation
von: Islam, Md. Farhadul, et al.
Veröffentlicht: (2024)
von: Islam, Md. Farhadul, et al.
Veröffentlicht: (2024)
Efficient and Privacy-Protecting Background Removal for 2D Video Streaming using iPhone 15 Pro Max LiDAR
von: Kinnevan, Jessica, et al.
Veröffentlicht: (2025)
von: Kinnevan, Jessica, et al.
Veröffentlicht: (2025)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
Person detection and re-identification in open-world settings of retail stores and public spaces
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection
von: Baek, Sunghwan, et al.
Veröffentlicht: (2026)
von: Baek, Sunghwan, et al.
Veröffentlicht: (2026)
Under-Canopy Terrain Reconstruction in Dense Forests Using RGB Imaging and Neural 3D Reconstruction
von: Sheffer, Refael, et al.
Veröffentlicht: (2026)
von: Sheffer, Refael, et al.
Veröffentlicht: (2026)
Slice-Consistent 3D Volumetric Brain CT-to-MRI Translation with 2D Brownian Bridge Diffusion Model
von: Choo, Kyobin, et al.
Veröffentlicht: (2024)
von: Choo, Kyobin, et al.
Veröffentlicht: (2024)
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
von: Egorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Egorov, Konstantin, et al.
Veröffentlicht: (2025)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
Ring Artifacts Removal Based on Implicit Neural Representation of Sinogram Data
von: Shi, Ligen, et al.
Veröffentlicht: (2024)
von: Shi, Ligen, et al.
Veröffentlicht: (2024)
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
Improving Visual Object Tracking through Visual Prompting
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2024)
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
von: Lu, Wanglong, et al.
Veröffentlicht: (2024) -
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
von: Lu, Wanglong, et al.
Veröffentlicht: (2024) -
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
von: Zafeiri, Maria, et al.
Veröffentlicht: (2024) -
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
von: Gerats, Beerend G. A., et al.
Veröffentlicht: (2024) -
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
von: Lu, Wanglong, et al.
Veröffentlicht: (2025)