SpinMeRound: Consistent Multi-View Identity Generation Using Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Galanakis, Stathis, Lattas, Alexandros, Moschoglou, Stylianos, Kainz, Bernhard, Zafeiriou, Stefanos |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2023)
by: Galanakis, Stathis, et al.
Published: (2023)
Arc2Face: A Foundation Model for ID-Consistent Human Faces
by: Papantoniou, Foivos Paraperas, et al.
Published: (2024)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2024)
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
AnimateMe: 4D Facial Expressions via Diffusion Models
by: Gerogiannis, Dimitrios, et al.
Published: (2024)
by: Gerogiannis, Dimitrios, et al.
Published: (2024)
DermaFlux: Synthetic Skin Lesion Generation with Rectified Flows for Enhanced Image Classification
by: Galanakis, Stathis, et al.
Published: (2026)
by: Galanakis, Stathis, et al.
Published: (2026)
ImHead: A Large-scale Implicit Morphable Model for Localized Head Modeling
by: Potamias, Rolandos Alexandros, et al.
Published: (2025)
by: Potamias, Rolandos Alexandros, et al.
Published: (2025)
Improving face generation quality and prompt following with synthetic captions
by: Tarasiou, Michail, et al.
Published: (2024)
by: Tarasiou, Michail, et al.
Published: (2024)
ID-to-3D: Expressive ID-guided 3D Heads via Score Distillation Sampling
by: Babiloni, Francesca, et al.
Published: (2024)
by: Babiloni, Francesca, et al.
Published: (2024)
Arc2Avatar: Generating Expressive 3D Avatars from a Single Image via ID Guidance
by: Gerogiannis, Dimitrios, et al.
Published: (2025)
by: Gerogiannis, Dimitrios, et al.
Published: (2025)
Geo-ID: Test-Time Geometric Consensus for Cross-View Consistent Intrinsics
by: Dirik, Alara, et al.
Published: (2026)
by: Dirik, Alara, et al.
Published: (2026)
ID-Consistent, Precise Expression Generation with Blendshape-Guided Diffusion
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
Physical Simulator In-the-Loop Video Generation
by: Foo, Lin Geng, et al.
Published: (2026)
by: Foo, Lin Geng, et al.
Published: (2026)
ShapeFusion: A 3D diffusion model for localized shape editing
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
Locally Adaptive Neural 3D Morphable Models
by: Tarasiou, Michail, et al.
Published: (2024)
by: Tarasiou, Michail, et al.
Published: (2024)
Design2Cloth: 3D Cloth Generation from 2D Masks
by: Zheng, Jiali, et al.
Published: (2024)
by: Zheng, Jiali, et al.
Published: (2024)
MaDiS: Taming Masked Diffusion Language Models for Sign Language Generation
by: Zuo, Ronglai, et al.
Published: (2026)
by: Zuo, Ronglai, et al.
Published: (2026)
Enabling PSO-Secure Synthetic Data Sharing Using Diversity-Aware Diffusion Models
by: Dombrowski, Mischa, et al.
Published: (2025)
by: Dombrowski, Mischa, et al.
Published: (2025)
UV-free Texture Generation with Denoising and Geodesic Heat Diffusions
by: Foti, Simone, et al.
Published: (2024)
by: Foti, Simone, et al.
Published: (2024)
WiLoR: End-to-end 3D Hand Localization and Reconstruction in-the-wild
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
Distribution Matching for Multi-Task Learning of Classification Tasks: a Large-Scale Study on Faces & Beyond
by: Kollias, Dimitrios, et al.
Published: (2024)
by: Kollias, Dimitrios, et al.
Published: (2024)
Dex2HOI: Dexterous Bimanual Two-Object Interaction Generation
by: Pratikaki, Chrysa, et al.
Published: (2026)
by: Pratikaki, Chrysa, et al.
Published: (2026)
Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language Generator
by: Zuo, Ronglai, et al.
Published: (2024)
by: Zuo, Ronglai, et al.
Published: (2024)
Uncovering Hidden Subspaces in Video Diffusion Models Using Re-Identification
by: Dombrowski, Mischa, et al.
Published: (2024)
by: Dombrowski, Mischa, et al.
Published: (2024)
JVID: Joint Video-Image Diffusion for Visual-Quality and Temporal-Consistency in Video Generation
by: Reynaud, Hadrien, et al.
Published: (2024)
by: Reynaud, Hadrien, et al.
Published: (2024)
Noise Crystallization and Liquid Noise: Zero-shot Video Generation using Image Diffusion Models
by: Khan, Muhammad Haaris, et al.
Published: (2024)
by: Khan, Muhammad Haaris, et al.
Published: (2024)
SAGS: Structure-Aware 3D Gaussian Splatting
by: Ververas, Evangelos, et al.
Published: (2024)
by: Ververas, Evangelos, et al.
Published: (2024)
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024)
by: Yu, Zhengdi, et al.
Published: (2024)
HO-Flow: Generalizable Hand-Object Interaction Generation with Latent Flow Matching
by: Chen, Zerui, et al.
Published: (2026)
by: Chen, Zerui, et al.
Published: (2026)
ShapePuri: Shape Guided and Appearance Generalized Adversarial Purification
by: Li, Zhe, et al.
Published: (2026)
by: Li, Zhe, et al.
Published: (2026)
The Learnability Gap in Medical Latent Diffusion
by: Dombrowski, Mischa, et al.
Published: (2026)
by: Dombrowski, Mischa, et al.
Published: (2026)
Neural Sign Actors: A diffusion model for 3D sign language production from text
by: Baltatzis, Vasileios, et al.
Published: (2023)
by: Baltatzis, Vasileios, et al.
Published: (2023)
Do You See What I Am Pointing At? Gesture-Based Egocentric Video Question Answering
by: Choi, Yura, et al.
Published: (2026)
by: Choi, Yura, et al.
Published: (2026)
MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion
by: Chen, Xuyang, et al.
Published: (2025)
by: Chen, Xuyang, et al.
Published: (2025)
Generate to Ground: Multimodal Text Conditioning Boosts Phrase Grounding in Medical Vision-Language Models
by: Nützel, Felix, et al.
Published: (2025)
by: Nützel, Felix, et al.
Published: (2025)
Leveraging Multi-Modal Information to Enhance Dataset Distillation
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Label-free Motion-Conditioned Diffusion Model for Cardiac Ultrasound Synthesis
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
by: Zhu, Ruishu, et al.
Published: (2025)
by: Zhu, Ruishu, et al.
Published: (2025)
Image Distillation for Safe Data Sharing in Histopathology
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
Context-self contrastive pretraining for crop type semantic segmentation
by: Tarasiou, Michail, et al.
Published: (2021)
by: Tarasiou, Michail, et al.
Published: (2021)
LCMem: A Universal Model for Robust Image Memorization Detection
by: Dombrowski, Mischa, et al.
Published: (2025)
by: Dombrowski, Mischa, et al.
Published: (2025)
Similar Items
-
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2023) -
Arc2Face: A Foundation Model for ID-Consistent Human Faces
by: Papantoniou, Foivos Paraperas, et al.
Published: (2024) -
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025) -
AnimateMe: 4D Facial Expressions via Diffusion Models
by: Gerogiannis, Dimitrios, et al.
Published: (2024) -
DermaFlux: Synthetic Skin Lesion Generation with Rectified Flows for Enhanced Image Classification
by: Galanakis, Stathis, et al.
Published: (2026)