Archon: A Unified Multimodal Model for Holistic Digital Human Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Bao, Chong, Liu, Shichen, Yu, Lijun, Futschik, David, Moschoglou, Stylianos, Srivastava, Shefali, Bai, Ziqian, Tan, Feitong, Zhang, Guofeng, Cui, Zhaopeng, Fanello, Sean, Zhang, Yinda |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular Videos
por: Li, Yuan, et al.
Publicado: (2025)
por: Li, Yuan, et al.
Publicado: (2025)
Efficient 3D Implicit Head Avatar with Mesh-anchored Hash Table Blendshapes
por: Bai, Ziqian, et al.
Publicado: (2024)
por: Bai, Ziqian, et al.
Publicado: (2024)
SVG: 3D Stereoscopic Video Generation via Denoising Frame Matrix
por: Dai, Peng, et al.
Publicado: (2024)
por: Dai, Peng, et al.
Publicado: (2024)
S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix
por: Dai, Peng, et al.
Publicado: (2025)
por: Dai, Peng, et al.
Publicado: (2025)
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
por: Yu, Zhixuan, et al.
Publicado: (2024)
por: Yu, Zhixuan, et al.
Publicado: (2024)
LightAvatar: Efficient Head Avatar as Dynamic Neural Light Field
por: Wang, Huan, et al.
Publicado: (2024)
por: Wang, Huan, et al.
Publicado: (2024)
Talking Together: Synthesizing Co-Located 3D Conversations from Audio
por: Shan, Mengyi, et al.
Publicado: (2026)
por: Shan, Mengyi, et al.
Publicado: (2026)
GeneAvatar: Generic Expression-Aware Volumetric Head Avatar Editing from a Single Image
por: Bao, Chong, et al.
Publicado: (2024)
por: Bao, Chong, et al.
Publicado: (2024)
LightCity: An Urban Dataset for Outdoor Inverse Rendering and Reconstruction under Multi-illumination Conditions
por: Wang, Jingjing, et al.
Publicado: (2026)
por: Wang, Jingjing, et al.
Publicado: (2026)
Free360: Layered Gaussian Splatting for Unbounded 360-Degree View Synthesis from Extremely Sparse and Unposed Views
por: Bao, Chong, et al.
Publicado: (2025)
por: Bao, Chong, et al.
Publicado: (2025)
AtlasGS: Atlanta-world Guided Surface Reconstruction with Implicit Structured Gaussians
por: Zhang, Xiyu, et al.
Publicado: (2025)
por: Zhang, Xiyu, et al.
Publicado: (2025)
GO-NeRF: Generating Objects in Neural Radiance Fields for Virtual Reality Content Creation
por: Dai, Peng, et al.
Publicado: (2024)
por: Dai, Peng, et al.
Publicado: (2024)
Improving face generation quality and prompt following with synthetic captions
por: Tarasiou, Michail, et al.
Publicado: (2024)
por: Tarasiou, Michail, et al.
Publicado: (2024)
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
por: Galanakis, Stathis, et al.
Publicado: (2023)
por: Galanakis, Stathis, et al.
Publicado: (2023)
CHOSEN: Contrastive Hypothesis Selection for Multi-View Depth Refinement
por: Qiu, Di, et al.
Publicado: (2024)
por: Qiu, Di, et al.
Publicado: (2024)
SpinMeRound: Consistent Multi-View Identity Generation Using Diffusion Models
por: Galanakis, Stathis, et al.
Publicado: (2025)
por: Galanakis, Stathis, et al.
Publicado: (2025)
PATS: Patch Area Transportation with Subdivision for Local Feature Matching
por: Ni, Junjie, et al.
Publicado: (2023)
por: Ni, Junjie, et al.
Publicado: (2023)
Archon: An Architecture Search Framework for Inference-Time Techniques
por: Saad-Falcon, Jon, et al.
Publicado: (2024)
por: Saad-Falcon, Jon, et al.
Publicado: (2024)
UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation
por: Li, Yi, et al.
Publicado: (2025)
por: Li, Yi, et al.
Publicado: (2025)
SplatLoc: 3D Gaussian Splatting-based Visual Localization for Augmented Reality
por: Zhai, Hongjia, et al.
Publicado: (2024)
por: Zhai, Hongjia, et al.
Publicado: (2024)
GaussianPrediction: Dynamic 3D Gaussian Prediction for Motion Extrapolation and Free View Synthesis
por: Zhao, Boming, et al.
Publicado: (2024)
por: Zhao, Boming, et al.
Publicado: (2024)
AnimateMe: 4D Facial Expressions via Diffusion Models
por: Gerogiannis, Dimitrios, et al.
Publicado: (2024)
por: Gerogiannis, Dimitrios, et al.
Publicado: (2024)
Random Reward Phase-Type Distributions with Applications in Latent Severity Modeling
por: Pauli, Simon, et al.
Publicado: (2026)
por: Pauli, Simon, et al.
Publicado: (2026)
CG-SLAM: Efficient Dense RGB-D SLAM in a Consistent Uncertainty-aware 3D Gaussian Field
por: Hu, Jiarui, et al.
Publicado: (2024)
por: Hu, Jiarui, et al.
Publicado: (2024)
BlinkFlow: A Dataset to Push the Limits of Event-based Optical Flow Estimation
por: Li, Yijin, et al.
Publicado: (2023)
por: Li, Yijin, et al.
Publicado: (2023)
D$^3$FlowSLAM: Self-Supervised Dynamic SLAM with Flow Motion Decomposition and DINO Guidance
por: Yu, Xingyuan, et al.
Publicado: (2022)
por: Yu, Xingyuan, et al.
Publicado: (2022)
NeuraLoc: Visual Localization in Neural Implicit Map with Dual Complementary Features
por: Zhai, Hongjia, et al.
Publicado: (2025)
por: Zhai, Hongjia, et al.
Publicado: (2025)
BlinkTrack: Feature Tracking over 80 FPS via Events and Images
por: Shen, Yichen, et al.
Publicado: (2024)
por: Shen, Yichen, et al.
Publicado: (2024)
EVER: Exact Volumetric Ellipsoid Rendering for Real-time View Synthesis
por: Mai, Alexander, et al.
Publicado: (2024)
por: Mai, Alexander, et al.
Publicado: (2024)
Arc2Face: A Foundation Model for ID-Consistent Human Faces
por: Papantoniou, Foivos Paraperas, et al.
Publicado: (2024)
por: Papantoniou, Foivos Paraperas, et al.
Publicado: (2024)
Physical Simulator In-the-Loop Video Generation
por: Foo, Lin Geng, et al.
Publicado: (2026)
por: Foo, Lin Geng, et al.
Publicado: (2026)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
por: Dong, Yitong, et al.
Publicado: (2024)
por: Dong, Yitong, et al.
Publicado: (2024)
Circular supply chains in manufacturing—Quo vadis? Accomplishments, challenges and future opportunities
por: Arijit Bhattacharya, et al.
Publicado: (2024)
por: Arijit Bhattacharya, et al.
Publicado: (2024)
MoManifold: Learning to Measure 3D Human Motion via Decoupled Joint Acceleration Manifolds
por: Dang, Ziqiang, et al.
Publicado: (2024)
por: Dang, Ziqiang, et al.
Publicado: (2024)
Text To 3D Object Generation For Scalable Room Assembly
por: Laguna, Sonia, et al.
Publicado: (2025)
por: Laguna, Sonia, et al.
Publicado: (2025)
ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation
por: Fu, Haoyu, et al.
Publicado: (2025)
por: Fu, Haoyu, et al.
Publicado: (2025)
Investigating the Interaction of Digital Capabilities, Sustainable Practices, Product Quality, and Customer Satisfaction in Perishable Food Supply Chains
por: Lakshmi Shetty, et al.
Publicado: (2026)
por: Lakshmi Shetty, et al.
Publicado: (2026)
StructuReiser: A Structure-preserving Video Stylization Method
por: Spetlik, Radim, et al.
Publicado: (2024)
por: Spetlik, Radim, et al.
Publicado: (2024)
Adaptive Multiple Comparisons With the Best
por: Haoyu Chen, et al.
Publicado: (2024)
por: Haoyu Chen, et al.
Publicado: (2024)
CUS-GS: A Compact Unified Structured Gaussian Splatting Framework for Multimodal Scene Representation
por: Ming, Yuhang, et al.
Publicado: (2025)
por: Ming, Yuhang, et al.
Publicado: (2025)
Ejemplares similares
-
IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular Videos
por: Li, Yuan, et al.
Publicado: (2025) -
Efficient 3D Implicit Head Avatar with Mesh-anchored Hash Table Blendshapes
por: Bai, Ziqian, et al.
Publicado: (2024) -
SVG: 3D Stereoscopic Video Generation via Denoising Frame Matrix
por: Dai, Peng, et al.
Publicado: (2024) -
S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix
por: Dai, Peng, et al.
Publicado: (2025) -
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
por: Yu, Zhixuan, et al.
Publicado: (2024)