Saved in:
| Main Authors: | Yang, Fan, Li, Heyuan, Li, Peihao, Yuan, Weihao, Qiu, Lingteng, Song, Chaoyue, Chen, Cheng, He, Yisheng, Zhang, Shifeng, Han, Xiaoguang, Hoi, Steven, Lin, Guosheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.07720 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViSA-Enhanced Aerial VLN: A Visual-Spatial Reasoning Enhanced Framework for Aerial Vision-Language Navigation
by: Tong, Haoyu, et al.
Published: (2026)
by: Tong, Haoyu, et al.
Published: (2026)
MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction
by: He, Yisheng, et al.
Published: (2026)
by: He, Yisheng, et al.
Published: (2026)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
by: Nakamura, Issa, et al.
Published: (2026)
by: Nakamura, Issa, et al.
Published: (2026)
ViSA-Flow: Accelerating Robot Skill Learning via Large-Scale Video Semantic Action Flow
by: Chen, Changhe, et al.
Published: (2025)
by: Chen, Changhe, et al.
Published: (2025)
HyPlaneHead: Rethinking Tri-plane-like Representations in Full-Head Image Synthesis
by: Li, Heyuan, et al.
Published: (2025)
by: Li, Heyuan, et al.
Published: (2025)
OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
by: Gan, Qijun, et al.
Published: (2025)
by: Gan, Qijun, et al.
Published: (2025)
Condition Matters in Full-head 3D GANs
by: Li, Heyuan, et al.
Published: (2026)
by: Li, Heyuan, et al.
Published: (2026)
PartNerFace: Part-based Neural Radiance Fields for Animatable Facial Avatar Reconstruction
by: Yu, Xianggang, et al.
Published: (2026)
by: Yu, Xianggang, et al.
Published: (2026)
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length
by: Huang, Yubo, et al.
Published: (2025)
by: Huang, Yubo, et al.
Published: (2025)
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
OMG-Avatar: One-shot Multi-LOD Gaussian Head Avatar
by: Ren, Jianqiang, et al.
Published: (2026)
by: Ren, Jianqiang, et al.
Published: (2026)
Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos
by: Luo, Xianrui, et al.
Published: (2025)
by: Luo, Xianrui, et al.
Published: (2025)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
by: Qiu, Lingteng, et al.
Published: (2024)
by: Qiu, Lingteng, et al.
Published: (2024)
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
by: Wu, Yushuang, et al.
Published: (2024)
by: Wu, Yushuang, et al.
Published: (2024)
GUAVA: Generalizable Upper Body 3D Gaussian Avatar
by: Zhang, Dongbin, et al.
Published: (2025)
by: Zhang, Dongbin, et al.
Published: (2025)
SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Instance Segmentation
by: Xu, Mutian, et al.
Published: (2023)
by: Xu, Mutian, et al.
Published: (2023)
ViMo: Generating Motions from Casual Videos
by: Qiu, Liangdong, et al.
Published: (2024)
by: Qiu, Liangdong, et al.
Published: (2024)
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
by: Han, Xiaoguang, et al.
Published: (2024)
by: Han, Xiaoguang, et al.
Published: (2024)
SphereHead: Stable 3D Full-head Synthesis with Spherical Tri-plane Representation
by: Li, Heyuan, et al.
Published: (2024)
by: Li, Heyuan, et al.
Published: (2024)
REACTO: Reconstructing Articulated Objects from a Single Video
by: Song, Chaoyue, et al.
Published: (2024)
by: Song, Chaoyue, et al.
Published: (2024)
HRM^2Avatar: High-Fidelity Real-Time Mobile Avatars from Monocular Phone Scans
by: Shi, Chao, et al.
Published: (2025)
by: Shi, Chao, et al.
Published: (2025)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
AvatarTex: High-Fidelity Facial Texture Reconstruction from Single-Image Stylized Avatars
by: Qiu, Yuda, et al.
Published: (2025)
by: Qiu, Yuda, et al.
Published: (2025)
BecomingLit: Relightable Gaussian Avatars with Hybrid Neural Shading
by: Schmidt, Jonathan, et al.
Published: (2025)
by: Schmidt, Jonathan, et al.
Published: (2025)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
by: He, Yisheng, et al.
Published: (2025)
by: He, Yisheng, et al.
Published: (2025)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
by: Li, Peng, et al.
Published: (2025)
by: Li, Peng, et al.
Published: (2025)
Democratizing the Creation of Animatable Facial Avatars
by: Zhu, Yilin, et al.
Published: (2024)
by: Zhu, Yilin, et al.
Published: (2024)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
by: Zhu, Shenhao, et al.
Published: (2024)
by: Zhu, Shenhao, et al.
Published: (2024)
A Novel Access Control and Privacy-Enhancing Approach for Models in Edge Computing
by: Li, Peihao
Published: (2024)
by: Li, Peihao
Published: (2024)
LinkXplore: A Framework for Affordable High-Quality Blockchain Data
by: Li, Peihao
Published: (2025)
by: Li, Peihao
Published: (2025)
RDPO: Real Data Preference Optimization for Physics Consistency Video Generation
by: Qian, Wenxu, et al.
Published: (2025)
by: Qian, Wenxu, et al.
Published: (2025)
Large Depth Completion Model from Sparse Observations
by: Yu, Zhu, et al.
Published: (2026)
by: Yu, Zhu, et al.
Published: (2026)
DreamDissector: Learning Disentangled Text-to-3D Generation from 2D Diffusion Priors
by: Yan, Zizheng, et al.
Published: (2024)
by: Yan, Zizheng, et al.
Published: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024)
by: Ye, Chongjie, et al.
Published: (2024)
SplattingAvatar: Realistic Real-Time Human Avatars with Mesh-Embedded Gaussian Splatting
by: Shao, Zhijing, et al.
Published: (2024)
by: Shao, Zhijing, et al.
Published: (2024)
MoDA: Modeling Deformable 3D Objects from Casual Videos
by: Song, Chaoyue, et al.
Published: (2023)
by: Song, Chaoyue, et al.
Published: (2023)
ReImagine: Rethinking Controllable High-Quality Human Video Generation via Image-First Synthesis
by: Sun, Zhengwentai, et al.
Published: (2026)
by: Sun, Zhengwentai, et al.
Published: (2026)
Similar Items
-
ViSA-Enhanced Aerial VLN: A Visual-Spatial Reasoning Enhanced Framework for Aerial Vision-Language Navigation
by: Tong, Haoyu, et al.
Published: (2026) -
MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction
by: He, Yisheng, et al.
Published: (2026) -
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
by: Qiu, Lingteng, et al.
Published: (2025) -
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
by: Nakamura, Issa, et al.
Published: (2026) -
ViSA-Flow: Accelerating Robot Skill Learning via Large-Scale Video Semantic Action Flow
by: Chen, Changhe, et al.
Published: (2025)