Saved in:
| Main Authors: | Yu, Zhu, Zhao, Zhengyi, Zhang, Runmin, Qiu, Lingteng, Qiu, Kejie, He, Yisheng, Zhu, Siyu, Dong, Zilong, Cao, Si-Yuan, Shen, Hui-Liang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.30115 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
by: Hu, Yingdong, et al.
Published: (2025)
by: Hu, Yingdong, et al.
Published: (2025)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
by: Zhu, Shenhao, et al.
Published: (2024)
by: Zhu, Shenhao, et al.
Published: (2024)
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
by: Yu, Zhu, et al.
Published: (2024)
by: Yu, Zhu, et al.
Published: (2024)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
by: He, Yisheng, et al.
Published: (2025)
by: He, Yisheng, et al.
Published: (2025)
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
by: He, Yisheng, et al.
Published: (2024)
by: He, Yisheng, et al.
Published: (2024)
Structure-Aware Radar-Camera Depth Estimation
by: Zhang, Fuyi, et al.
Published: (2025)
by: Zhang, Fuyi, et al.
Published: (2025)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
by: Li, Peng, et al.
Published: (2025)
by: Li, Peng, et al.
Published: (2025)
Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction
by: Zhang, Runmin, et al.
Published: (2025)
by: Zhang, Runmin, et al.
Published: (2025)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency Constraint
by: Zhang, Runmin, et al.
Published: (2025)
by: Zhang, Runmin, et al.
Published: (2025)
HyPlaneHead: Rethinking Tri-plane-like Representations in Full-Head Image Synthesis
by: Li, Heyuan, et al.
Published: (2025)
by: Li, Heyuan, et al.
Published: (2025)
EDFFDNet: Towards Accurate and Efficient Unsupervised Multi-Grid Image Registration
by: Zhu, Haokai, et al.
Published: (2025)
by: Zhu, Haokai, et al.
Published: (2025)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
by: Qiu, Lingteng, et al.
Published: (2024)
by: Qiu, Lingteng, et al.
Published: (2024)
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
by: Wu, Yushuang, et al.
Published: (2024)
by: Wu, Yushuang, et al.
Published: (2024)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
by: Yuan, Weihao, et al.
Published: (2024)
by: Yuan, Weihao, et al.
Published: (2024)
SSHNet: Unsupervised Cross-modal Homography Estimation via Problem Reformulation and Split Optimization
by: Yu, Junchen, et al.
Published: (2024)
by: Yu, Junchen, et al.
Published: (2024)
Freeplane: Unlocking Free Lunch in Triplane-Based Sparse-View Reconstruction Models
by: Sun, Wenqiang, et al.
Published: (2024)
by: Sun, Wenqiang, et al.
Published: (2024)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
by: Han, Xiaoguang, et al.
Published: (2024)
by: Han, Xiaoguang, et al.
Published: (2024)
SGDFormer: One-stage Transformer-based Architecture for Cross-Spectral Stereo Image Guided Denoising
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024)
by: Ye, Chongjie, et al.
Published: (2024)
Position Engineering: Boosting Large Language Models through Positional Information Manipulation
by: He, Zhiyuan, et al.
Published: (2024)
by: He, Zhiyuan, et al.
Published: (2024)
Language Driven Occupancy Prediction
by: Yu, Zhu, et al.
Published: (2024)
by: Yu, Zhu, et al.
Published: (2024)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Hierarchical Federated Unlearning for Large Language Models
by: Zhong, Yisheng, et al.
Published: (2025)
by: Zhong, Yisheng, et al.
Published: (2025)
Head-wise Adaptive Rotary Positional Encoding for Fine-Grained Image Generation
by: Li, Jiaye, et al.
Published: (2025)
by: Li, Jiaye, et al.
Published: (2025)
GIC: Gaussian-Informed Continuum for Physical Property Identification and Simulation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
by: Zhang, Xue, et al.
Published: (2024)
by: Zhang, Xue, et al.
Published: (2024)
LLM-Metrics: Measuring Research Impact Through Large Language Model Memory
by: Shen, Si, et al.
Published: (2026)
by: Shen, Si, et al.
Published: (2026)
Application and Optimization of Large Models Based on Prompt Tuning for Fact-Check-Worthiness Estimation
by: Yu, Yinglong, et al.
Published: (2025)
by: Yu, Yinglong, et al.
Published: (2025)
Condition Matters in Full-head 3D GANs
by: Li, Heyuan, et al.
Published: (2026)
by: Li, Heyuan, et al.
Published: (2026)
DicFace: Dirichlet-Constrained Variational Codebook Learning for Temporally Coherent Video Face Restoration
by: Chen, Yan, et al.
Published: (2025)
by: Chen, Yan, et al.
Published: (2025)
The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents
by: Sun, Yuwei, et al.
Published: (2026)
by: Sun, Yuwei, et al.
Published: (2026)
S-BEVLoc: BEV-based Self-supervised Framework for Large-scale LiDAR Global Localization
by: Zhang, Chenghao, et al.
Published: (2025)
by: Zhang, Chenghao, et al.
Published: (2025)
Similar Items
-
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
by: Hu, Yingdong, et al.
Published: (2025) -
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024) -
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
by: Li, Zhe, et al.
Published: (2024) -
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025) -
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
by: Zhu, Shenhao, et al.
Published: (2024)