SIGMAN:Scaling 3D Human Gaussian Generation with Millions of Assets
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Yuhang, Liu, Fengqi, Lu, Yixing, Zhao, Qin, Wu, Pingyu, Zhai, Wei, Yi, Ran, Cao, Yang, Ma, Lizhuang, Zha, Zheng-Jun, Dong, Junting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HERO: Human Reaction Generation from Videos
di: Yu, Chengjun, et al.
Pubblicazione: (2025)
di: Yu, Chengjun, et al.
Pubblicazione: (2025)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
di: Yang, Yuhang, et al.
Pubblicazione: (2023)
di: Yang, Yuhang, et al.
Pubblicazione: (2023)
AdR-Gaussian: Accelerating Gaussian Splatting with Adaptive Radius
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
di: Yang, Yuhang, et al.
Pubblicazione: (2024)
di: Yang, Yuhang, et al.
Pubblicazione: (2024)
Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data
di: Fan, Ke, et al.
Pubblicazione: (2025)
di: Fan, Ke, et al.
Pubblicazione: (2025)
End-to-End Spatial-Temporal Transformer for Real-time 4D HOI Reconstruction
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
di: Han, Guangyi, et al.
Pubblicazione: (2025)
di: Han, Guangyi, et al.
Pubblicazione: (2025)
Improved Video VAE for Latent Video Diffusion Model
di: Wu, Pingyu, et al.
Pubblicazione: (2024)
di: Wu, Pingyu, et al.
Pubblicazione: (2024)
Scaling Large Motion Models with Million-Level Human Motions
di: Wang, Ye, et al.
Pubblicazione: (2024)
di: Wang, Ye, et al.
Pubblicazione: (2024)
GRACE: Estimating Geometry-level 3D Human-Scene Contact from 2D Images
di: Wang, Chengfeng, et al.
Pubblicazione: (2025)
di: Wang, Chengfeng, et al.
Pubblicazione: (2025)
Continuous Piecewise-Affine Based Motion Model for Image Animation
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
From Frames to Sequences: Temporally Consistent Human-Centric Dense Prediction
di: Miao, Xingyu, et al.
Pubblicazione: (2026)
di: Miao, Xingyu, et al.
Pubblicazione: (2026)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
di: Shao, Yawen, et al.
Pubblicazione: (2024)
di: Shao, Yawen, et al.
Pubblicazione: (2024)
Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
di: Wu, Pingyu, et al.
Pubblicazione: (2025)
di: Wu, Pingyu, et al.
Pubblicazione: (2025)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
Grounding 3D Scene Affordance From Egocentric Interactions
di: Liu, Cuiyu, et al.
Pubblicazione: (2024)
di: Liu, Cuiyu, et al.
Pubblicazione: (2024)
EF-3DGS: Event-Aided Free-Trajectory 3D Gaussian Splatting
di: Liao, Bohao, et al.
Pubblicazione: (2024)
di: Liao, Bohao, et al.
Pubblicazione: (2024)
Gloria: Consistent Character Video Generation via Content Anchors
di: Yang, Yuhang, et al.
Pubblicazione: (2026)
di: Yang, Yuhang, et al.
Pubblicazione: (2026)
Event Stream Filtering via Probability Flux Estimation
di: Chen, Jinze, et al.
Pubblicazione: (2025)
di: Chen, Jinze, et al.
Pubblicazione: (2025)
Visual-Geometric Collaborative Guidance for Affordance Learning
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
Leverage Task Context for Object Affordance Ranking
di: Huang, Haojie, et al.
Pubblicazione: (2024)
di: Huang, Haojie, et al.
Pubblicazione: (2024)
GAS: Generative Avatar Synthesis from a Single Image
di: Lu, Yixing, et al.
Pubblicazione: (2025)
di: Lu, Yixing, et al.
Pubblicazione: (2025)
PointDGMamba: Domain Generalization of Point Cloud Classification via Generalized State Space Model
di: Yang, Hao, et al.
Pubblicazione: (2024)
di: Yang, Hao, et al.
Pubblicazione: (2024)
VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video Generation
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
3D Gaussian Head Avatars with Expressive Dynamic Appearances by Compact Tensorial Representations
di: Wang, Yating, et al.
Pubblicazione: (2025)
di: Wang, Yating, et al.
Pubblicazione: (2025)
\$OneMillion-Bench: How Far are Language Agents from Human Experts?
di: Yang, Qianyu, et al.
Pubblicazione: (2026)
di: Yang, Qianyu, et al.
Pubblicazione: (2026)
Fishbone: From One 3D Asset to a Million Controllable Edits
di: He, Yumeng, et al.
Pubblicazione: (2026)
di: He, Yumeng, et al.
Pubblicazione: (2026)
ArtGS:3D Gaussian Splatting for Interactive Visual-Physical Modeling and Manipulation of Articulated Objects
di: Yu, Qiaojun, et al.
Pubblicazione: (2025)
di: Yu, Qiaojun, et al.
Pubblicazione: (2025)
Unbiased Gradient Estimation for Event Binning via Functional Backpropagation
di: Chen, Jinze, et al.
Pubblicazione: (2026)
di: Chen, Jinze, et al.
Pubblicazione: (2026)
Event-based Visual Deformation Measurement
di: Wu, Yuliang, et al.
Pubblicazione: (2026)
di: Wu, Yuliang, et al.
Pubblicazione: (2026)
Event-based Asynchronous HDR Imaging by Temporal Incident Light Modulation
di: Wu, Yuliang, et al.
Pubblicazione: (2024)
di: Wu, Yuliang, et al.
Pubblicazione: (2024)
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
di: Yang, Jian, et al.
Pubblicazione: (2024)
di: Yang, Jian, et al.
Pubblicazione: (2024)
EMoTive: Event-guided Trajectory Modeling for 3D Motion Estimation
di: Wan, Zengyu, et al.
Pubblicazione: (2025)
di: Wan, Zengyu, et al.
Pubblicazione: (2025)
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation
di: Wang, Wenhao, et al.
Pubblicazione: (2024)
di: Wang, Wenhao, et al.
Pubblicazione: (2024)
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction
di: Yi, Ran, et al.
Pubblicazione: (2025)
di: Yi, Ran, et al.
Pubblicazione: (2025)
Gaussian Splatting with Discretized SDF for Relightable Assets
di: Zhu, Zuo-Liang, et al.
Pubblicazione: (2025)
di: Zhu, Zuo-Liang, et al.
Pubblicazione: (2025)
$\text{S}^{3}$Mamba: Arbitrary-Scale Super-Resolution via Scaleable State Space Model
di: Xia, Peizhe, et al.
Pubblicazione: (2024)
di: Xia, Peizhe, et al.
Pubblicazione: (2024)
MATE: Motion-Augmented Temporal Consistency for Event-based Point Tracking
di: Han, Han, et al.
Pubblicazione: (2024)
di: Han, Han, et al.
Pubblicazione: (2024)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
HERO: Human Reaction Generation from Videos
di: Yu, Chengjun, et al.
Pubblicazione: (2025) -
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
di: Yang, Yuhang, et al.
Pubblicazione: (2023) -
AdR-Gaussian: Accelerating Gaussian Splatting with Adaptive Radius
di: Wang, Xinzhe, et al.
Pubblicazione: (2024) -
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
di: Yang, Yuhang, et al.
Pubblicazione: (2024) -
Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data
di: Fan, Ke, et al.
Pubblicazione: (2025)