HART: Human Aligned Reconstruction Transformer
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Xiyi, Wang, Shaofei, Mihajlovic, Marko, Kang, Taewon, Prokudin, Sergey, Lin, Ming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Morphable Diffusion: 3D-Consistent Diffusion for Single-image Avatar Creation
por: Chen, Xiyi, et al.
Publicado: (2024)
por: Chen, Xiyi, et al.
Publicado: (2024)
SplatFormer: Point Transformer for Robust 3D Gaussian Splatting
por: Chen, Yutong, et al.
Publicado: (2024)
por: Chen, Yutong, et al.
Publicado: (2024)
ResFields: Residual Neural Fields for Spatiotemporal Signals
por: Mihajlovic, Marko, et al.
Publicado: (2023)
por: Mihajlovic, Marko, et al.
Publicado: (2023)
Neural Texture Splatting: Expressive 3D Gaussian Splatting for View Synthesis, Geometry, and Dynamic Reconstruction
por: Wang, Yiming, et al.
Publicado: (2025)
por: Wang, Yiming, et al.
Publicado: (2025)
RISE-SDF: a Relightable Information-Shared Signed Distance Field for Glossy Object Inverse Rendering
por: Zhang, Deheng, et al.
Publicado: (2024)
por: Zhang, Deheng, et al.
Publicado: (2024)
SplatFields: Neural Gaussian Splats for Sparse 3D and 4D Reconstruction
por: Mihajlovic, Marko, et al.
Publicado: (2024)
por: Mihajlovic, Marko, et al.
Publicado: (2024)
Degrees of Freedom Matter: Inferring Dynamics from Point Trajectories
por: Zhang, Yan, et al.
Publicado: (2024)
por: Zhang, Yan, et al.
Publicado: (2024)
3DGS-Avatar: Animatable Avatars via Deformable 3D Gaussian Splatting
por: Qian, Zhiyin, et al.
Publicado: (2023)
por: Qian, Zhiyin, et al.
Publicado: (2023)
Character-Centered Dialogue Generation from Scene-Level Prompts
por: Kang, Taewon, et al.
Publicado: (2025)
por: Kang, Taewon, et al.
Publicado: (2025)
NEGATE: Constrained Semantic Guidance for Linguistic Negation in Text-to-Video Diffusion
por: Kang, Taewon, et al.
Publicado: (2026)
por: Kang, Taewon, et al.
Publicado: (2026)
Multi-View 3D Point Tracking
por: Rajič, Frano, et al.
Publicado: (2025)
por: Rajič, Frano, et al.
Publicado: (2025)
GGPT: Geometry Grounded Point Transformer
por: Chen, Yutong, et al.
Publicado: (2026)
por: Chen, Yutong, et al.
Publicado: (2026)
Trajectory-Guided Diffusion for Foreground-Preserving Background Generation in Multi-Layer Documents
por: Kang, Taewon
Publicado: (2026)
por: Kang, Taewon
Publicado: (2026)
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling
por: Kang, Taewon, et al.
Publicado: (2025)
por: Kang, Taewon, et al.
Publicado: (2025)
3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance
por: Kang, Taewon, et al.
Publicado: (2024)
por: Kang, Taewon, et al.
Publicado: (2024)
DCR: Counterfactual Attractor Guidance for Rare Compositional Generation
por: Kang, Taewon, et al.
Publicado: (2026)
por: Kang, Taewon, et al.
Publicado: (2026)
SonoWorld: From One Image to a 3D Audio-Visual Scene
por: Jin, Derong, et al.
Publicado: (2026)
por: Jin, Derong, et al.
Publicado: (2026)
SHARE: Scene-Human Aligned Reconstruction
por: Li, Joshua, et al.
Publicado: (2025)
por: Li, Joshua, et al.
Publicado: (2025)
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
por: Tang, Haotian, et al.
Publicado: (2024)
por: Tang, Haotian, et al.
Publicado: (2024)
Spline Deformation Field
por: Song, Mingyang, et al.
Publicado: (2025)
por: Song, Mingyang, et al.
Publicado: (2025)
Towards Transformer-Based Aligned Generation with Self-Coherence Guidance
por: Wang, Shulei, et al.
Publicado: (2025)
por: Wang, Shulei, et al.
Publicado: (2025)
HDRTransDC: High Dynamic Range Image Reconstruction with Transformer Deformation Convolution
por: Shang, Shuaikang, et al.
Publicado: (2024)
por: Shang, Shuaikang, et al.
Publicado: (2024)
HALO: Human-Aligned End-to-end Image Retargeting with Layered Transformations
por: Xu, Yiran, et al.
Publicado: (2025)
por: Xu, Yiran, et al.
Publicado: (2025)
STMR: Spiral Transformer for Hand Mesh Reconstruction
por: Xie, Huilong, et al.
Publicado: (2024)
por: Xie, Huilong, et al.
Publicado: (2024)
Local Patches Meet Global Context: Scalable 3D Diffusion Priors for Computed Tomography Reconstruction
por: Yang, Taewon, et al.
Publicado: (2025)
por: Yang, Taewon, et al.
Publicado: (2025)
AAformer: Auto-Aligned Transformer for Person Re-Identification
por: Zhu, Kuan, et al.
Publicado: (2021)
por: Zhu, Kuan, et al.
Publicado: (2021)
VolumetricSMPL: A Neural Volumetric Body Model for Efficient Interactions, Contacts, and Collisions
por: Mihajlovic, Marko, et al.
Publicado: (2025)
por: Mihajlovic, Marko, et al.
Publicado: (2025)
IntegratedPIFu: Integrated Pixel Aligned Implicit Function for Single-view Human Reconstruction
por: Chan, Kennard Yanting, et al.
Publicado: (2022)
por: Chan, Kennard Yanting, et al.
Publicado: (2022)
B-Cos Aligned Transformers Learn Human-Interpretable Features
por: Tran, Manuel, et al.
Publicado: (2024)
por: Tran, Manuel, et al.
Publicado: (2024)
IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray Tracing
por: Wang, Shaofei, et al.
Publicado: (2023)
por: Wang, Shaofei, et al.
Publicado: (2023)
Human Hair Reconstruction with Strand-Aligned 3D Gaussians
por: Zakharov, Egor, et al.
Publicado: (2024)
por: Zakharov, Egor, et al.
Publicado: (2024)
Text-Conditioned Background Generation for Editable Multi-Layer Documents
por: Kang, Taewon, et al.
Publicado: (2025)
por: Kang, Taewon, et al.
Publicado: (2025)
Index-Aligned Query Distillation for Transformer-based Incremental Object Detection
por: Ma, Mingxiao, et al.
Publicado: (2025)
por: Ma, Mingxiao, et al.
Publicado: (2025)
GRAFT: Geometric Refinement and Fitting Transformer for Human Scene Reconstruction
por: YM, Pradyumna, et al.
Publicado: (2026)
por: YM, Pradyumna, et al.
Publicado: (2026)
Video2BEV: Transforming Drone Videos to BEVs for Video-based Geo-localization
por: Ju, Hao, et al.
Publicado: (2024)
por: Ju, Hao, et al.
Publicado: (2024)
FIOVA: A Multi-Annotator Benchmark for Human-Aligned Video Captioning
por: Hu, Shiyu, et al.
Publicado: (2024)
por: Hu, Shiyu, et al.
Publicado: (2024)
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing
por: Li, Yan, et al.
Publicado: (2026)
por: Li, Yan, et al.
Publicado: (2026)
PointForward: Feedforward Driving Reconstruction through Point-Aligned Representations
por: Chi, Cheng, et al.
Publicado: (2026)
por: Chi, Cheng, et al.
Publicado: (2026)
DNF-Avatar: Distilling Neural Fields for Real-time Animatable Avatar Relighting
por: Jiang, Zeren, et al.
Publicado: (2025)
por: Jiang, Zeren, et al.
Publicado: (2025)
Aligning Human Motion Generation with Human Perceptions
por: Wang, Haoru, et al.
Publicado: (2024)
por: Wang, Haoru, et al.
Publicado: (2024)
Ejemplares similares
-
Morphable Diffusion: 3D-Consistent Diffusion for Single-image Avatar Creation
por: Chen, Xiyi, et al.
Publicado: (2024) -
SplatFormer: Point Transformer for Robust 3D Gaussian Splatting
por: Chen, Yutong, et al.
Publicado: (2024) -
ResFields: Residual Neural Fields for Spatiotemporal Signals
por: Mihajlovic, Marko, et al.
Publicado: (2023) -
Neural Texture Splatting: Expressive 3D Gaussian Splatting for View Synthesis, Geometry, and Dynamic Reconstruction
por: Wang, Yiming, et al.
Publicado: (2025) -
RISE-SDF: a Relightable Information-Shared Signed Distance Field for Glossy Object Inverse Rendering
por: Zhang, Deheng, et al.
Publicado: (2024)