EG4D: Explicit Generation of 4D Object without Score Distillation
Fuente:
arXiv
Guardado en:
| Autores principales: | Sun, Qi, Guo, Zhiyang, Wan, Ziyu, Yan, Jing Nathan, Yin, Shengming, Zhou, Wengang, Liao, Jing, Li, Houqiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction
por: Guo, Zhiyang, et al.
Publicado: (2024)
por: Guo, Zhiyang, et al.
Publicado: (2024)
HeadArtist: Text-conditioned 3D Head Generation with Self Score Distillation
por: Liu, Hongyu, et al.
Publicado: (2023)
por: Liu, Hongyu, et al.
Publicado: (2023)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
por: Sun, Qi, et al.
Publicado: (2024)
por: Sun, Qi, et al.
Publicado: (2024)
Make-It-Animatable: An Efficient Framework for Authoring Animation-Ready 3D Characters
por: Guo, Zhiyang, et al.
Publicado: (2024)
por: Guo, Zhiyang, et al.
Publicado: (2024)
Make-It-Poseable: Feed-forward Latent Posing Model for 3D Characters
por: Guo, Zhiyang, et al.
Publicado: (2025)
por: Guo, Zhiyang, et al.
Publicado: (2025)
Animus3D: Text-driven 3D Animation via Motion Score Distillation
por: Sun, Qi, et al.
Publicado: (2025)
por: Sun, Qi, et al.
Publicado: (2025)
Optimizing Distributional Geometry Alignment with Optimal Transport for Generative Dataset Distillation
por: Cui, Xiao, et al.
Publicado: (2025)
por: Cui, Xiao, et al.
Publicado: (2025)
Self-Classification Enhancement and Correction for Weakly Supervised Object Detection
por: Yin, Yufei, et al.
Publicado: (2025)
por: Yin, Yufei, et al.
Publicado: (2025)
Recurrent Generic Contour-based Instance Segmentation with Progressive Learning
por: Feng, Hao, et al.
Publicado: (2023)
por: Feng, Hao, et al.
Publicado: (2023)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
por: Sun, Qi, et al.
Publicado: (2024)
por: Sun, Qi, et al.
Publicado: (2024)
Video-based Sign Language Recognition without Temporal Segmentation
por: Huang, Jie, et al.
Publicado: (2018)
por: Huang, Jie, et al.
Publicado: (2018)
Structural Action Transformer for 3D Dexterous Manipulation
por: Lei, Xiaohan, et al.
Publicado: (2026)
por: Lei, Xiaohan, et al.
Publicado: (2026)
Text2NeRF: Text-Driven 3D Scene Generation with Neural Radiance Fields
por: Zhang, Jingbo, et al.
Publicado: (2023)
por: Zhang, Jingbo, et al.
Publicado: (2023)
4D-fy: Text-to-4D Generation Using Hybrid Score Distillation Sampling
por: Bahmani, Sherwin, et al.
Publicado: (2023)
por: Bahmani, Sherwin, et al.
Publicado: (2023)
MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning
por: Liu, Xiaoyang, et al.
Publicado: (2024)
por: Liu, Xiaoyang, et al.
Publicado: (2024)
RoFIR: Robust Fisheye Image Rectification Framework Impervious to Optical Center Deviation
por: Liao, Zhaokang, et al.
Publicado: (2024)
por: Liao, Zhaokang, et al.
Publicado: (2024)
RaFE: Generative Radiance Fields Restoration
por: Wu, Zhongkai, et al.
Publicado: (2024)
por: Wu, Zhongkai, et al.
Publicado: (2024)
Rethinking Long-tailed Dataset Distillation: A Uni-Level Framework with Unbiased Recovery and Relabeling
por: Cui, Xiao, et al.
Publicado: (2025)
por: Cui, Xiao, et al.
Publicado: (2025)
Learning Generalizable Human Motion Generator with Reinforcement Learning
por: Mao, Yunyao, et al.
Publicado: (2024)
por: Mao, Yunyao, et al.
Publicado: (2024)
An Object is Worth 64x64 Pixels: Generating 3D Object via Image Diffusion
por: Yan, Xingguang, et al.
Publicado: (2024)
por: Yan, Xingguang, et al.
Publicado: (2024)
GaussNav: Gaussian Splatting for Visual Navigation
por: Lei, Xiaohan, et al.
Publicado: (2024)
por: Lei, Xiaohan, et al.
Publicado: (2024)
Revisiting Shadow Detection from a Vision-Language Perspective
por: Wang, Yonghui, et al.
Publicado: (2026)
por: Wang, Yonghui, et al.
Publicado: (2026)
AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding
por: Wang, Yonghui, et al.
Publicado: (2024)
por: Wang, Yonghui, et al.
Publicado: (2024)
StepVAR: Structure-Texture Guided Pruning for Visual Autoregressive Models
por: Liu, Keli, et al.
Publicado: (2026)
por: Liu, Keli, et al.
Publicado: (2026)
Dynamic Realms: 4D Content Analysis, Recovery and Generation with Geometric, Topological and Physical Priors
por: Dou, Zhiyang
Publicado: (2024)
por: Dou, Zhiyang
Publicado: (2024)
Hybrid Fourier Score Distillation for Efficient One Image to 3D Object Generation
por: Yang, Shuzhou, et al.
Publicado: (2024)
por: Yang, Shuzhou, et al.
Publicado: (2024)
Robust Multimodal Large Language Models Against Modality Conflict
por: Zhang, Zongmeng, et al.
Publicado: (2025)
por: Zhang, Zongmeng, et al.
Publicado: (2025)
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
por: Chu, Xiaomeng, et al.
Publicado: (2023)
por: Chu, Xiaomeng, et al.
Publicado: (2023)
Text-Animator: Controllable Visual Text Video Generation
por: Liu, Lin, et al.
Publicado: (2024)
por: Liu, Lin, et al.
Publicado: (2024)
SimDistill: Simulated Multi-modal Distillation for BEV 3D Object Detection
por: Zhao, Haimei, et al.
Publicado: (2023)
por: Zhao, Haimei, et al.
Publicado: (2023)
DeepEraser: Deep Iterative Context Mining for Generic Text Eraser
por: Feng, Hao, et al.
Publicado: (2024)
por: Feng, Hao, et al.
Publicado: (2024)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
por: Wang, Zhendong, et al.
Publicado: (2025)
por: Wang, Zhendong, et al.
Publicado: (2025)
Connecting Consistency Distillation to Score Distillation for Text-to-3D Generation
por: Li, Zongrui, et al.
Publicado: (2024)
por: Li, Zongrui, et al.
Publicado: (2024)
ScaleWeaver: Weaving Efficient Controllable T2I Generation with Multi-Scale Reference Attention
por: Liu, Keli, et al.
Publicado: (2025)
por: Liu, Keli, et al.
Publicado: (2025)
DynaPose4D: High-Quality 4D Dynamic Content Generation via Pose Alignment Loss
por: Yang, Jing, et al.
Publicado: (2025)
por: Yang, Jing, et al.
Publicado: (2025)
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection
por: Wang, Yonghui, et al.
Publicado: (2024)
por: Wang, Yonghui, et al.
Publicado: (2024)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
por: Zhao, Weichao, et al.
Publicado: (2024)
por: Zhao, Weichao, et al.
Publicado: (2024)
Exploiting Spatial-Temporal Context for Interacting Hand Reconstruction on Monocular RGB Video
por: Zhao, Weichao, et al.
Publicado: (2023)
por: Zhao, Weichao, et al.
Publicado: (2023)
Progressive Multi-modal Conditional Prompt Tuning
por: Qiu, Xiaoyu, et al.
Publicado: (2024)
por: Qiu, Xiaoyu, et al.
Publicado: (2024)
Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
por: Du, Yongchao, et al.
Publicado: (2024)
por: Du, Yongchao, et al.
Publicado: (2024)
Ejemplares similares
-
Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction
por: Guo, Zhiyang, et al.
Publicado: (2024) -
HeadArtist: Text-conditioned 3D Head Generation with Self Score Distillation
por: Liu, Hongyu, et al.
Publicado: (2023) -
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
por: Sun, Qi, et al.
Publicado: (2024) -
Make-It-Animatable: An Efficient Framework for Authoring Animation-Ready 3D Characters
por: Guo, Zhiyang, et al.
Publicado: (2024) -
Make-It-Poseable: Feed-forward Latent Posing Model for 3D Characters
por: Guo, Zhiyang, et al.
Publicado: (2025)