Seeing World Dynamics in a Nutshell
Fuente:
arXiv
Guardado en:
| Autores principales: | Shen, Qiuhong, Yi, Xuanyu, Lin, Mingbao, Zhang, Hanwang, Yan, Shuicheng, Wang, Xinchao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
por: Shen, Qiuhong, et al.
Publicado: (2024)
por: Shen, Qiuhong, et al.
Publicado: (2024)
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
por: Xu, Xiangyu, et al.
Publicado: (2024)
por: Xu, Xiangyu, et al.
Publicado: (2024)
Instant3D: Instant Text-to-3D Generation
por: Li, Ming, et al.
Publicado: (2023)
por: Li, Ming, et al.
Publicado: (2023)
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
por: Dai, Yudi, et al.
Publicado: (2024)
por: Dai, Yudi, et al.
Publicado: (2024)
Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction
por: Shen, Qiuhong, et al.
Publicado: (2024)
por: Shen, Qiuhong, et al.
Publicado: (2024)
A Survey on 3D Gaussian Splatting
por: Chen, Guikun, et al.
Publicado: (2024)
por: Chen, Guikun, et al.
Publicado: (2024)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
por: Lyu, Tianle, et al.
Publicado: (2025)
por: Lyu, Tianle, et al.
Publicado: (2025)
Lester: rotoscope animation through video object segmentation and tracking
por: Tous, Ruben
Publicado: (2024)
por: Tous, Ruben
Publicado: (2024)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
por: Sar, Ayan, et al.
Publicado: (2025)
por: Sar, Ayan, et al.
Publicado: (2025)
Extreme Compression of Adaptive Neural Images
por: Hoshikawa, Leo, et al.
Publicado: (2024)
por: Hoshikawa, Leo, et al.
Publicado: (2024)
Freehand Sketch Generation from Mechanical Components
por: Liao, Zhichao, et al.
Publicado: (2024)
por: Liao, Zhichao, et al.
Publicado: (2024)
ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
por: Wang, Xuanchen, et al.
Publicado: (2025)
por: Wang, Xuanchen, et al.
Publicado: (2025)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
por: Sun, Zeyi, et al.
Publicado: (2024)
por: Sun, Zeyi, et al.
Publicado: (2024)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
por: Guo, Xin, et al.
Publicado: (2025)
por: Guo, Xin, et al.
Publicado: (2025)
ArchGPT: Understanding the World's Architectures with Large Multimodal Models
por: Wang, Yuze, et al.
Publicado: (2025)
por: Wang, Yuze, et al.
Publicado: (2025)
Vista3D: Unravel the 3D Darkside of a Single Image
por: Shen, Qiuhong, et al.
Publicado: (2024)
por: Shen, Qiuhong, et al.
Publicado: (2024)
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models
por: S, Sridhar, et al.
Publicado: (2025)
por: S, Sridhar, et al.
Publicado: (2025)
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
por: Zhang, Zewei, et al.
Publicado: (2024)
por: Zhang, Zewei, et al.
Publicado: (2024)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
por: Liu, Ziyuan, et al.
Publicado: (2026)
por: Liu, Ziyuan, et al.
Publicado: (2026)
Laplacian Analysis Meets Dynamics Modelling: Gaussian Splatting for 4D Reconstruction
por: Zhou, Yifan, et al.
Publicado: (2025)
por: Zhou, Yifan, et al.
Publicado: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
por: Park, Inkyu, et al.
Publicado: (2023)
por: Park, Inkyu, et al.
Publicado: (2023)
ToonAging: Face Re-Aging upon Artistic Portrait Style Transfer
por: Kim, Bumsoo, et al.
Publicado: (2024)
por: Kim, Bumsoo, et al.
Publicado: (2024)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
por: Singer, Assaf, et al.
Publicado: (2025)
por: Singer, Assaf, et al.
Publicado: (2025)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
por: Kim, Bumsoo, et al.
Publicado: (2024)
por: Kim, Bumsoo, et al.
Publicado: (2024)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
por: Chiu, Pin-Yen, et al.
Publicado: (2025)
por: Chiu, Pin-Yen, et al.
Publicado: (2025)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
por: Girdhar, Rohit, et al.
Publicado: (2023)
por: Girdhar, Rohit, et al.
Publicado: (2023)
Instruction-Driven 3D Facial Expression Generation and Transition
por: Vo, Anh H., et al.
Publicado: (2026)
por: Vo, Anh H., et al.
Publicado: (2026)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
por: Bilecen, Bahri Batuhan, et al.
Publicado: (2024)
por: Bilecen, Bahri Batuhan, et al.
Publicado: (2024)
Cross-Scenario Deraining Adaptation with Unpaired Data: Superpixel Structural Priors and Multi-Stage Pseudo-Rain Synthesis
por: Zhao, Kangbo, et al.
Publicado: (2026)
por: Zhao, Kangbo, et al.
Publicado: (2026)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
por: Zheng, Shuhong, et al.
Publicado: (2026)
por: Zheng, Shuhong, et al.
Publicado: (2026)
Poison-splat: Computation Cost Attack on 3D Gaussian Splatting
por: Lu, Jiahao, et al.
Publicado: (2024)
por: Lu, Jiahao, et al.
Publicado: (2024)
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
por: Xu, Chuanzhi, et al.
Publicado: (2026)
por: Xu, Chuanzhi, et al.
Publicado: (2026)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
por: Lin, Jiantao, et al.
Publicado: (2025)
por: Lin, Jiantao, et al.
Publicado: (2025)
Casual3DHDR: Deblurring High Dynamic Range 3D Gaussian Splatting from Casually Captured Videos
por: Gong, Shucheng, et al.
Publicado: (2025)
por: Gong, Shucheng, et al.
Publicado: (2025)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
por: Rakesh, Vineet Kumar, et al.
Publicado: (2025)
por: Rakesh, Vineet Kumar, et al.
Publicado: (2025)
Photoreal Scene Reconstruction from an Egocentric Device
por: Lv, Zhaoyang, et al.
Publicado: (2025)
por: Lv, Zhaoyang, et al.
Publicado: (2025)
Coral Model Generation from Single Images for Virtual Reality Applications
por: Fu, Jie, et al.
Publicado: (2024)
por: Fu, Jie, et al.
Publicado: (2024)
SVGS: Enhancing Gaussian Splatting Using Primitives with Spatially Varying Colors
por: Xu, Rui, et al.
Publicado: (2024)
por: Xu, Rui, et al.
Publicado: (2024)
Perceive-Sample-Compress: Towards Real-Time 3D Gaussian Splatting
por: Wang, Zijian, et al.
Publicado: (2025)
por: Wang, Zijian, et al.
Publicado: (2025)
MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching
por: Xie, Shuzhao, et al.
Publicado: (2026)
por: Xie, Shuzhao, et al.
Publicado: (2026)
Ejemplares similares
-
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
por: Shen, Qiuhong, et al.
Publicado: (2024) -
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
por: Xu, Xiangyu, et al.
Publicado: (2024) -
Instant3D: Instant Text-to-3D Generation
por: Li, Ming, et al.
Publicado: (2023) -
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
por: Dai, Yudi, et al.
Publicado: (2024) -
Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction
por: Shen, Qiuhong, et al.
Publicado: (2024)