Mono4DEditor: Text-Driven 4D Scene Editing from Monocular Video via Point-Level Localization of Language-Embedded Gaussians
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Jin-Chuan, Su, Chengye, Wang, Jiajun, Shamir, Ariel, Wang, Miao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MonoDream: Monocular Vision-Language Navigation with Panoramic Dreaming
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
Global-Local Aware Scene Text Editing
von: Yang, Fuxiang, et al.
Veröffentlicht: (2025)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2025)
VINGS-Mono: Visual-Inertial Gaussian Splatting Monocular SLAM in Large Scenes
von: Wu, Ke, et al.
Veröffentlicht: (2025)
von: Wu, Ke, et al.
Veröffentlicht: (2025)
Mono4DGS-HDR: High Dynamic Range 4D Gaussian Splatting from Alternating-exposure Monocular Videos
von: Liu, Jinfeng, et al.
Veröffentlicht: (2025)
von: Liu, Jinfeng, et al.
Veröffentlicht: (2025)
PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting
von: Miao, Qiaowei, et al.
Veröffentlicht: (2024)
von: Miao, Qiaowei, et al.
Veröffentlicht: (2024)
Kinematics-Driven Gaussian Shape Deformation for Blurry Monocular Dynamic Scenes
von: Song, Yeon-Ji, et al.
Veröffentlicht: (2026)
von: Song, Yeon-Ji, et al.
Veröffentlicht: (2026)
LatentEditor: Text Driven Local Editing of 3D Scenes
von: Khalid, Umar, et al.
Veröffentlicht: (2023)
von: Khalid, Umar, et al.
Veröffentlicht: (2023)
MonoFusion: Sparse-View 4D Reconstruction via Monocular Fusion
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
SIMSplat: Predictive Driving Scene Editing with Language-aligned 4D Gaussian Splatting
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
MonoCloth: Reconstruction and Animation of Cloth-Decoupled Human Avatars from Monocular Videos
von: Jin, Daisheng, et al.
Veröffentlicht: (2025)
von: Jin, Daisheng, et al.
Veröffentlicht: (2025)
MonoGS++: Fast and Accurate Monocular RGB Gaussian SLAM
von: Li, Renwu, et al.
Veröffentlicht: (2025)
von: Li, Renwu, et al.
Veröffentlicht: (2025)
Endo-4DGS: Endoscopic Monocular Scene Reconstruction with 4D Gaussian Splatting
von: Huang, Yiming, et al.
Veröffentlicht: (2024)
von: Huang, Yiming, et al.
Veröffentlicht: (2024)
REED-VAE: RE-Encode Decode Training for Iterative Image Editing with Diffusion Models
von: Almog, Gal, et al.
Veröffentlicht: (2025)
von: Almog, Gal, et al.
Veröffentlicht: (2025)
AvatarBrush: Monocular Reconstruction of Gaussian Avatars with Intuitive Local Editing
von: Li, Mengtian, et al.
Veröffentlicht: (2025)
von: Li, Mengtian, et al.
Veröffentlicht: (2025)
MonoNPHM: Dynamic Head Reconstruction from Monocular Videos
von: Giebenhain, Simon, et al.
Veröffentlicht: (2023)
von: Giebenhain, Simon, et al.
Veröffentlicht: (2023)
Uncertainty Matters in Dynamic Gaussian Splatting for Monocular 4D Reconstruction
von: Guo, Fengzhi, et al.
Veröffentlicht: (2025)
von: Guo, Fengzhi, et al.
Veröffentlicht: (2025)
Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos
von: Yuan, Chengbo, et al.
Veröffentlicht: (2024)
von: Yuan, Chengbo, et al.
Veröffentlicht: (2024)
MonoMobility: Zero-Shot 3D Mobility Analysis from Monocular Videos
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
Flash-Mono: Feed-Forward Accelerated Gaussian Splatting Monocular SLAM
von: Zhang, Zicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2026)
MonoEM-GS: Monocular Expectation-Maximization Gaussian Splatting SLAM
von: Kruzhkov, Evgenii, et al.
Veröffentlicht: (2026)
von: Kruzhkov, Evgenii, et al.
Veröffentlicht: (2026)
A Survey on Monocular Re-Localization: From the Perspective of Scene Map Representation
von: Miao, Jinyu, et al.
Veröffentlicht: (2023)
von: Miao, Jinyu, et al.
Veröffentlicht: (2023)
DreamScene4D: Dynamic Multi-Object Scene Generation from Monocular Videos
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2024)
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2024)
FATE: Full-head Gaussian Avatar with Textural Editing from Monocular Video
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
The Scene Language: Representing Scenes with Programs, Words, and Embeddings
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
TextMastero: Mastering High-Quality Scene Text Editing in Diverse Languages and Styles
von: Wang, Tong, et al.
Veröffentlicht: (2024)
von: Wang, Tong, et al.
Veröffentlicht: (2024)
ProDyG: Progressive Dynamic Scene Reconstruction via Gaussian Splatting from Monocular Videos
von: Chen, Shi, et al.
Veröffentlicht: (2025)
von: Chen, Shi, et al.
Veröffentlicht: (2025)
Point'n Move: Interactive Scene Object Manipulation on Gaussian Splatting Radiance Fields
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
DragText: Rethinking Text Embedding in Point-based Image Editing
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
Deblur4DGS: 4D Gaussian Splatting from Blurry Monocular Video
von: Wu, Renlong, et al.
Veröffentlicht: (2024)
von: Wu, Renlong, et al.
Veröffentlicht: (2024)
AudioScenic: Audio-Driven Video Scene Editing
von: Shen, Kaixin, et al.
Veröffentlicht: (2024)
von: Shen, Kaixin, et al.
Veröffentlicht: (2024)
MonoRace: Winning Champion-Level Drone Racing with Robust Monocular AI
von: Bahnam, Stavrow A., et al.
Veröffentlicht: (2026)
von: Bahnam, Stavrow A., et al.
Veröffentlicht: (2026)
Feature4X: Bridging Any Monocular Video to 4D Agentic AI with Versatile Gaussian Feature Fields
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
EfficientMonoHair: Fast Strand-Level Reconstruction from Monocular Video via Multi-View Direction Fusion
von: Li, Da, et al.
Veröffentlicht: (2026)
von: Li, Da, et al.
Veröffentlicht: (2026)
MonoOcc: Digging into Monocular Semantic Occupancy Prediction
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
Sparse4DGS: 4D Gaussian Splatting for Sparse-Frame Dynamic Scene Reconstruction
von: Shi, Changyue, et al.
Veröffentlicht: (2025)
von: Shi, Changyue, et al.
Veröffentlicht: (2025)
MonoSLAM: Robust Monocular SLAM with Global Structure Optimization
von: Jiang, Bingzheng, et al.
Veröffentlicht: (2025)
von: Jiang, Bingzheng, et al.
Veröffentlicht: (2025)
MonoMSK: Monocular 3D Musculoskeletal Dynamics Estimation
von: Koleini, Farnoosh, et al.
Veröffentlicht: (2025)
von: Koleini, Farnoosh, et al.
Veröffentlicht: (2025)
MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos
von: Rho, Daniel, et al.
Veröffentlicht: (2026)
von: Rho, Daniel, et al.
Veröffentlicht: (2026)
MonoHair: High-Fidelity Hair Modeling from a Monocular Video
von: Wu, Keyu, et al.
Veröffentlicht: (2024)
von: Wu, Keyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MonoDream: Monocular Vision-Language Navigation with Panoramic Dreaming
von: Wang, Shuo, et al.
Veröffentlicht: (2025) -
Global-Local Aware Scene Text Editing
von: Yang, Fuxiang, et al.
Veröffentlicht: (2025) -
VINGS-Mono: Visual-Inertial Gaussian Splatting Monocular SLAM in Large Scenes
von: Wu, Ke, et al.
Veröffentlicht: (2025) -
Mono4DGS-HDR: High Dynamic Range 4D Gaussian Splatting from Alternating-exposure Monocular Videos
von: Liu, Jinfeng, et al.
Veröffentlicht: (2025) -
PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting
von: Miao, Qiaowei, et al.
Veröffentlicht: (2024)