Geometry-Guided Reinforcement Learning for Multi-view Consistent 3D Scene Editing
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Jiyuan, Lin, Chunyu, Sun, Lei, Cao, Zhi, Yin, Yuyang, Nie, Lang, Yuan, Zhenlong, Chu, Xiangxiang, Wei, Yunchao, Liao, Kang, Lin, Guosheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Editor to Dense Geometry Estimator
di: Wang, JiYuan, et al.
Pubblicazione: (2025)
di: Wang, JiYuan, et al.
Pubblicazione: (2025)
360 Layout Estimation via Orthogonal Planes Disentanglement and Multi-view Geometric Consistency Perception
di: Shen, Zhijie, et al.
Pubblicazione: (2023)
di: Shen, Zhijie, et al.
Pubblicazione: (2023)
Digging into contrastive learning for robust depth estimation with diffusion models
di: Wang, Jiyuan, et al.
Pubblicazione: (2024)
di: Wang, Jiyuan, et al.
Pubblicazione: (2024)
Beyond Wide-Angle Images: Structure-to-Detail Video Portrait Correction via Unsupervised Spatiotemporal Adaptation
di: Nie, Wenbo, et al.
Pubblicazione: (2025)
di: Nie, Wenbo, et al.
Pubblicazione: (2025)
Advancing Real-World Parking Slot Detection with Large-Scale Dataset and Semi-Supervised Baseline
di: Zhang, Zhihao, et al.
Pubblicazione: (2025)
di: Zhang, Zhihao, et al.
Pubblicazione: (2025)
Semi-Supervised Coupled Thin-Plate Spline Model for Rotation Correction and Beyond
di: Nie, Lang, et al.
Pubblicazione: (2024)
di: Nie, Lang, et al.
Pubblicazione: (2024)
UniStitch: Unifying Semantic and Geometric Features for Image Stitching
di: Mei, Yuan, et al.
Pubblicazione: (2026)
di: Mei, Yuan, et al.
Pubblicazione: (2026)
Robust Image Stitching with Optimal Plane
di: Nie, Lang, et al.
Pubblicazione: (2025)
di: Nie, Lang, et al.
Pubblicazione: (2025)
Jasmine: Harnessing Diffusion Prior for Self-supervised Depth Estimation
di: Wang, Jiyuan, et al.
Pubblicazione: (2025)
di: Wang, Jiyuan, et al.
Pubblicazione: (2025)
SGFormer: Spherical Geometry Transformer for 360 Depth Estimation
di: Zhang, Junsong, et al.
Pubblicazione: (2024)
di: Zhang, Junsong, et al.
Pubblicazione: (2024)
Revisiting 360 Depth Estimation with PanoGabor: A New Fusion Perspective
di: Shen, Zhijie, et al.
Pubblicazione: (2024)
di: Shen, Zhijie, et al.
Pubblicazione: (2024)
StabStitch++: Unsupervised Online Video Stitching with Spatiotemporal Bidirectional Warps
di: Nie, Lang, et al.
Pubblicazione: (2025)
di: Nie, Lang, et al.
Pubblicazione: (2025)
Semi-Supervised 360 Layout Estimation with Panoramic Collaborative Perturbations
di: Zhang, Junsong, et al.
Pubblicazione: (2025)
di: Zhang, Junsong, et al.
Pubblicazione: (2025)
You Need a Transition Plane: Bridging Continuous Panoramic 3D Reconstruction with Perspective Gaussian Splatting
di: Shen, Zhijie, et al.
Pubblicazione: (2025)
di: Shen, Zhijie, et al.
Pubblicazione: (2025)
FLUX-Text: A Simple and Advanced Diffusion Transformer Baseline for Scene Text Editing
di: Lan, Rui, et al.
Pubblicazione: (2025)
di: Lan, Rui, et al.
Pubblicazione: (2025)
SceneLCM: End-to-End Layout-Guided Interactive Indoor Scene Generation with Latent Consistency Model
di: Lin, Yangkai, et al.
Pubblicazione: (2025)
di: Lin, Yangkai, et al.
Pubblicazione: (2025)
WeatherDepth: Curriculum Contrastive Learning for Self-Supervised Depth Estimation under Adverse Weather Conditions
di: Wang, Jiyuan, et al.
Pubblicazione: (2023)
di: Wang, Jiyuan, et al.
Pubblicazione: (2023)
DeCo: Decoupled Human-Centered Diffusion Video Editing with Motion Consistency
di: Zhong, Xiaojing, et al.
Pubblicazione: (2024)
di: Zhong, Xiaojing, et al.
Pubblicazione: (2024)
4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency
di: Yin, Yuyang, et al.
Pubblicazione: (2023)
di: Yin, Yuyang, et al.
Pubblicazione: (2023)
Eliminating Warping Shakes for Unsupervised Online Video Stitching
di: Nie, Lang, et al.
Pubblicazione: (2024)
di: Nie, Lang, et al.
Pubblicazione: (2024)
TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking
di: Hu, Jiyuan, et al.
Pubblicazione: (2026)
di: Hu, Jiyuan, et al.
Pubblicazione: (2026)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
di: Gu, Chenghao, et al.
Pubblicazione: (2024)
di: Gu, Chenghao, et al.
Pubblicazione: (2024)
A Deep Ordinal Distortion Estimation Approach for Distortion Rectification
di: Liao, Kang, et al.
Pubblicazione: (2020)
di: Liao, Kang, et al.
Pubblicazione: (2020)
Generative Photographic Control for Scene-Consistent Video Cinematic Editing
di: Sun, Huiqiang, et al.
Pubblicazione: (2025)
di: Sun, Huiqiang, et al.
Pubblicazione: (2025)
From Scale to Speed: Adaptive Test-Time Scaling for Image Editing
di: Qu, Xiangyan, et al.
Pubblicazione: (2026)
di: Qu, Xiangyan, et al.
Pubblicazione: (2026)
MVPortrait: Text-Guided Motion and Emotion Control for Multi-view Vivid Portrait Animation
di: Lin, Yukang, et al.
Pubblicazione: (2025)
di: Lin, Yukang, et al.
Pubblicazione: (2025)
Style-Consistent 3D Indoor Scene Synthesis with Decoupled Objects
di: Zhang, Yunfan, et al.
Pubblicazione: (2024)
di: Zhang, Yunfan, et al.
Pubblicazione: (2024)
Deep Learning for Camera Calibration and Beyond: A Survey
di: Liao, Kang, et al.
Pubblicazione: (2023)
di: Liao, Kang, et al.
Pubblicazione: (2023)
Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools
di: Yuan, Zhenlong, et al.
Pubblicazione: (2025)
di: Yuan, Zhenlong, et al.
Pubblicazione: (2025)
Fast Multi-view Consistent 3D Editing with Video Priors
di: Chen, Liyi, et al.
Pubblicazione: (2025)
di: Chen, Liyi, et al.
Pubblicazione: (2025)
In-Context Learning with Unpaired Clips for Instruction-based Video Editing
di: Liao, Xinyao, et al.
Pubblicazione: (2025)
di: Liao, Xinyao, et al.
Pubblicazione: (2025)
Revisiting Monocular 3D Object Detection with Depth Thickness Field
di: Zhang, Qiude, et al.
Pubblicazione: (2024)
di: Zhang, Qiude, et al.
Pubblicazione: (2024)
IntegratedPIFu: Integrated Pixel Aligned Implicit Function for Single-view Human Reconstruction
di: Chan, Kennard Yanting, et al.
Pubblicazione: (2022)
di: Chan, Kennard Yanting, et al.
Pubblicazione: (2022)
TSAR-MVS: Textureless-aware Segmentation and Correlative Refinement Guided Multi-View Stereo
di: Yuan, Zhenlong, et al.
Pubblicazione: (2023)
di: Yuan, Zhenlong, et al.
Pubblicazione: (2023)
Text-to-3D Generation by 2D Editing
di: Li, Haoran, et al.
Pubblicazione: (2024)
di: Li, Haoran, et al.
Pubblicazione: (2024)
3D-VirtFusion: Synthetic 3D Data Augmentation through Generative Diffusion Models and Controllable Editing
di: Dong, Shichao, et al.
Pubblicazione: (2024)
di: Dong, Shichao, et al.
Pubblicazione: (2024)
SIM: A mapping framework for built environment auditing based on street view imagery
di: Ning, Huan, et al.
Pubblicazione: (2025)
di: Ning, Huan, et al.
Pubblicazione: (2025)
TiP4GEN: Text to Immersive Panorama 4D Scene Generation
di: Xing, Ke, et al.
Pubblicazione: (2025)
di: Xing, Ke, et al.
Pubblicazione: (2025)
DGE: Direct Gaussian 3D Editing by Consistent Multi-view Editing
di: Chen, Minghao, et al.
Pubblicazione: (2024)
di: Chen, Minghao, et al.
Pubblicazione: (2024)
Robotic Manipulation is Vision-to-Geometry Mapping ($f(v) \rightarrow G$): Vision-Geometry Backbones over Language and Video Models
di: Song, Zijian, et al.
Pubblicazione: (2026)
di: Song, Zijian, et al.
Pubblicazione: (2026)
Documenti analoghi
-
From Editor to Dense Geometry Estimator
di: Wang, JiYuan, et al.
Pubblicazione: (2025) -
360 Layout Estimation via Orthogonal Planes Disentanglement and Multi-view Geometric Consistency Perception
di: Shen, Zhijie, et al.
Pubblicazione: (2023) -
Digging into contrastive learning for robust depth estimation with diffusion models
di: Wang, Jiyuan, et al.
Pubblicazione: (2024) -
Beyond Wide-Angle Images: Structure-to-Detail Video Portrait Correction via Unsupervised Spatiotemporal Adaptation
di: Nie, Wenbo, et al.
Pubblicazione: (2025) -
Advancing Real-World Parking Slot Detection with Large-Scale Dataset and Semi-Supervised Baseline
di: Zhang, Zhihao, et al.
Pubblicazione: (2025)