Geometry-as-context: Modulating Explicit 3D in Scene-consistent Video Generation to Geometry Context
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, JiaKui, Liu, Jialun, Yang, Liying, Zhang, Xinliang, Li, Kaiwen, Zeng, Shuang, Li, Yuanwei, Huang, Haibin, Zhang, Chi, Lu, Yanye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Auto-Regressively Generating Multi-View Consistent Images
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
Bridging Degradation Discrimination and Generation for Universal Image Restoration
von: Hu, JiaKui, et al.
Veröffentlicht: (2026)
von: Hu, JiaKui, et al.
Veröffentlicht: (2026)
Universal Image Restoration Pre-training via Degradation Classification
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
Enhancing Image Restoration Transformer via Adaptive Translation Equivariance
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
Universal Image Restoration Pre-training via Masked Degradation Classification
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
Omni-View: Unlocking How Generation Facilitates Understanding in Unified 3D Model based on Multiview images
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration
von: Yao, Zhengjian, et al.
Veröffentlicht: (2026)
von: Yao, Zhengjian, et al.
Veröffentlicht: (2026)
RADA: Region-Aware Dual-encoder Auxiliary learning for Barely-supervised Medical Image Segmentation
von: Zeng, Shuang, et al.
Veröffentlicht: (2026)
von: Zeng, Shuang, et al.
Veröffentlicht: (2026)
SuperCL: Superpixel Guided Contrastive Learning for Medical Image Segmentation Pre-training
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
Inter- and Intra-image Refinement for Few Shot Segmentation
von: Fu, Ourui, et al.
Veröffentlicht: (2025)
von: Fu, Ourui, et al.
Veröffentlicht: (2025)
V2C-CBM: Building Concept Bottlenecks with Vision-to-Concept Tokenizer
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
Chat-CBM: Towards Interactive Concept Bottleneck Models with Frozen Large Language Models
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
Low-Rank Mixture-of-Experts for Continual Medical Image Segmentation
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
AdaTok: Adaptive Token Compression with Object-Aware Representations for Efficient Multimodal LLMs
von: Zhang, Xinliang, et al.
Veröffentlicht: (2025)
von: Zhang, Xinliang, et al.
Veröffentlicht: (2025)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
Exploiting Inherent Class Label: Towards Robust Scribble Supervised Semantic Segmentation
von: Zhang, Xinliang, et al.
Veröffentlicht: (2025)
von: Zhang, Xinliang, et al.
Veröffentlicht: (2025)
Spatial-Temporal State Propagation Autoregressive Model for 4D Object Generation
von: Yang, Liying, et al.
Veröffentlicht: (2026)
von: Yang, Liying, et al.
Veröffentlicht: (2026)
Scribble Hides Class: Promoting Scribble-Based Weakly-Supervised Semantic Segmentation with Its Class Label
von: Zhang, Xinliang, et al.
Veröffentlicht: (2024)
von: Zhang, Xinliang, et al.
Veröffentlicht: (2024)
VEIGAR: View-consistent Explicit Inpainting and Geometry Alignment for 3D object Removal
von: Do, Pham Khai Nguyen, et al.
Veröffentlicht: (2025)
von: Do, Pham Khai Nguyen, et al.
Veröffentlicht: (2025)
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
von: Xu, Tian-Xing, et al.
Veröffentlicht: (2025)
von: Xu, Tian-Xing, et al.
Veröffentlicht: (2025)
Training-free Test-time Improvement for Explainable Medical Image Classification
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
Scalable Scene Modeling from Perspective Imaging: Physics-based Appearance and Geometry Inference
von: Song, Shuang
Veröffentlicht: (2024)
von: Song, Shuang
Veröffentlicht: (2024)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Towards Explicit Geometry-Reflectance Collaboration for Generalized LiDAR Segmentation in Adverse Weather
von: Yang, Longyu, et al.
Veröffentlicht: (2025)
von: Yang, Longyu, et al.
Veröffentlicht: (2025)
Geometry-Aware Normalizing Wasserstein Flows for Optimal Causal Inference
von: Hou, Kaiwen
Veröffentlicht: (2023)
von: Hou, Kaiwen
Veröffentlicht: (2023)
Explicit Temporal-Semantic Modeling for Dense Video Captioning via Context-Aware Cross-Modal Interaction
von: Jia, Mingda, et al.
Veröffentlicht: (2025)
von: Jia, Mingda, et al.
Veröffentlicht: (2025)
StructuredField: Unifying Structured Geometry and Radiance Field
von: Song, Kaiwen, et al.
Veröffentlicht: (2025)
von: Song, Kaiwen, et al.
Veröffentlicht: (2025)
Geometry-aware 4D Video Generation for Robot Manipulation
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking
von: Hu, Jiyuan, et al.
Veröffentlicht: (2026)
von: Hu, Jiyuan, et al.
Veröffentlicht: (2026)
Explicit Stair Geometry Conditioning for Robust Humanoid Locomotion
von: Zhang, Jianguo, et al.
Veröffentlicht: (2026)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2026)
Geometry and Perception Guided Gaussians for Multiview-consistent 3D Generation from a Single Image
von: Li, Pufan, et al.
Veröffentlicht: (2025)
von: Li, Pufan, et al.
Veröffentlicht: (2025)
Improve Retinal Artery/Vein Classification via Channel Couplin
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
World-consistent Video Diffusion with Explicit 3D Modeling
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
TensoSDF: Roughness-aware Tensorial Representation for Robust Geometry and Material Reconstruction
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
TAPESTRY: From Geometry to Appearance via Consistent Turntable Videos
von: Zeng, Yan, et al.
Veröffentlicht: (2026)
von: Zeng, Yan, et al.
Veröffentlicht: (2026)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
von: Li, Bohan, et al.
Veröffentlicht: (2023)
von: Li, Bohan, et al.
Veröffentlicht: (2023)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation
von: Zhang, Xinliang Frederick, et al.
Veröffentlicht: (2026)
von: Zhang, Xinliang Frederick, et al.
Veröffentlicht: (2026)
RieMind: Geometry-Grounded Spatial Agent for Scene Understanding
von: Ropero, Fernando, et al.
Veröffentlicht: (2026)
von: Ropero, Fernando, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Auto-Regressively Generating Multi-View Consistent Images
von: Hu, JiaKui, et al.
Veröffentlicht: (2025) -
Bridging Degradation Discrimination and Generation for Universal Image Restoration
von: Hu, JiaKui, et al.
Veröffentlicht: (2026) -
Universal Image Restoration Pre-training via Degradation Classification
von: Hu, JiaKui, et al.
Veröffentlicht: (2025) -
Enhancing Image Restoration Transformer via Adaptive Translation Equivariance
von: Hu, JiaKui, et al.
Veröffentlicht: (2025) -
Universal Image Restoration Pre-training via Masked Degradation Classification
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)