Gaga: Group Any Gaussians via 3D-aware Memory Bank
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lyu, Weijie, Li, Xueting, Kundu, Abhijit, Tsai, Yi-Hsuan, Yang, Ming-Hsuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
Edit3r: Instant 3D Scene Editing from Sparse Unposed Images
von: Liu, Jiageng, et al.
Veröffentlicht: (2025)
von: Liu, Jiageng, et al.
Veröffentlicht: (2025)
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning
von: He, Jixuan, et al.
Veröffentlicht: (2026)
von: He, Jixuan, et al.
Veröffentlicht: (2026)
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
von: Lyu, Weijie, et al.
Veröffentlicht: (2026)
von: Lyu, Weijie, et al.
Veröffentlicht: (2026)
Tex4D: Zero-shot 4D Scene Texturing with Video Diffusion Models
von: Bao, Jingzhi, et al.
Veröffentlicht: (2024)
von: Bao, Jingzhi, et al.
Veröffentlicht: (2024)
FaceLift: Learning Generalizable Single Image 3D Face Reconstruction from Synthetic Heads
von: Lyu, Weijie, et al.
Veröffentlicht: (2024)
von: Lyu, Weijie, et al.
Veröffentlicht: (2024)
Layout-your-3D: Controllable and Precise 3D Generation with 2D Blueprint
von: Zhou, Junwei, et al.
Veröffentlicht: (2024)
von: Zhou, Junwei, et al.
Veröffentlicht: (2024)
Ranking-aware adapter for text-driven image ordering with CLIP
von: Yu, Wei-Hsiang, et al.
Veröffentlicht: (2024)
von: Yu, Wei-Hsiang, et al.
Veröffentlicht: (2024)
Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts
von: Fang, Shuangkang, et al.
Veröffentlicht: (2024)
von: Fang, Shuangkang, et al.
Veröffentlicht: (2024)
Synthesizing Consistent Novel Views via 3D Epipolar Attention without Re-Training
von: Ye, Botao, et al.
Veröffentlicht: (2025)
von: Ye, Botao, et al.
Veröffentlicht: (2025)
CoCo4D: Comprehensive and Complex 4D Scene Generation
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
No Pose, No Problem: Surprisingly Simple 3D Gaussian Splats from Sparse Unposed Images
von: Ye, Botao, et al.
Veröffentlicht: (2024)
von: Ye, Botao, et al.
Veröffentlicht: (2024)
InstaInpaint: Instant 3D-Scene Inpainting with Masked Large Reconstruction Model
von: You, Junqi, et al.
Veröffentlicht: (2025)
von: You, Junqi, et al.
Veröffentlicht: (2025)
Pyramid Diffusion for Fine 3D Large Scene Generation
von: Liu, Yuheng, et al.
Veröffentlicht: (2023)
von: Liu, Yuheng, et al.
Veröffentlicht: (2023)
Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting
von: Lu, Shu-Wei, et al.
Veröffentlicht: (2025)
von: Lu, Shu-Wei, et al.
Veröffentlicht: (2025)
ReCoSplat: Autoregressive Feed-Forward Gaussian Splatting Using Render-and-Compare
von: Cheng, Freeman, et al.
Veröffentlicht: (2026)
von: Cheng, Freeman, et al.
Veröffentlicht: (2026)
Self-training Room Layout Estimation via Geometry-aware Ray-casting
von: Solarte, Bolivar, et al.
Veröffentlicht: (2024)
von: Solarte, Bolivar, et al.
Veröffentlicht: (2024)
Controllable 3D Outdoor Scene Generation via Scene Graphs
von: Liu, Yuheng, et al.
Veröffentlicht: (2025)
von: Liu, Yuheng, et al.
Veröffentlicht: (2025)
Text-Driven Image Editing via Learnable Regions
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
MeshLLM: Empowering Large Language Models to Progressively Understand and Generate 3D Mesh
von: Fang, Shuangkang, et al.
Veröffentlicht: (2025)
von: Fang, Shuangkang, et al.
Veröffentlicht: (2025)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
HoliGS: Holistic Gaussian Splatting for Embodied View Synthesis
von: Wang, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyuan, et al.
Veröffentlicht: (2025)
GALA3D: Towards Text-to-3D Complex Scene Generation via Layout-guided Generative Gaussian Splatting
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2024)
RTracker: Recoverable Tracking via PN Tree Structured Memory
von: Huang, Yuqing, et al.
Veröffentlicht: (2024)
von: Huang, Yuqing, et al.
Veröffentlicht: (2024)
Segment Any 3D Gaussians
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
Any Image Restoration via Efficient Spatial-Frequency Degradation Adaptation
von: Ren, Bin, et al.
Veröffentlicht: (2025)
von: Ren, Bin, et al.
Veröffentlicht: (2025)
IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation
von: Lin, Yuanze, et al.
Veröffentlicht: (2025)
von: Lin, Yuanze, et al.
Veröffentlicht: (2025)
Restage4D: Reanimating Deformable 3D Reconstruction from a Single Video
von: He, Jixuan, et al.
Veröffentlicht: (2025)
von: He, Jixuan, et al.
Veröffentlicht: (2025)
Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling
von: Qiu, Xiaowen, et al.
Veröffentlicht: (2025)
von: Qiu, Xiaowen, et al.
Veröffentlicht: (2025)
HENet++: Hybrid Encoding and Multi-task Learning for 3D Perception and End-to-end Autonomous Driving
von: Xia, Zhongyu, et al.
Veröffentlicht: (2025)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2025)
Segment Any 4D Gaussians
von: Ji, Shengxiang, et al.
Veröffentlicht: (2024)
von: Ji, Shengxiang, et al.
Veröffentlicht: (2024)
Depth Any Panoramas: A Foundation Model for Panoramic Depth Estimation
von: Lin, Xin, et al.
Veröffentlicht: (2025)
von: Lin, Xin, et al.
Veröffentlicht: (2025)
DrivingGaussian: Composite Gaussian Splatting for Surrounding Dynamic Autonomous Driving Scenes
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2023)
Efficient Video Object Segmentation via Modulated Cross-Attention Memory
von: Shaker, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Shaker, Abdelrahman, et al.
Veröffentlicht: (2024)
Confronting Ambiguity in 6D Object Pose Estimation via Score-Based Diffusion on SE(3)
von: Hsiao, Tsu-Ching, et al.
Veröffentlicht: (2023)
von: Hsiao, Tsu-Ching, et al.
Veröffentlicht: (2023)
AutoOcc: Automatic Open-Ended Semantic Occupancy Annotation via Vision-Language Guided Gaussian Splatting
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
DrivingGaussian++: Towards Realistic Reconstruction and Editable Simulation for Surrounding Dynamic Driving Scenes
von: Xiong, Yajiao, et al.
Veröffentlicht: (2025)
von: Xiong, Yajiao, et al.
Veröffentlicht: (2025)
EA3D: Online Open-World 3D Object Extraction from Streaming Videos
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaoyu, et al.
Veröffentlicht: (2025)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023) -
Edit3r: Instant 3D Scene Editing from Sparse Unposed Images
von: Liu, Jiageng, et al.
Veröffentlicht: (2025) -
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023) -
Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning
von: He, Jixuan, et al.
Veröffentlicht: (2026) -
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
von: Lyu, Weijie, et al.
Veröffentlicht: (2026)