DELTAv2: Accelerating Dense 3D Tracking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ngo, Tuan Duc, Mirzaei, Ashkan, Qian, Guocheng, Liang, Hanwen, Gan, Chuang, Kalogerakis, Evangelos, Wonka, Peter, Wang, Chaoyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026)
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026)
DELTA: Dense Efficient Long-range 3D Tracking for any video
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2024)
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
DAGE: Dual-Stream Architecture for Efficient and Fine-Grained Geometry Estimation
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026)
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026)
ShapeGen4D: Towards High Quality 4D Shape Generation from Videos
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2025)
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2025)
Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation
von: Chatzis, Nikitas, et al.
Veröffentlicht: (2026)
von: Chatzis, Nikitas, et al.
Veröffentlicht: (2026)
SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials
von: Cao, Junyi, et al.
Veröffentlicht: (2025)
von: Cao, Junyi, et al.
Veröffentlicht: (2025)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
von: Zheng, Shuhong, et al.
Veröffentlicht: (2025)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2025)
GaussianCut: Interactive segmentation via graph cut for 3D Gaussian Splatting
von: Jain, Umangi, et al.
Veröffentlicht: (2024)
von: Jain, Umangi, et al.
Veröffentlicht: (2024)
EventSplat: 3D Gaussian Splatting from Moving Event Cameras for Real-time Rendering
von: Yura, Toshiya, et al.
Veröffentlicht: (2024)
von: Yura, Toshiya, et al.
Veröffentlicht: (2024)
4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
PatchAlign3D: Local Feature Alignment for Dense 3D Shape understanding
von: Hadgi, Souhail, et al.
Veröffentlicht: (2026)
von: Hadgi, Souhail, et al.
Veröffentlicht: (2026)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
OmniView: An All-Seeing Diffusion Model for 3D and 4D View Synthesis
von: Fan, Xiang, et al.
Veröffentlicht: (2025)
von: Fan, Xiang, et al.
Veröffentlicht: (2025)
Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features
von: Wimmer, Thomas, et al.
Veröffentlicht: (2023)
von: Wimmer, Thomas, et al.
Veröffentlicht: (2023)
PrEditor3D: Fast and Precise 3D Shape Editing
von: Erkoç, Ziya, et al.
Veröffentlicht: (2024)
von: Erkoç, Ziya, et al.
Veröffentlicht: (2024)
Helix4D: Complex 4D Mesh Generation
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
Wonderland: Navigating 3D Scenes from a Single Image
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
FROST-STA: Frozen Dense Features for the Ego4D Short-Term Object Interaction Anticipation
von: Wang, Chaoyang, et al.
Veröffentlicht: (2026)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2026)
Human Geometry Distribution for 3D Animation Generation
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
DePT3R: Joint Dense Point Tracking and 3D Reconstruction of Dynamic Scenes in a Single Forward Pass
von: Alumootil, Vivek, et al.
Veröffentlicht: (2025)
von: Alumootil, Vivek, et al.
Veröffentlicht: (2025)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
Pix4Point: Image Pretrained Standard Transformers for 3D Point Cloud Understanding
von: Qian, Guocheng, et al.
Veröffentlicht: (2022)
von: Qian, Guocheng, et al.
Veröffentlicht: (2022)
LASPA: Latent Spatial Alignment for Fast Training-free Single Image Editing
von: Alharbi, Yazeed, et al.
Veröffentlicht: (2024)
von: Alharbi, Yazeed, et al.
Veröffentlicht: (2024)
EditCLIP: Representation Learning for Image Editing
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
NIVeL: Neural Implicit Vector Layers for Text-to-Vector Generation
von: Thamizharasan, Vikas, et al.
Veröffentlicht: (2024)
von: Thamizharasan, Vikas, et al.
Veröffentlicht: (2024)
VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
von: Nam, Jisu, et al.
Veröffentlicht: (2026)
von: Nam, Jisu, et al.
Veröffentlicht: (2026)
EfficientViT: Multi-Scale Linear Attention for High-Resolution Dense Prediction
von: Cai, Han, et al.
Veröffentlicht: (2022)
von: Cai, Han, et al.
Veröffentlicht: (2022)
Dense Matchers for Dense Tracking
von: Jelínek, Tomáš, et al.
Veröffentlicht: (2024)
von: Jelínek, Tomáš, et al.
Veröffentlicht: (2024)
efunc: An Efficient Function Representation without Neural Networks
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
LaRI: Layered Ray Intersections for Single-view 3D Geometric Reasoning
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Track4World: Feedforward World-centric Dense 3D Tracking of All Pixels
von: Lu, Jiahao, et al.
Veröffentlicht: (2026)
von: Lu, Jiahao, et al.
Veröffentlicht: (2026)
GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
von: Petrov, Dmitry, et al.
Veröffentlicht: (2024)
Stratified Domain Adaptation: A Progressive Self-Training Approach for Scene Text Recognition
von: Le, Kha Nhat, et al.
Veröffentlicht: (2024)
von: Le, Kha Nhat, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026) -
DELTA: Dense Efficient Long-range 3D Tracking for any video
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2024) -
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023) -
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025) -
DAGE: Dual-Stream Architecture for Efficient and Fine-Grained Geometry Estimation
von: Ngo, Tuan Duc, et al.
Veröffentlicht: (2026)