DELTAv2: Accelerating Dense 3D Tracking
Fuente:
arXiv
Salvato in:
| Autori principali: | Ngo, Tuan Duc, Mirzaei, Ashkan, Qian, Guocheng, Liang, Hanwen, Gan, Chuang, Kalogerakis, Evangelos, Wonka, Peter, Wang, Chaoyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026)
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026)
DELTA: Dense Efficient Long-range 3D Tracking for any video
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2024)
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023)
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023)
EasyV2V: A High-quality Instruction-based Video Editing Framework
di: Mai, Jinjie, et al.
Pubblicazione: (2025)
di: Mai, Jinjie, et al.
Pubblicazione: (2025)
DAGE: Dual-Stream Architecture for Efficient and Fine-Grained Geometry Estimation
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026)
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026)
ShapeGen4D: Towards High Quality 4D Shape Generation from Videos
di: Yenphraphai, Jiraphon, et al.
Pubblicazione: (2025)
di: Yenphraphai, Jiraphon, et al.
Pubblicazione: (2025)
Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation
di: Chatzis, Nikitas, et al.
Pubblicazione: (2026)
di: Chatzis, Nikitas, et al.
Pubblicazione: (2026)
SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials
di: Cao, Junyi, et al.
Pubblicazione: (2025)
di: Cao, Junyi, et al.
Pubblicazione: (2025)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
di: Zheng, Shuhong, et al.
Pubblicazione: (2025)
di: Zheng, Shuhong, et al.
Pubblicazione: (2025)
GaussianCut: Interactive segmentation via graph cut for 3D Gaussian Splatting
di: Jain, Umangi, et al.
Pubblicazione: (2024)
di: Jain, Umangi, et al.
Pubblicazione: (2024)
EventSplat: 3D Gaussian Splatting from Moving Event Cameras for Real-time Rendering
di: Yura, Toshiya, et al.
Pubblicazione: (2024)
di: Yura, Toshiya, et al.
Pubblicazione: (2024)
4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
PatchAlign3D: Local Feature Alignment for Dense 3D Shape understanding
di: Hadgi, Souhail, et al.
Pubblicazione: (2026)
di: Hadgi, Souhail, et al.
Pubblicazione: (2026)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2024)
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2024)
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
di: Zhang, Biao, et al.
Pubblicazione: (2024)
di: Zhang, Biao, et al.
Pubblicazione: (2024)
OmniView: An All-Seeing Diffusion Model for 3D and 4D View Synthesis
di: Fan, Xiang, et al.
Pubblicazione: (2025)
di: Fan, Xiang, et al.
Pubblicazione: (2025)
Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features
di: Wimmer, Thomas, et al.
Pubblicazione: (2023)
di: Wimmer, Thomas, et al.
Pubblicazione: (2023)
PrEditor3D: Fast and Precise 3D Shape Editing
di: Erkoç, Ziya, et al.
Pubblicazione: (2024)
di: Erkoç, Ziya, et al.
Pubblicazione: (2024)
Helix4D: Complex 4D Mesh Generation
di: Yenphraphai, Jiraphon, et al.
Pubblicazione: (2026)
di: Yenphraphai, Jiraphon, et al.
Pubblicazione: (2026)
Wonderland: Navigating 3D Scenes from a Single Image
di: Liang, Hanwen, et al.
Pubblicazione: (2024)
di: Liang, Hanwen, et al.
Pubblicazione: (2024)
FROST-STA: Frozen Dense Features for the Ego4D Short-Term Object Interaction Anticipation
di: Wang, Chaoyang, et al.
Pubblicazione: (2026)
di: Wang, Chaoyang, et al.
Pubblicazione: (2026)
Human Geometry Distribution for 3D Animation Generation
di: Tang, Xiangjun, et al.
Pubblicazione: (2025)
di: Tang, Xiangjun, et al.
Pubblicazione: (2025)
DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models
di: Wu, Ziyi, et al.
Pubblicazione: (2025)
di: Wu, Ziyi, et al.
Pubblicazione: (2025)
DePT3R: Joint Dense Point Tracking and 3D Reconstruction of Dynamic Scenes in a Single Forward Pass
di: Alumootil, Vivek, et al.
Pubblicazione: (2025)
di: Alumootil, Vivek, et al.
Pubblicazione: (2025)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
Pix4Point: Image Pretrained Standard Transformers for 3D Point Cloud Understanding
di: Qian, Guocheng, et al.
Pubblicazione: (2022)
di: Qian, Guocheng, et al.
Pubblicazione: (2022)
LASPA: Latent Spatial Alignment for Fast Training-free Single Image Editing
di: Alharbi, Yazeed, et al.
Pubblicazione: (2024)
di: Alharbi, Yazeed, et al.
Pubblicazione: (2024)
EditCLIP: Representation Learning for Image Editing
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
NIVeL: Neural Implicit Vector Layers for Text-to-Vector Generation
di: Thamizharasan, Vikas, et al.
Pubblicazione: (2024)
di: Thamizharasan, Vikas, et al.
Pubblicazione: (2024)
VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control
di: Bahmani, Sherwin, et al.
Pubblicazione: (2024)
di: Bahmani, Sherwin, et al.
Pubblicazione: (2024)
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
di: Nam, Jisu, et al.
Pubblicazione: (2026)
di: Nam, Jisu, et al.
Pubblicazione: (2026)
EfficientViT: Multi-Scale Linear Attention for High-Resolution Dense Prediction
di: Cai, Han, et al.
Pubblicazione: (2022)
di: Cai, Han, et al.
Pubblicazione: (2022)
Dense Matchers for Dense Tracking
di: Jelínek, Tomáš, et al.
Pubblicazione: (2024)
di: Jelínek, Tomáš, et al.
Pubblicazione: (2024)
efunc: An Efficient Function Representation without Neural Networks
di: Zhang, Biao, et al.
Pubblicazione: (2025)
di: Zhang, Biao, et al.
Pubblicazione: (2025)
LaRI: Layered Ray Intersections for Single-view 3D Geometric Reasoning
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
Track4World: Feedforward World-centric Dense 3D Tracking of All Pixels
di: Lu, Jiahao, et al.
Pubblicazione: (2026)
di: Lu, Jiahao, et al.
Pubblicazione: (2026)
GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
di: Petrov, Dmitry, et al.
Pubblicazione: (2024)
Stratified Domain Adaptation: A Progressive Self-Training Approach for Scene Text Recognition
di: Le, Kha Nhat, et al.
Pubblicazione: (2024)
di: Le, Kha Nhat, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026) -
DELTA: Dense Efficient Long-range 3D Tracking for any video
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2024) -
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023) -
EasyV2V: A High-quality Instruction-based Video Editing Framework
di: Mai, Jinjie, et al.
Pubblicazione: (2025) -
DAGE: Dual-Stream Architecture for Efficient and Fine-Grained Geometry Estimation
di: Ngo, Tuan Duc, et al.
Pubblicazione: (2026)