Efficiently Reconstructing Dynamic Scenes One D4RT at a Time
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Chuhan, Moing, Guillaume Le, Koppula, Skanda, Rocco, Ignacio, Momeni, Liliane, Xie, Junyu, Sun, Shuyang, Sukthankar, Rahul, Barral, Joëlle K., Hadsell, Raia, Ghahramani, Zoubin, Zisserman, Andrew, Zhang, Junlin, Sajjadi, Mehdi S. M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
by: Hasson, Yana, et al.
Published: (2025)
by: Hasson, Yana, et al.
Published: (2025)
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
by: Koppula, Skanda, et al.
Published: (2024)
by: Koppula, Skanda, et al.
Published: (2024)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025)
by: Zholus, Artem, et al.
Published: (2025)
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
by: Jang, Youngjoon, et al.
Published: (2025)
by: Jang, Youngjoon, et al.
Published: (2025)
BootsTAP: Bootstrapped Training for Tracking-Any-Point
by: Doersch, Carl, et al.
Published: (2024)
by: Doersch, Carl, et al.
Published: (2024)
Segment, Embed, and Align: A Universal Recipe for Aligning Subtitles to Signing
by: Jiang, Zifan, et al.
Published: (2025)
by: Jiang, Zifan, et al.
Published: (2025)
A Simple Recipe for Contrastively Pre-training Video-First Encoders Beyond 16 Frames
by: Papalampidi, Pinelopi, et al.
Published: (2023)
by: Papalampidi, Pinelopi, et al.
Published: (2023)
Lost in Translation, Found in Embeddings: Sign Language Translation and Alignment
by: Jang, Youngjoon, et al.
Published: (2025)
by: Jang, Youngjoon, et al.
Published: (2025)
Dense Optical Tracking: Connecting the Dots
by: Moing, Guillaume Le, et al.
Published: (2023)
by: Moing, Guillaume Le, et al.
Published: (2023)
Memory Consolidation Enables Long-Context Video Understanding
by: Balažević, Ivana, et al.
Published: (2024)
by: Balažević, Ivana, et al.
Published: (2024)
A Tale of Two Languages: Large-Vocabulary Continuous Sign Language Recognition from Spoken Language Supervision
by: Raude, Charles, et al.
Published: (2024)
by: Raude, Charles, et al.
Published: (2024)
Scaling 4D Representations
by: Carreira, João, et al.
Published: (2024)
by: Carreira, João, et al.
Published: (2024)
Appearance-Based Refinement for Object-Centric Motion Segmentation
by: Xie, Junyu, et al.
Published: (2023)
by: Xie, Junyu, et al.
Published: (2023)
Critical Perspectives on Interreligious Education
by: Syeed, Najeeba, et al.
Published: (2020)
by: Syeed, Najeeba, et al.
Published: (2020)
Reseña de Kaufmant, Marie-Eugénie, «Le cheval au théâtre dans l’Espagne du Siècle d’Or. Fondements idéologiques et mécanismes d’une poétique dans la comedia nueva», Binges, Éditions Orbis Tertius, 2018, 576 pp. ISBN: 978-2-36783-113-8
by: Morgane Kappès-Le Moing
Published: (2019)
by: Morgane Kappès-Le Moing
Published: (2019)
From Image to Video: An Empirical Study of Diffusion Representations
by: Vélez, Pedro, et al.
Published: (2025)
by: Vélez, Pedro, et al.
Published: (2025)
SC-Phi2: A Fine-tuned Small Language Model for StarCraft II Macromanagement Tasks
by: Khan, Muhammad Junaid, et al.
Published: (2024)
by: Khan, Muhammad Junaid, et al.
Published: (2024)
DORSal: Diffusion for Object-centric Representations of Scenes et al
by: Jabri, Allan, et al.
Published: (2023)
by: Jabri, Allan, et al.
Published: (2023)
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
by: Seitzer, Maximilian, et al.
Published: (2023)
by: Seitzer, Maximilian, et al.
Published: (2023)
Resource-Efficient Neural Networks for Embedded Systems
by: Roth, Wolfgang, et al.
Published: (2020)
by: Roth, Wolfgang, et al.
Published: (2020)
Moving Object Segmentation: All You Need Is SAM (and Flow)
by: Xie, Junyu, et al.
Published: (2024)
by: Xie, Junyu, et al.
Published: (2024)
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
by: Xie, Junyu, et al.
Published: (2026)
by: Xie, Junyu, et al.
Published: (2026)
LayerLock: Non-collapsing Representation Learning with Progressive Freezing
by: Erdogan, Goker, et al.
Published: (2025)
by: Erdogan, Goker, et al.
Published: (2025)
A Mixed Diet Makes DINO An Omnivorous Vision Encoder
by: Kabra, Rishabh, et al.
Published: (2026)
by: Kabra, Rishabh, et al.
Published: (2026)
On Lie n-centralizers, n-commuting linear maps and related mappings
by: Ghahramani, Hoger
Published: (2024)
by: Ghahramani, Hoger
Published: (2024)
Privacy Dream: A Malware Propagation‐Inspired Attack on Authentication Protocols
by: Meysam Ghahramani
Published: (2025)
by: Meysam Ghahramani
Published: (2025)
Pre-trained Gaussian Processes for Bayesian Optimization
by: Wang, Zi, et al.
Published: (2021)
by: Wang, Zi, et al.
Published: (2021)
Administración por objetivos / Anthony P. Raia ; traducción de Francisco Reséndiz Gonz lez
by: Raia, Anthony P
by: Raia, Anthony P
Percepção dos discentes e docentes sobre estágio supervisionado no curso de ciências contábeis da UNEMAT - campus de Tanga-rá da Serra
by: Franciene Souza Raia
Published: (2011)
by: Franciene Souza Raia
Published: (2011)
Evolution of Panchayati Raj System in India with Special Reference to Telangana: Issues and Challenges
by: Mallesham, Koppula, et al.
Published: (2025)
by: Mallesham, Koppula, et al.
Published: (2025)
Synthesis and evaluation of PVC‐Cu/Al2O3 nanocomposite membranes for removing of natural organic matter from the wastewater
by: Seyed Mehdi Sajjadi, et al.
Published: (2024)
by: Seyed Mehdi Sajjadi, et al.
Published: (2024)
Securing Healthcare with Deep Learning: A CNN-Based Model for medical IoT Threat Detection
by: Mohamadi, Alireza, et al.
Published: (2024)
by: Mohamadi, Alireza, et al.
Published: (2024)
Smart Sampling: Self-Attention and Bootstrapping for Improved Ensembled Q-Learning
by: Khan, Muhammad Junaid, et al.
Published: (2024)
by: Khan, Muhammad Junaid, et al.
Published: (2024)
Enhanced Multimodal Content Moderation of Children's Videos using Audiovisual Fusion
by: Ahmed, Syed Hammad, et al.
Published: (2024)
by: Ahmed, Syed Hammad, et al.
Published: (2024)
Character-Centric Understanding of Animated Movies
by: Gui, Zhongrui, et al.
Published: (2025)
by: Gui, Zhongrui, et al.
Published: (2025)
Interaction-Centric Knowledge Infusion and Transfer for Open-Vocabulary Scene Graph Generation
by: Li, Lin, et al.
Published: (2025)
by: Li, Lin, et al.
Published: (2025)
RadarSplat: Radar Gaussian Splatting for High-Fidelity Data Synthesis and 3D Reconstruction of Autonomous Driving Scenes
by: Kung, Pou-Chun, et al.
Published: (2025)
by: Kung, Pou-Chun, et al.
Published: (2025)
Einsum Networks: Fast and Scalable Learning of Tractable Probabilistic Circuits
by: Peharz, Robert, et al.
Published: (2020)
by: Peharz, Robert, et al.
Published: (2020)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
by: Wang, Chuhan, et al.
Published: (2026)
by: Wang, Chuhan, et al.
Published: (2026)
Recurrence-based Vanishing Point Detection
by: Bharadwaj, Skanda, et al.
Published: (2024)
by: Bharadwaj, Skanda, et al.
Published: (2024)
Similar Items
-
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
by: Hasson, Yana, et al.
Published: (2025) -
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
by: Koppula, Skanda, et al.
Published: (2024) -
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025) -
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
by: Jang, Youngjoon, et al.
Published: (2025) -
BootsTAP: Bootstrapped Training for Tracking-Any-Point
by: Doersch, Carl, et al.
Published: (2024)