Moving Off-the-Grid: Scene-Grounded Video Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | van Steenkiste, Sjoerd, Zoran, Daniel, Yang, Yi, Rubanova, Yulia, Kabra, Rishabh, Doersch, Carl, Gokay, Dilara, Heyward, Joseph, Pot, Etienne, Greff, Klaus, Hudson, Drew A., Keck, Thomas Albert, Carreira, Joao, Dosovitskiy, Alexey, Sajjadi, Mehdi S. M., Kipf, Thomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
von: Seitzer, Maximilian, et al.
Veröffentlicht: (2023)
von: Seitzer, Maximilian, et al.
Veröffentlicht: (2023)
Neural Assets: 3D-Aware Multi-Object Scene Synthesis with Image Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2024)
von: Wu, Ziyi, et al.
Veröffentlicht: (2024)
Direct Motion Models for Assessing Generated Videos
von: Allen, Kelsey, et al.
Veröffentlicht: (2025)
von: Allen, Kelsey, et al.
Veröffentlicht: (2025)
How to Spin an Object: First, Get the Shape Right
von: Kabra, Rishabh, et al.
Veröffentlicht: (2024)
von: Kabra, Rishabh, et al.
Veröffentlicht: (2024)
DORSal: Diffusion for Object-centric Representations of Scenes et al
von: Jabri, Allan, et al.
Veröffentlicht: (2023)
von: Jabri, Allan, et al.
Veröffentlicht: (2023)
Scaling 4D Representations
von: Carreira, João, et al.
Veröffentlicht: (2024)
von: Carreira, João, et al.
Veröffentlicht: (2024)
Neural USD: An object-centric framework for iterative editing and control
von: Escontrela, Alejandro, et al.
Veröffentlicht: (2025)
von: Escontrela, Alejandro, et al.
Veröffentlicht: (2025)
Learning from One Continuous Video Stream
von: Carreira, João, et al.
Veröffentlicht: (2023)
von: Carreira, João, et al.
Veröffentlicht: (2023)
Learning from Streaming Video with Orthogonal Gradients
von: Han, Tengda, et al.
Veröffentlicht: (2025)
von: Han, Tengda, et al.
Veröffentlicht: (2025)
BootsTAP: Bootstrapped Training for Tracking-Any-Point
von: Doersch, Carl, et al.
Veröffentlicht: (2024)
von: Doersch, Carl, et al.
Veröffentlicht: (2024)
A Mixed Diet Makes DINO An Omnivorous Vision Encoder
von: Kabra, Rishabh, et al.
Veröffentlicht: (2026)
von: Kabra, Rishabh, et al.
Veröffentlicht: (2026)
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
von: Hasson, Yana, et al.
Veröffentlicht: (2025)
von: Hasson, Yana, et al.
Veröffentlicht: (2025)
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
LayerLock: Non-collapsing Representation Learning with Progressive Freezing
von: Erdogan, Goker, et al.
Veröffentlicht: (2025)
von: Erdogan, Goker, et al.
Veröffentlicht: (2025)
Recurrent Video Masked Autoencoders
von: Zoran, Daniel, et al.
Veröffentlicht: (2025)
von: Zoran, Daniel, et al.
Veröffentlicht: (2025)
From Image to Video: An Empirical Study of Diffusion Representations
von: Vélez, Pedro, et al.
Veröffentlicht: (2025)
von: Vélez, Pedro, et al.
Veröffentlicht: (2025)
How Does Code Pretraining Affect Language Model Task Performance?
von: Petty, Jackson, et al.
Veröffentlicht: (2024)
von: Petty, Jackson, et al.
Veröffentlicht: (2024)
Frozen Forecasting: A Unified Evaluation
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
OpenWorldSAM: Extending SAM2 for Universal Image Segmentation with Language Prompts
von: Xiao, Shiting, et al.
Veröffentlicht: (2025)
von: Xiao, Shiting, et al.
Veröffentlicht: (2025)
Unique Lives, Shared World: Learning from Single-Life Videos
von: Han, Tengda, et al.
Veröffentlicht: (2025)
von: Han, Tengda, et al.
Veröffentlicht: (2025)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
Forecasting Motion in the Wild
von: Thakkar, Neerja, et al.
Veröffentlicht: (2026)
von: Thakkar, Neerja, et al.
Veröffentlicht: (2026)
Home in A Hybrid World
von: Pot, Martin
Veröffentlicht: (2022)
von: Pot, Martin
Veröffentlicht: (2022)
Osons être des littéraires
von: Oliver Pot
Veröffentlicht: (2016)
von: Oliver Pot
Veröffentlicht: (2016)
Making sense of the protests in Turkey (and Brazil): contesting neo-liberal urbanism in ‘Rebel Cities’
von: Bulent Gokay
Veröffentlicht: (2015)
von: Bulent Gokay
Veröffentlicht: (2015)
TRecViT: A Recurrent Video Transformer
von: Pătrăucean, Viorica, et al.
Veröffentlicht: (2024)
von: Pătrăucean, Viorica, et al.
Veröffentlicht: (2024)
A Systematic Comparison of Syllogistic Reasoning in Humans and Language Models
von: Eisape, Tiwalayo, et al.
Veröffentlicht: (2023)
von: Eisape, Tiwalayo, et al.
Veröffentlicht: (2023)
The Impact of Depth on Compositional Generalization in Transformer Language Models
von: Petty, Jackson, et al.
Veröffentlicht: (2023)
von: Petty, Jackson, et al.
Veröffentlicht: (2023)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
von: Zholus, Artem, et al.
Veröffentlicht: (2025)
von: Zholus, Artem, et al.
Veröffentlicht: (2025)
Anfitriones, El Deparatamento de Steward
von: Chi Pot, Raúl
Veröffentlicht: (2008)
von: Chi Pot, Raúl
Veröffentlicht: (2008)
Leveraging VLM-Based Pipelines to Annotate 3D Objects
von: Kabra, Rishabh, et al.
Veröffentlicht: (2023)
von: Kabra, Rishabh, et al.
Veröffentlicht: (2023)
DPconv: Super-Polynomially Faster Join Ordering
von: Stoian, Mihail, et al.
Veröffentlicht: (2024)
von: Stoian, Mihail, et al.
Veröffentlicht: (2024)
Synthesis and evaluation of PVC‐Cu/Al2O3 nanocomposite membranes for removing of natural organic matter from the wastewater
von: Seyed Mehdi Sajjadi, et al.
Veröffentlicht: (2024)
von: Seyed Mehdi Sajjadi, et al.
Veröffentlicht: (2024)
Learning rigid-body simulators over implicit shapes for large-scale scenes and vision
von: Rubanova, Yulia, et al.
Veröffentlicht: (2024)
von: Rubanova, Yulia, et al.
Veröffentlicht: (2024)
Scaling Face Interaction Graph Networks to Real World Scenes
von: Lopez-Guevara, Tatiana, et al.
Veröffentlicht: (2024)
von: Lopez-Guevara, Tatiana, et al.
Veröffentlicht: (2024)
Olophrum assimile Pk., an addition to the British list
von: Beare, T. Hudson (Thomas Hudson)
Veröffentlicht: (1908)
von: Beare, T. Hudson (Thomas Hudson)
Veröffentlicht: (1908)
Tuning load redistribution and damage near heterogeneous interfaces
von: Greff, Christian, et al.
Veröffentlicht: (2024)
von: Greff, Christian, et al.
Veröffentlicht: (2024)
U.S. drug policy and supply-side strategies : ssessing effectiveness and results = La política antidrogas de Estados Unidos y las estrategias de control de oferta : Una evaluación de su efectividad y resultados / Michelle Keck, Guadalupe Correa Cabrera
von: Keck, Michelle
von: Keck, Michelle
Population growth, shifting cultivation, and unsustainable agricultural development : a case study in Madagascar / Andrew Keck, Narendra P. Sharma, Gershon Feder
von: Keck, Andrew
Veröffentlicht: (1994)
von: Keck, Andrew
Veröffentlicht: (1994)
Ähnliche Einträge
-
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
von: Seitzer, Maximilian, et al.
Veröffentlicht: (2023) -
Neural Assets: 3D-Aware Multi-Object Scene Synthesis with Image Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2024) -
Direct Motion Models for Assessing Generated Videos
von: Allen, Kelsey, et al.
Veröffentlicht: (2025) -
How to Spin an Object: First, Get the Shape Right
von: Kabra, Rishabh, et al.
Veröffentlicht: (2024) -
DORSal: Diffusion for Object-centric Representations of Scenes et al
von: Jabri, Allan, et al.
Veröffentlicht: (2023)