Dynamic Reflections: Probing Video Representations with Text Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Tyler, Han, Tengda, Guibas, Leonidas, Pătrăucean, Viorica, Ovsjanikov, Maks |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unique Lives, Shared World: Learning from Single-Life Videos
by: Han, Tengda, et al.
Published: (2025)
by: Han, Tengda, et al.
Published: (2025)
Shape Non-rigid Kinematics (SNK): A Zero-Shot Method for Non-Rigid Shape Matching via Unsupervised Functional Map Regularized Reconstruction
by: Attaiki, Souhaib, et al.
Published: (2024)
by: Attaiki, Souhaib, et al.
Published: (2024)
FILTR: Extracting Topological Features from Pretrained 3D Models
by: Martinez, Louis, et al.
Published: (2026)
by: Martinez, Louis, et al.
Published: (2026)
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
by: Maillard, Léopold, et al.
Published: (2026)
by: Maillard, Léopold, et al.
Published: (2026)
Learning from Streaming Video with Orthogonal Gradients
by: Han, Tengda, et al.
Published: (2025)
by: Han, Tengda, et al.
Published: (2025)
Memory-Scalable and Simplified Functional Map Learning
by: Magnet, Robin, et al.
Published: (2024)
by: Magnet, Robin, et al.
Published: (2024)
PoNQ: a Neural QEM-based Mesh Representation
by: Maruani, Nissim, et al.
Published: (2024)
by: Maruani, Nissim, et al.
Published: (2024)
Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication
by: Behmanesh, Maysam, et al.
Published: (2025)
by: Behmanesh, Maysam, et al.
Published: (2025)
FourieRF: Few-Shot NeRFs via Progressive Fourier Frequency Control
by: Gomez, Diego, et al.
Published: (2025)
by: Gomez, Diego, et al.
Published: (2025)
Self-Supervised Dual Contouring
by: Sundararaman, Ramana, et al.
Published: (2024)
by: Sundararaman, Ramana, et al.
Published: (2024)
To Supervise or Not to Supervise: Understanding and Addressing the Key Challenges of Point Cloud Transfer Learning
by: Hadgi, Souhail, et al.
Published: (2024)
by: Hadgi, Souhail, et al.
Published: (2024)
Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features
by: Wimmer, Thomas, et al.
Published: (2023)
by: Wimmer, Thomas, et al.
Published: (2023)
Escaping Plato's Cave: Towards the Alignment of 3D and Text Latent Spaces
by: Hadgi, Souhail, et al.
Published: (2025)
by: Hadgi, Souhail, et al.
Published: (2025)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
by: Heyward, Joseph, et al.
Published: (2024)
by: Heyward, Joseph, et al.
Published: (2024)
DiffuMatch: Category-Agnostic Spectral Diffusion Priors for Robust Non-rigid Shape Matching
by: Pierson, Emery, et al.
Published: (2025)
by: Pierson, Emery, et al.
Published: (2025)
Perception Test 2025: Challenge Summary and a Unified VQA Extension
by: Heyward, Joseph, et al.
Published: (2026)
by: Heyward, Joseph, et al.
Published: (2026)
LACONIC: A 3D Layout Adapter for Controllable Image Creation
by: Maillard, Léopold, et al.
Published: (2025)
by: Maillard, Léopold, et al.
Published: (2025)
Beyond Prompts: Unconditional 3D Inversion for Out-of-Distribution Shapes
by: Chen, Victoria Yue, et al.
Published: (2026)
by: Chen, Victoria Yue, et al.
Published: (2026)
DeBaRA: Denoising-Based 3D Room Arrangement Generation
by: Maillard, Léopold, et al.
Published: (2024)
by: Maillard, Léopold, et al.
Published: (2024)
GANFusion: Feed-Forward Text-to-3D with Diffusion in GAN Space
by: Attaiki, Souhaib, et al.
Published: (2024)
by: Attaiki, Souhaib, et al.
Published: (2024)
VideoLifter: Lifting Videos to 3D with Fast Hierarchical Stereo Alignment
by: Cong, Wenyan, et al.
Published: (2025)
by: Cong, Wenyan, et al.
Published: (2025)
PatchAlign3D: Local Feature Alignment for Dense 3D Shape understanding
by: Hadgi, Souhail, et al.
Published: (2026)
by: Hadgi, Souhail, et al.
Published: (2026)
From Blobs to Spokes: High-Fidelity Surface Reconstruction via Oriented Gaussians
by: Gomez, Diego, et al.
Published: (2026)
by: Gomez, Diego, et al.
Published: (2026)
PaNDaS: Learnable Deformation Modeling with Localized Control
by: Besnier, Thomas, et al.
Published: (2024)
by: Besnier, Thomas, et al.
Published: (2024)
MoSca: Dynamic Gaussian Fusion from Casual Videos via 4D Motion Scaffolds
by: Lei, Jiahui, et al.
Published: (2024)
by: Lei, Jiahui, et al.
Published: (2024)
InfoGaussian: Structure-Aware Dynamic Gaussians through Lightweight Information Shaping
by: Zhang, Yunchao, et al.
Published: (2024)
by: Zhang, Yunchao, et al.
Published: (2024)
Dynamic Gaussian Marbles for Novel View Synthesis of Casual Monocular Videos
by: Stearns, Colton, et al.
Published: (2024)
by: Stearns, Colton, et al.
Published: (2024)
MILo: Mesh-In-the-Loop Gaussian Splatting for Detailed and Efficient Surface Reconstruction
by: Guédon, Antoine, et al.
Published: (2025)
by: Guédon, Antoine, et al.
Published: (2025)
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
by: Huang, Ian, et al.
Published: (2024)
by: Huang, Ian, et al.
Published: (2024)
Zero-Shot Image Feature Consensus with Deep Functional Maps
by: Cheng, Xinle, et al.
Published: (2024)
by: Cheng, Xinle, et al.
Published: (2024)
A Simple Recipe for Contrastively Pre-training Video-First Encoders Beyond 16 Frames
by: Papalampidi, Pinelopi, et al.
Published: (2023)
by: Papalampidi, Pinelopi, et al.
Published: (2023)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
by: Gong, Bingchen, et al.
Published: (2024)
by: Gong, Bingchen, et al.
Published: (2024)
OCH3R: Object-Centric Holistic 3D Reconstruction
by: Du, Yi, et al.
Published: (2026)
by: Du, Yi, et al.
Published: (2026)
It's Just Another Day: Unique Video Captioning by Discriminative Prompting
by: Perrett, Toby, et al.
Published: (2024)
by: Perrett, Toby, et al.
Published: (2024)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
GARLIC: GAussian Representation LearnIng for spaCe partitioning
by: Rigas, Panagiotis, et al.
Published: (2025)
by: Rigas, Panagiotis, et al.
Published: (2025)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025)
by: Lascheit, Kai, et al.
Published: (2025)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
by: Lei, Jiahui, et al.
Published: (2025)
by: Lei, Jiahui, et al.
Published: (2025)
Similar Items
-
Unique Lives, Shared World: Learning from Single-Life Videos
by: Han, Tengda, et al.
Published: (2025) -
Shape Non-rigid Kinematics (SNK): A Zero-Shot Method for Non-Rigid Shape Matching via Unsupervised Functional Map Regularized Reconstruction
by: Attaiki, Souhaib, et al.
Published: (2024) -
FILTR: Extracting Topological Features from Pretrained 3D Models
by: Martinez, Louis, et al.
Published: (2026) -
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
by: Maillard, Léopold, et al.
Published: (2026) -
Learning from Streaming Video with Orthogonal Gradients
by: Han, Tengda, et al.
Published: (2025)