GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Leslie, Hur, Junhwa, Herrmann, Charles, Zhan, Fangneng, Zickler, Todd, Sun, Deqing, Pfister, Hanspeter |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boundary Attention: Learning curves, corners, junctions and grouping
by: Polansky, Mia Gaia, et al.
Published: (2024)
by: Polansky, Mia Gaia, et al.
Published: (2024)
UFO-4D: Unposed Feedforward 4D Reconstruction from Two Images
by: Hur, Junhwa, et al.
Published: (2026)
by: Hur, Junhwa, et al.
Published: (2026)
LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory
by: Zhang, Junyi, et al.
Published: (2026)
by: Zhang, Junyi, et al.
Published: (2026)
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance
by: Xu, Tianling, et al.
Published: (2025)
by: Xu, Tianling, et al.
Published: (2025)
Motion Prompting: Controlling Video Generation with Motion Trajectories
by: Geng, Daniel, et al.
Published: (2024)
by: Geng, Daniel, et al.
Published: (2024)
RoboTAG: End-to-end Robot Configuration Estimation via Topological Alignment Graph
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion
by: Zhang, Junyi, et al.
Published: (2024)
by: Zhang, Junyi, et al.
Published: (2024)
Telling Left from Right: Identifying Geometry-Aware Semantic Correspondence
by: Zhang, Junyi, et al.
Published: (2023)
by: Zhang, Junyi, et al.
Published: (2023)
Structural Inhomogeneities and Suppressed Magneto-Structural Coupling in Mn-Substituted GeCo2O4
by: Sharma, Shivani, et al.
Published: (2025)
by: Sharma, Shivani, et al.
Published: (2025)
Generative Perception of Shape and Material from Differential Motion
by: Han, Xinran Nicole, et al.
Published: (2025)
by: Han, Xinran Nicole, et al.
Published: (2025)
Abstract 3D Perception for Spatial Intelligence in Vision-Language Models
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
Emergent Vortex Ordering in a Multiflavor Pyrochlore-Lattice Compound GeCo$_2$O$_4$
by: Mo, Jiajun, et al.
Published: (2026)
by: Mo, Jiajun, et al.
Published: (2026)
Vortex order in magnetic frustrated GeNi$_2$O$_4$ and GeCo$_2$O$_4$ spinels
by: Beauvois, K., et al.
Published: (2026)
by: Beauvois, K., et al.
Published: (2026)
High-Resolution Frame Interpolation with Patch-based Cascaded Diffusion
by: Hur, Junhwa, et al.
Published: (2024)
by: Hur, Junhwa, et al.
Published: (2024)
GeCo-SRT: Geometry-aware Continual Adaptation for Robotic Cross-Task Sim-to-Real Transfer
by: Yu, Wenbo, et al.
Published: (2026)
by: Yu, Wenbo, et al.
Published: (2026)
LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs
by: Kim, Jihwan, et al.
Published: (2026)
by: Kim, Jihwan, et al.
Published: (2026)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
WonderJourney: Going from Anywhere to Everywhere
by: Yu, Hong-Xing, et al.
Published: (2023)
by: Yu, Hong-Xing, et al.
Published: (2023)
Force Prompting: Video Generation Models Can Learn and Generalize Physics-based Control Signals
by: Gillman, Nate, et al.
Published: (2025)
by: Gillman, Nate, et al.
Published: (2025)
GenLens: A Systematic Evaluation of Visual GenAI Model Outputs
by: Lin, Tica, et al.
Published: (2024)
by: Lin, Tica, et al.
Published: (2024)
Grayscale to Hyperspectral at Any Resolution Using a Phase-Only Lens
by: Hazineh, Dean, et al.
Published: (2024)
by: Hazineh, Dean, et al.
Published: (2024)
CObL: Toward Zero-Shot Ordinal Layering without User Prompting
by: Damaraju, Aneel, et al.
Published: (2025)
by: Damaraju, Aneel, et al.
Published: (2025)
MASIV: Toward Material-Agnostic System Identification from Videos
by: Zhao, Yizhou, et al.
Published: (2025)
by: Zhao, Yizhou, et al.
Published: (2025)
Sportify: Question Answering with Embedded Visualizations and Personified Narratives for Sports Video
by: Lee, Chunggi, et al.
Published: (2024)
by: Lee, Chunggi, et al.
Published: (2024)
Visual Acoustic Fields
by: Li, Yuelei, et al.
Published: (2025)
by: Li, Yuelei, et al.
Published: (2025)
Multistable Shape from Shading Emerges from Patch Diffusion
by: Han, Xinran Nicole, et al.
Published: (2024)
by: Han, Xinran Nicole, et al.
Published: (2024)
RiGS: Rigid-aware 4D Gaussian Splatting from a Single Monocular Video
by: Wu, Chenyu, et al.
Published: (2026)
by: Wu, Chenyu, et al.
Published: (2026)
PanoCoach: Enhancing Tactical Coaching and Communication in Soccer with Mixed-Reality Telepresence
by: Kang, Andrew, et al.
Published: (2024)
by: Kang, Andrew, et al.
Published: (2024)
CTRL-GS: Cascaded Temporal Residue Learning for 4D Gaussian Splatting
by: Hou, Karly, et al.
Published: (2025)
by: Hou, Karly, et al.
Published: (2025)
General Neural Gauge Fields
by: Zhan, Fangneng, et al.
Published: (2023)
by: Zhan, Fangneng, et al.
Published: (2023)
CoMusion: Towards Consistent Stochastic Human Motion Prediction via Motion Diffusion
by: Sun, Jiarui, et al.
Published: (2023)
by: Sun, Jiarui, et al.
Published: (2023)
Lite2Relight: 3D-aware Single Image Portrait Relighting
by: Rao, Pramod, et al.
Published: (2024)
by: Rao, Pramod, et al.
Published: (2024)
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding
by: Kim, Namho, et al.
Published: (2025)
by: Kim, Namho, et al.
Published: (2025)
Emergent Temporal Correspondences from Video Diffusion Transformers
by: Nam, Jisu, et al.
Published: (2025)
by: Nam, Jisu, et al.
Published: (2025)
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
by: Nam, Jisu, et al.
Published: (2026)
by: Nam, Jisu, et al.
Published: (2026)
Fréchet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos
by: Liu, Jiahe, et al.
Published: (2024)
by: Liu, Jiahe, et al.
Published: (2024)
Under One Sun: Multi-Object Generative Perception of Materials and Illumination
by: Yoshii, Nobuo, et al.
Published: (2026)
by: Yoshii, Nobuo, et al.
Published: (2026)
MechVerse: Evaluating Physical Motion Consistency in Video Generation Models
by: Jain, Rahul, et al.
Published: (2026)
by: Jain, Rahul, et al.
Published: (2026)
Sora Generates Videos with Stunning Geometrical Consistency
by: Li, Xuanyi, et al.
Published: (2024)
by: Li, Xuanyi, et al.
Published: (2024)
3DPR: Single Image 3D Portrait Relight using Generative Priors
by: Rao, Pramod, et al.
Published: (2025)
by: Rao, Pramod, et al.
Published: (2025)
Similar Items
-
Boundary Attention: Learning curves, corners, junctions and grouping
by: Polansky, Mia Gaia, et al.
Published: (2024) -
UFO-4D: Unposed Feedforward 4D Reconstruction from Two Images
by: Hur, Junhwa, et al.
Published: (2026) -
LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory
by: Zhang, Junyi, et al.
Published: (2026) -
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance
by: Xu, Tianling, et al.
Published: (2025) -
Motion Prompting: Controlling Video Generation with Motion Trajectories
by: Geng, Daniel, et al.
Published: (2024)