UFM: A Simple Path towards Unified Dense Correspondence with Flow
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuchen, Keetha, Nikhil, Lyu, Chenwei, Jhamb, Bhuvan, Chen, Yutian, Qiu, Yuheng, Karhade, Jay, Jha, Shreyas, Hu, Yaoyu, Ramanan, Deva, Scherer, Sebastian, Wang, Wenshan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Any4D: Unified Feed-Forward Metric 4D Reconstruction
by: Karhade, Jay, et al.
Published: (2025)
by: Karhade, Jay, et al.
Published: (2025)
SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM
by: Keetha, Nikhil, et al.
Published: (2023)
by: Keetha, Nikhil, et al.
Published: (2023)
MAC-VO: Metrics-aware Covariance for Learning-based Stereo Visual Odometry
by: Qiu, Yuheng, et al.
Published: (2024)
by: Qiu, Yuheng, et al.
Published: (2024)
RayFronts: Open-Set Semantic Ray Frontiers for Online Scene Understanding and Exploration
by: Alama, Omar, et al.
Published: (2025)
by: Alama, Omar, et al.
Published: (2025)
Towards Foundational Models for Single-Chip Radar
by: Huang, Tianshu, et al.
Published: (2025)
by: Huang, Tianshu, et al.
Published: (2025)
RAVEN: Resilient Aerial Navigation via Open-Set Semantic Memory and Behavior Adaptation
by: Kim, Seungchan, et al.
Published: (2025)
by: Kim, Seungchan, et al.
Published: (2025)
Predicting Long-horizon Futures by Conditioning on Geometry and Time
by: Khurana, Tarasha, et al.
Published: (2024)
by: Khurana, Tarasha, et al.
Published: (2024)
Demonstrating ViSafe: Vision-enabled Safety for High-speed Detect and Avoid
by: Kapoor, Parv, et al.
Published: (2025)
by: Kapoor, Parv, et al.
Published: (2025)
UFM: Unified Feature Matching Pre-training with Multi-Modal Image Assistants
by: Di, Yide, et al.
Published: (2025)
by: Di, Yide, et al.
Published: (2025)
Not All Memories Age the Same: Autodiscovery of Adaptive Decay in Knowledge Graphs
by: Karhade, Mandar
Published: (2026)
by: Karhade, Mandar
Published: (2026)
Using Diffusion Priors for Video Amodal Segmentation
by: Chen, Kaihua, et al.
Published: (2024)
by: Chen, Kaihua, et al.
Published: (2024)
RefAV: Towards Planning-Centric Scenario Mining
by: Davidson, Cainan, et al.
Published: (2025)
by: Davidson, Cainan, et al.
Published: (2025)
Reconstruct, Inpaint, Test-Time Finetune: Dynamic Novel-view Synthesis from Monocular Videos
by: Chen, Kaihua, et al.
Published: (2025)
by: Chen, Kaihua, et al.
Published: (2025)
AnyThermal: Towards Learning Universal Representations for Thermal Perception
by: Maheshwari, Parv, et al.
Published: (2026)
by: Maheshwari, Parv, et al.
Published: (2026)
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
by: Keetha, Nikhil, et al.
Published: (2025)
by: Keetha, Nikhil, et al.
Published: (2025)
Co-Me: Confidence-Guided Token Merging for Visual Geometric Transformers
by: Chen, Yutian, et al.
Published: (2025)
by: Chen, Yutian, et al.
Published: (2025)
TartanAviation: Image, Speech, and ADS-B Trajectory Datasets for Terminal Airspace Operations
by: Patrikar, Jay, et al.
Published: (2024)
by: Patrikar, Jay, et al.
Published: (2024)
Lead Nitrate (Pb(NO)) Toxicity Effects on DNA Structure and Histopathological Damage in Gills of Common Carp (Cyprinus carpio).
by: Sharma, Ritu, et al.
Published: (2025)
by: Sharma, Ritu, et al.
Published: (2025)
Global AI Governance in Healthcare: A Cross-Jurisdictional Regulatory Analysis
by: Chakraborty, Attrayee, et al.
Published: (2024)
by: Chakraborty, Attrayee, et al.
Published: (2024)
Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection
by: Khurana, Mehar, et al.
Published: (2024)
by: Khurana, Mehar, et al.
Published: (2024)
Revisiting Few-Shot Object Detection with Vision-Language Models
by: Madan, Anish, et al.
Published: (2023)
by: Madan, Anish, et al.
Published: (2023)
SMORE: Simultaneous Map and Object REconstruction
by: Chodosh, Nathaniel, et al.
Published: (2024)
by: Chodosh, Nathaniel, et al.
Published: (2024)
Evaluating a VR System for Collecting Safety-Critical Vehicle-Pedestrian Interactions
by: Weng, Erica, et al.
Published: (2023)
by: Weng, Erica, et al.
Published: (2023)
UFM: A Community Learning Center. Agency Report [1982-83].
Published: (1983)
Published: (1983)
UFM‐Based Simulation of Competitive Multi‐Fracture Propagation in Horizontal Wells
by: Xiaojia Xue, et al.
Published: (2025)
by: Xiaojia Xue, et al.
Published: (2025)
AirIO: Learning Inertial Odometry with Enhanced IMU Feature Observability
by: Qiu, Yuheng, et al.
Published: (2025)
by: Qiu, Yuheng, et al.
Published: (2025)
Simulating Field Experiments with Large Language Models
by: Chen, Yaoyu, et al.
Published: (2024)
by: Chen, Yaoyu, et al.
Published: (2024)
Predicting Field Experiments with Large Language Models
by: Chen, Yaoyu, et al.
Published: (2025)
by: Chen, Yaoyu, et al.
Published: (2025)
FlowR: Flowing from Sparse to Dense 3D Reconstructions
by: Fischer, Tobias, et al.
Published: (2025)
by: Fischer, Tobias, et al.
Published: (2025)
Planning with Adaptive World Models for Autonomous Driving
by: Vasudevan, Arun Balajee, et al.
Published: (2024)
by: Vasudevan, Arun Balajee, et al.
Published: (2024)
Depth-supervised NeRF: Fewer Views and Faster Training for Free
by: Deng, Kangle, et al.
Published: (2021)
by: Deng, Kangle, et al.
Published: (2021)
TartanGround: A Large-Scale Dataset for Ground Robot Perception and Navigation
by: Patel, Manthan, et al.
Published: (2025)
by: Patel, Manthan, et al.
Published: (2025)
EndoUFM: Utilizing Foundation Models for Monocular depth estimation of endoscopic images
by: Yao, Xinning, et al.
Published: (2025)
by: Yao, Xinning, et al.
Published: (2025)
Towards Understanding Camera Motions in Any Video
by: Lin, Zhiqiu, et al.
Published: (2025)
by: Lin, Zhiqiu, et al.
Published: (2025)
HDMI: Learning Interactive Humanoid Whole-Body Control from Human Videos
by: Weng, Haoyang, et al.
Published: (2025)
by: Weng, Haoyang, et al.
Published: (2025)
A PTC‐Inspired Energy Reduction Technique for SAR ADCs Used in CMOS Image Sensor Readouts
by: Arun K., et al.
Published: (2025)
by: Arun K., et al.
Published: (2025)
MonoFusion: Sparse-View 4D Reconstruction via Monocular Fusion
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
PAI-Bench: A Comprehensive Benchmark For Physical AI
by: Zhou, Fengzhe, et al.
Published: (2025)
by: Zhou, Fengzhe, et al.
Published: (2025)
Multimodality Helps Unimodality: Cross-Modal Few-Shot Learning with Multimodal Models
by: Lin, Zhiqiu, et al.
Published: (2023)
by: Lin, Zhiqiu, et al.
Published: (2023)
I Can't Believe It's Not Scene Flow!
by: Khatri, Ishan, et al.
Published: (2024)
by: Khatri, Ishan, et al.
Published: (2024)
Similar Items
-
Any4D: Unified Feed-Forward Metric 4D Reconstruction
by: Karhade, Jay, et al.
Published: (2025) -
SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM
by: Keetha, Nikhil, et al.
Published: (2023) -
MAC-VO: Metrics-aware Covariance for Learning-based Stereo Visual Odometry
by: Qiu, Yuheng, et al.
Published: (2024) -
RayFronts: Open-Set Semantic Ray Frontiers for Online Scene Understanding and Exploration
by: Alama, Omar, et al.
Published: (2025) -
Towards Foundational Models for Single-Chip Radar
by: Huang, Tianshu, et al.
Published: (2025)