Saved in:
| Main Authors: | Jiang, Hanwen, Karpur, Arjun, Cao, Bingyi, Huang, Qixing, Araujo, Andre |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.12979 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LFM-3D: Learnable Feature Matching Across Wide Baselines Using 3D Signals
by: Karpur, Arjun, et al.
Published: (2023)
by: Karpur, Arjun, et al.
Published: (2023)
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
by: Jiang, Hanwen, et al.
Published: (2024)
by: Jiang, Hanwen, et al.
Published: (2024)
Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation Models
by: Jiang, Yu, et al.
Published: (2026)
by: Jiang, Yu, et al.
Published: (2026)
TIPS: Text-Image Pretraining with Spatial awareness
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
CoFie: Learning Compact Neural Surface Representations with Coordinate Fields
by: Jiang, Hanwen, et al.
Published: (2024)
by: Jiang, Hanwen, et al.
Published: (2024)
MambaGlue: Fast and Robust Local Feature Matching With Mamba
by: Ryoo, Kihwan, et al.
Published: (2025)
by: Ryoo, Kihwan, et al.
Published: (2025)
WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling
by: Fang, Shaoheng, et al.
Published: (2025)
by: Fang, Shaoheng, et al.
Published: (2025)
Global-to-Local or Local-to-Global? Enhancing Image Retrieval with Efficient Local Search and Effective Global Re-ranking
by: Aiger, Dror, et al.
Published: (2025)
by: Aiger, Dror, et al.
Published: (2025)
LightGlueStick: a Fast and Robust Glue for Joint Point-Line Matching
by: Ubingazhibov, Aidyn, et al.
Published: (2025)
by: Ubingazhibov, Aidyn, et al.
Published: (2025)
Atlas Gaussians Diffusion for 3D Generation
by: Yang, Haitao, et al.
Published: (2024)
by: Yang, Haitao, et al.
Published: (2024)
SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation
by: Du, Songlin, et al.
Published: (2026)
by: Du, Songlin, et al.
Published: (2026)
TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment
by: Cao, Bingyi, et al.
Published: (2026)
by: Cao, Bingyi, et al.
Published: (2026)
GenCorres: Consistent Shape Matching via Coupled Implicit-Explicit Shape Generative Models
by: Yang, Haitao, et al.
Published: (2023)
by: Yang, Haitao, et al.
Published: (2023)
MapGlue: Multimodal Remote Sensing Image Matching
by: Wu, Peihao, et al.
Published: (2025)
by: Wu, Peihao, et al.
Published: (2025)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
by: Shi, Chen, et al.
Published: (2025)
by: Shi, Chen, et al.
Published: (2025)
XFeat: Accelerated Features for Lightweight Image Matching
by: Potje, Guilherme, et al.
Published: (2024)
by: Potje, Guilherme, et al.
Published: (2024)
Learning Convex Decomposition via Feature Fields
by: Yang, Yuezhi, et al.
Published: (2026)
by: Yang, Yuezhi, et al.
Published: (2026)
To Glue or Not to Glue? Classical vs Learned Image Matching for Mobile Mapping Cameras to Textured Semantic 3D Building Models
by: Gaisbauer, Simone, et al.
Published: (2025)
by: Gaisbauer, Simone, et al.
Published: (2025)
FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation Models
by: Ye, Jianglong, et al.
Published: (2023)
by: Ye, Jianglong, et al.
Published: (2023)
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
by: Liao, Chao, et al.
Published: (2025)
by: Liao, Chao, et al.
Published: (2025)
ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding
by: Guan, Yiran, et al.
Published: (2026)
by: Guan, Yiran, et al.
Published: (2026)
MatChA: Cross-Algorithm Matching with Feature Augmentation
by: Cubero, Paula Carbó, et al.
Published: (2025)
by: Cubero, Paula Carbó, et al.
Published: (2025)
DEFOM-Stereo: Depth Foundation Model Based Stereo Matching
by: Jiang, Hualie, et al.
Published: (2025)
by: Jiang, Hualie, et al.
Published: (2025)
VA-Adapter: Adapting Ultrasound Foundation Model to Echocardiography Probe Guidance
by: Wang, Teng, et al.
Published: (2025)
by: Wang, Teng, et al.
Published: (2025)
FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation
by: Bai, Yunpeng, et al.
Published: (2024)
by: Bai, Yunpeng, et al.
Published: (2024)
Information-Regularized Constrained Inversion for Stable Avatar Editing from Sparse Supervision
by: Liang, Zhenxiao, et al.
Published: (2026)
by: Liang, Zhenxiao, et al.
Published: (2026)
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
by: Liu, Yepeng, et al.
Published: (2025)
by: Liu, Yepeng, et al.
Published: (2025)
Geometry-Guided Modeling of Foundation Features Enables Generalizable Object Shape Deformation Learning
by: Ma, Yiyao, et al.
Published: (2026)
by: Ma, Yiyao, et al.
Published: (2026)
BenchDepth: Are We on the Right Way to Evaluate Depth Foundation Models?
by: Li, Zhenyu, et al.
Published: (2025)
by: Li, Zhenyu, et al.
Published: (2025)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
RayZer: A Self-supervised Large View Synthesis Model
by: Jiang, Hanwen, et al.
Published: (2025)
by: Jiang, Hanwen, et al.
Published: (2025)
Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
Autoregressive Denoising Score Matching is a Good Video Anomaly Detector
by: Zhang, Hanwen, et al.
Published: (2025)
by: Zhang, Hanwen, et al.
Published: (2025)
CFDNet: A Generalizable Foggy Stereo Matching Network with Contrastive Feature Distillation
by: Liu, Zihua, et al.
Published: (2024)
by: Liu, Zihua, et al.
Published: (2024)
RGM: A Robust Generalizable Matching Model
by: Zhang, Songyan, et al.
Published: (2023)
by: Zhang, Songyan, et al.
Published: (2023)
3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography
by: Zhu, Weicheng, et al.
Published: (2025)
by: Zhu, Weicheng, et al.
Published: (2025)
Harnessing Foundation Models for Robust and Generalizable 6-DOF Bronchoscopy Localization
by: Tian, Qingyao, et al.
Published: (2025)
by: Tian, Qingyao, et al.
Published: (2025)
RGBD-Glue: General Feature Combination for Robust RGB-D Point Cloud Registration
by: Chen, Congjia, et al.
Published: (2024)
by: Chen, Congjia, et al.
Published: (2024)
PPLNs: Parametric Piecewise Linear Networks for Event-Based Temporal Modeling and Beyond
by: Song, Chen, et al.
Published: (2024)
by: Song, Chen, et al.
Published: (2024)
Jigsaw++: Imagining Complete Shape Priors for Object Reassembly
by: Lu, Jiaxin, et al.
Published: (2024)
by: Lu, Jiaxin, et al.
Published: (2024)
Similar Items
-
LFM-3D: Learnable Feature Matching Across Wide Baselines Using 3D Signals
by: Karpur, Arjun, et al.
Published: (2023) -
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
by: Jiang, Hanwen, et al.
Published: (2024) -
Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation Models
by: Jiang, Yu, et al.
Published: (2026) -
TIPS: Text-Image Pretraining with Spatial awareness
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024) -
CoFie: Learning Compact Neural Surface Representations with Coordinate Fields
by: Jiang, Hanwen, et al.
Published: (2024)