GIM: Learning Generalizable Image Matcher From Internet Videos
Fuente:
arXiv
Salvato in:
| Autori principali: | Shen, Xuelun, Cai, Zhipeng, Yin, Wei, Müller, Matthias, Li, Zijun, Wang, Kaixuan, Chen, Xiaozhi, Wang, Cheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ConDo: Continual Domain Expansion for Absolute Pose Regression
di: Li, Zijun, et al.
Pubblicazione: (2024)
di: Li, Zijun, et al.
Pubblicazione: (2024)
EndoMatcher: Generalizable Endoscopic Image Matcher via Multi-Domain Pre-training for Robot-Assisted Surgery
di: Yang, Bingyu, et al.
Pubblicazione: (2025)
di: Yang, Bingyu, et al.
Pubblicazione: (2025)
A Detector-oblivious Multi-arm Network for Keypoint Matching
di: Shen, Xuelun, et al.
Pubblicazione: (2021)
di: Shen, Xuelun, et al.
Pubblicazione: (2021)
GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization
di: Chen, Yirui, et al.
Pubblicazione: (2024)
di: Chen, Yirui, et al.
Pubblicazione: (2024)
Adaptive Fusion of Single-View and Multi-View Depth for Autonomous Driving
di: Cheng, JunDa, et al.
Pubblicazione: (2024)
di: Cheng, JunDa, et al.
Pubblicazione: (2024)
EvalGIM: A Library for Evaluating Generative Image Models
di: Hall, Melissa, et al.
Pubblicazione: (2024)
di: Hall, Melissa, et al.
Pubblicazione: (2024)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
di: Liu, Yang, et al.
Pubblicazione: (2023)
di: Liu, Yang, et al.
Pubblicazione: (2023)
Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation
di: Hu, Mu, et al.
Pubblicazione: (2024)
di: Hu, Mu, et al.
Pubblicazione: (2024)
ImageAttributionBench: How Far Are We from Generalizable Attribution?
di: Mou, Tingshu, et al.
Pubblicazione: (2026)
di: Mou, Tingshu, et al.
Pubblicazione: (2026)
RoMeO: Robust Metric Visual Odometry
di: Cheng, Junda, et al.
Pubblicazione: (2024)
di: Cheng, Junda, et al.
Pubblicazione: (2024)
Dense Matchers for Dense Tracking
di: Jelínek, Tomáš, et al.
Pubblicazione: (2024)
di: Jelínek, Tomáš, et al.
Pubblicazione: (2024)
Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?
di: Corley, Isaac, et al.
Pubblicazione: (2026)
di: Corley, Isaac, et al.
Pubblicazione: (2026)
GGS: Generalizable Gaussian Splatting for Lane Switching in Autonomous Driving
di: Han, Huasong, et al.
Pubblicazione: (2024)
di: Han, Huasong, et al.
Pubblicazione: (2024)
Low-Contrast-Enhanced Contrastive Learning for Semi-Supervised Endoscopic Image Segmentation
di: Cai, Lingcong, et al.
Pubblicazione: (2024)
di: Cai, Lingcong, et al.
Pubblicazione: (2024)
Better Matching, Less Forgetting: A Quality-Guided Matcher for Transformer-based Incremental Object Detection
di: Wu, Qirui, et al.
Pubblicazione: (2026)
di: Wu, Qirui, et al.
Pubblicazione: (2026)
DreamMatcher: Appearance Matching Self-Attention for Semantically-Consistent Text-to-Image Personalization
di: Nam, Jisu, et al.
Pubblicazione: (2024)
di: Nam, Jisu, et al.
Pubblicazione: (2024)
Texture-aware Intrinsic Image Decomposition with Model- and Learning-based Priors
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
Learning Universal Features for Generalizable Image Forgery Localization
di: Zhao, Hengrun, et al.
Pubblicazione: (2025)
di: Zhao, Hengrun, et al.
Pubblicazione: (2025)
AffordMatcher: Affordance Learning in 3D Scenes from Visual Signifiers
di: Vu, Nghia, et al.
Pubblicazione: (2026)
di: Vu, Nghia, et al.
Pubblicazione: (2026)
HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation
di: Wang, Xiaolong, et al.
Pubblicazione: (2024)
di: Wang, Xiaolong, et al.
Pubblicazione: (2024)
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis
di: Yin, Shicheng, et al.
Pubblicazione: (2024)
di: Yin, Shicheng, et al.
Pubblicazione: (2024)
Monte Carlo Diffusion for Generalizable Learning-Based RANSAC
di: Wang, Jiale, et al.
Pubblicazione: (2025)
di: Wang, Jiale, et al.
Pubblicazione: (2025)
Frequency-based Matcher for Long-tailed Semantic Segmentation
di: Li, Shan, et al.
Pubblicazione: (2024)
di: Li, Shan, et al.
Pubblicazione: (2024)
Generalizable Implicit Motion Modeling for Video Frame Interpolation
di: Guo, Zujin, et al.
Pubblicazione: (2024)
di: Guo, Zujin, et al.
Pubblicazione: (2024)
ShapeMatcher: Self-Supervised Joint Shape Canonicalization, Segmentation, Retrieval and Deformation
di: Di, Yan, et al.
Pubblicazione: (2023)
di: Di, Yan, et al.
Pubblicazione: (2023)
Generalizable and Adaptive Continual Learning Framework for AI-generated Image Detection
di: Wang, Hanyi, et al.
Pubblicazione: (2026)
di: Wang, Hanyi, et al.
Pubblicazione: (2026)
Generalizable Two-Branch Framework for Image Class-Incremental Learning
di: Wu, Chao, et al.
Pubblicazione: (2024)
di: Wu, Chao, et al.
Pubblicazione: (2024)
From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos
di: Chen, Yin, et al.
Pubblicazione: (2023)
di: Chen, Yin, et al.
Pubblicazione: (2023)
SAIDO: Generalizable Detection of AI-Generated Images via Scene-Aware and Importance-Guided Dynamic Optimization in Continual Learning
di: Hu, Yongkang, et al.
Pubblicazione: (2025)
di: Hu, Yongkang, et al.
Pubblicazione: (2025)
RealRestorer: Towards Generalizable Real-World Image Restoration with Large-Scale Image Editing Models
di: Yang, Yufeng, et al.
Pubblicazione: (2026)
di: Yang, Yufeng, et al.
Pubblicazione: (2026)
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2025)
di: Ma, Junyi, et al.
Pubblicazione: (2025)
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
di: Wang, Hanlin, et al.
Pubblicazione: (2024)
di: Wang, Hanlin, et al.
Pubblicazione: (2024)
GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
di: Fu, Xiao, et al.
Pubblicazione: (2024)
di: Fu, Xiao, et al.
Pubblicazione: (2024)
Generalizable Video Quality Assessment via Weak-to-Strong Learning
di: Cao, Linhan, et al.
Pubblicazione: (2025)
di: Cao, Linhan, et al.
Pubblicazione: (2025)
MotionMatcher: Motion Customization of Text-to-Video Diffusion Models via Motion Feature Matching
di: Wu, Yen-Siang, et al.
Pubblicazione: (2025)
di: Wu, Yen-Siang, et al.
Pubblicazione: (2025)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
di: Wang, Chaoyang, et al.
Pubblicazione: (2024)
DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo
di: Zhu, Junzhe, et al.
Pubblicazione: (2024)
di: Zhu, Junzhe, et al.
Pubblicazione: (2024)
Exploit the Leak: Understanding Risks in Biometric Matchers
di: Durbet, Axel, et al.
Pubblicazione: (2023)
di: Durbet, Axel, et al.
Pubblicazione: (2023)
Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video
di: Wang, Yifan, et al.
Pubblicazione: (2026)
di: Wang, Yifan, et al.
Pubblicazione: (2026)
GenS: Generalizable Neural Surface Reconstruction from Multi-View Images
di: Peng, Rui, et al.
Pubblicazione: (2024)
di: Peng, Rui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ConDo: Continual Domain Expansion for Absolute Pose Regression
di: Li, Zijun, et al.
Pubblicazione: (2024) -
EndoMatcher: Generalizable Endoscopic Image Matcher via Multi-Domain Pre-training for Robot-Assisted Surgery
di: Yang, Bingyu, et al.
Pubblicazione: (2025) -
A Detector-oblivious Multi-arm Network for Keypoint Matching
di: Shen, Xuelun, et al.
Pubblicazione: (2021) -
GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization
di: Chen, Yirui, et al.
Pubblicazione: (2024) -
Adaptive Fusion of Single-View and Multi-View Depth for Autonomous Driving
di: Cheng, JunDa, et al.
Pubblicazione: (2024)