SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Dünkel, Olaf, Sunagad, Basavaraj, Wang, Haoran, Hoffmann, David T., Theobalt, Christian, Kortylewski, Adam |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Category-Level 3D Correspondence in Camera Space via Morphable Object Priors
by: Sommer, Leonhard, et al.
Published: (2026)
by: Sommer, Leonhard, et al.
Published: (2026)
Do It Yourself: Learning Semantic Correspondence from Pseudo-Labels
by: Dünkel, Olaf, et al.
Published: (2025)
by: Dünkel, Olaf, et al.
Published: (2025)
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence
by: Jesslen, Artur, et al.
Published: (2026)
by: Jesslen, Artur, et al.
Published: (2026)
Common3D: Self-Supervised Learning of 3D Morphable Models for Common Objects in Neural Feature Space
by: Sommer, Leonhard, et al.
Published: (2025)
by: Sommer, Leonhard, et al.
Published: (2025)
TEDRA: Text-based Editing of Dynamic and Photoreal Actors
by: Sunagad, Basavaraj, et al.
Published: (2024)
by: Sunagad, Basavaraj, et al.
Published: (2024)
CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models
by: Begiristain, León, et al.
Published: (2026)
by: Begiristain, León, et al.
Published: (2026)
CNS-Bench: Benchmarking Image Classifier Robustness Under Continuous Nuisance Shifts
by: Dünkel, Olaf, et al.
Published: (2025)
by: Dünkel, Olaf, et al.
Published: (2025)
FaceGPT: Self-supervised Learning to Chat about 3D Human Faces
by: Wang, Haoran, et al.
Published: (2024)
by: Wang, Haoran, et al.
Published: (2024)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
General Neural Gauge Fields
by: Zhan, Fangneng, et al.
Published: (2023)
by: Zhan, Fangneng, et al.
Published: (2023)
DatasetNeRF: Efficient 3D-aware Data Factory with Generative Radiance Fields
by: Chi, Yu, et al.
Published: (2023)
by: Chi, Yu, et al.
Published: (2023)
ASH: Animatable Gaussian Splats for Efficient and Photoreal Human Rendering
by: Pang, Haokai, et al.
Published: (2023)
by: Pang, Haokai, et al.
Published: (2023)
Relightable Neural Actor with Intrinsic Decomposition and Pose Control
by: Luvizon, Diogo, et al.
Published: (2023)
by: Luvizon, Diogo, et al.
Published: (2023)
GRMM: Real-Time High-Fidelity Gaussian Morphable Head Model with Learned Residuals
by: Mendiratta, Mohit, et al.
Published: (2025)
by: Mendiratta, Mohit, et al.
Published: (2025)
Attention (as Discrete-Time Markov) Chains
by: Erel, Yotam, et al.
Published: (2025)
by: Erel, Yotam, et al.
Published: (2025)
PocoLoco: A Point Cloud Diffusion Model of Human Shape in Loose Clothing
by: Seth, Siddharth, et al.
Published: (2024)
by: Seth, Siddharth, et al.
Published: (2024)
Evolutive Rendering Models
by: Zhan, Fangneng, et al.
Published: (2024)
by: Zhan, Fangneng, et al.
Published: (2024)
Every9D-21M: Large-Scale Real-World 9D Canonicalization of Everyday Objects
by: Sommer, Leonhard, et al.
Published: (2026)
by: Sommer, Leonhard, et al.
Published: (2026)
Leveraging Semantic Cues from Foundation Vision Models for Enhanced Local Feature Correspondence
by: Cadar, Felipe, et al.
Published: (2024)
by: Cadar, Felipe, et al.
Published: (2024)
NOVUM: Neural Object Volumes for Robust Object Classification
by: Jesslen, Artur, et al.
Published: (2023)
by: Jesslen, Artur, et al.
Published: (2023)
Unsupervised Learning of Category-Level 3D Pose from Object-Centric Videos
by: Sommer, Leonhard, et al.
Published: (2024)
by: Sommer, Leonhard, et al.
Published: (2024)
Normalizing Flows on the Product Space of SO(3) Manifolds for Probabilistic Human Pose Modeling
by: Dünkel, Olaf, et al.
Published: (2024)
by: Dünkel, Olaf, et al.
Published: (2024)
Interpretable 3D Neural Object Volumes for Robust Conceptual Reasoning
by: Pham, Nhi, et al.
Published: (2025)
by: Pham, Nhi, et al.
Published: (2025)
Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose Estimation
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
Independently Keypoint Learning for Small Object Semantic Correspondence
by: Jin, Hailong, et al.
Published: (2024)
by: Jin, Hailong, et al.
Published: (2024)
PractiLight: Practical Light Control Using Foundational Diffusion Models
by: Erel, Yotam, et al.
Published: (2025)
by: Erel, Yotam, et al.
Published: (2025)
Semantic Correspondence: Unified Benchmarking and a Strong Baseline
by: Zhang, Kaiyan, et al.
Published: (2025)
by: Zhang, Kaiyan, et al.
Published: (2025)
Towards Robust Semantic Correspondence: A Benchmark and Insights
by: Chong, Wenyue
Published: (2025)
by: Chong, Wenyue
Published: (2025)
Learning a Category-level Object Pose Estimator without Pose Annotations
by: Tian, Fengrui, et al.
Published: (2024)
by: Tian, Fengrui, et al.
Published: (2024)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
by: Gopal, Suhas, et al.
Published: (2025)
by: Gopal, Suhas, et al.
Published: (2025)
Shakti-VLMs: Scalable Vision-Language Models for Enterprise AI
by: Shakhadri, Syed Abdul Gaffar, et al.
Published: (2025)
by: Shakhadri, Syed Abdul Gaffar, et al.
Published: (2025)
Vector-Quantized Vision Foundation Models for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2025)
by: Zhao, Rongzhen, et al.
Published: (2025)
GLASS: Graph and Vision-Language Assisted Semantic Shape Correspondence
by: Xiao, Qinfeng, et al.
Published: (2026)
by: Xiao, Qinfeng, et al.
Published: (2026)
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
by: Aytekin, Ayce Idil, et al.
Published: (2025)
by: Aytekin, Ayce Idil, et al.
Published: (2025)
Sample-Specific Output Constraints for Neural Networks
by: Brosowsky, Mathis, et al.
Published: (2020)
by: Brosowsky, Mathis, et al.
Published: (2020)
Benchmarking Computational Pathology Foundation Models For Semantic Segmentation
by: Ramchandani, Lavish, et al.
Published: (2026)
by: Ramchandani, Lavish, et al.
Published: (2026)
Non-Rigid 3D Shape Correspondences: From Foundations to Open Challenges and Opportunities
by: Zhuravlev, Aleksei, et al.
Published: (2026)
by: Zhuravlev, Aleksei, et al.
Published: (2026)
Annotation Free Semantic Segmentation with Vision Foundation Models
by: Seifi, Soroush, et al.
Published: (2024)
by: Seifi, Soroush, et al.
Published: (2024)
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
by: Lee, Hansang, et al.
Published: (2017)
by: Lee, Hansang, et al.
Published: (2017)
A Bayesian Approach to OOD Robustness in Image Classification
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
Similar Items
-
Category-Level 3D Correspondence in Camera Space via Morphable Object Priors
by: Sommer, Leonhard, et al.
Published: (2026) -
Do It Yourself: Learning Semantic Correspondence from Pseudo-Labels
by: Dünkel, Olaf, et al.
Published: (2025) -
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence
by: Jesslen, Artur, et al.
Published: (2026) -
Common3D: Self-Supervised Learning of 3D Morphable Models for Common Objects in Neural Feature Space
by: Sommer, Leonhard, et al.
Published: (2025) -
TEDRA: Text-based Editing of Dynamic and Photoreal Actors
by: Sunagad, Basavaraj, et al.
Published: (2024)