Multi-View Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Segre, Leo, Hirschorn, Or, Avidan, Shai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Graph-Based Approach for Category-Agnostic Pose Estimation
by: Hirschorn, Or, et al.
Published: (2023)
by: Hirschorn, Or, et al.
Published: (2023)
Edge Weight Prediction For Category-Agnostic Pose Estimation
by: Hirschorn, Or, et al.
Published: (2024)
by: Hirschorn, Or, et al.
Published: (2024)
Optimize the Unseen -- Fast NeRF Cleanup with Free Space Prior
by: Segre, Leo, et al.
Published: (2024)
by: Segre, Leo, et al.
Published: (2024)
VF-NeRF: Viewshed Fields for Rigid NeRF Registration
by: Segre, Leo, et al.
Published: (2024)
by: Segre, Leo, et al.
Published: (2024)
CapeX: Category-Agnostic Pose Estimation from Textual Point Explanation
by: Rusanovsky, Matan, et al.
Published: (2024)
by: Rusanovsky, Matan, et al.
Published: (2024)
Frequency-Aware Gaussian Splatting Decomposition
by: Lavi, Yishai, et al.
Published: (2025)
by: Lavi, Yishai, et al.
Published: (2025)
Scene Grounding In the Wild
by: Cohen, Tamir, et al.
Published: (2026)
by: Cohen, Tamir, et al.
Published: (2026)
Securing Neural Networks with Knapsack Optimization
by: Gorski, Yakir, et al.
Published: (2023)
by: Gorski, Yakir, et al.
Published: (2023)
TextureSAM: Towards a Texture Aware Foundation Model for Segmentation
by: Cohen, Inbal, et al.
Published: (2025)
by: Cohen, Inbal, et al.
Published: (2025)
Talking Points: Describing and Localizing Pixels
by: Rusanovsky, Matan, et al.
Published: (2025)
by: Rusanovsky, Matan, et al.
Published: (2025)
Coordinate Descent for Network Linearization
by: Rakhlin, Vlad, et al.
Published: (2025)
by: Rakhlin, Vlad, et al.
Published: (2025)
Memories of Forgotten Concepts
by: Rusanovsky, Matan, et al.
Published: (2024)
by: Rusanovsky, Matan, et al.
Published: (2024)
Splatent: Splatting Diffusion Latents for Novel View Synthesis
by: Hirschorn, Or, et al.
Published: (2025)
by: Hirschorn, Or, et al.
Published: (2025)
Lightning-Fast Image Inversion and Editing for Text-to-Image Diffusion Models
by: Samuel, Dvir, et al.
Published: (2023)
by: Samuel, Dvir, et al.
Published: (2023)
Toward a Multi-View Brain Network Foundation Model: Cross-View Consistency Learning Across Arbitrary Atlases
by: Xu, Jiaxing, et al.
Published: (2026)
by: Xu, Jiaxing, et al.
Published: (2026)
Boosting Multi-View Stereo with Depth Foundation Model in the Absence of Real-World Labels
by: Zhu, Jie, et al.
Published: (2025)
by: Zhu, Jie, et al.
Published: (2025)
Decoding Functional Networks for Visual Categories via GNNs
by: Karmi, Shira, et al.
Published: (2026)
by: Karmi, Shira, et al.
Published: (2026)
M^3: Dense Matching Meets Multi-View Foundation Models for Monocular Gaussian Splatting SLAM
by: Ren, Kerui, et al.
Published: (2026)
by: Ren, Kerui, et al.
Published: (2026)
Emergent Extreme-View Geometry in 3D Foundation Models
by: Zhang, Yiwen, et al.
Published: (2025)
by: Zhang, Yiwen, et al.
Published: (2025)
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
by: Bozic, Vukasin, et al.
Published: (2026)
by: Bozic, Vukasin, et al.
Published: (2026)
Multi-View Consistent Wound Segmentation With Neural Fields
by: Chierchia, Remi, et al.
Published: (2026)
by: Chierchia, Remi, et al.
Published: (2026)
GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization
by: Song, Zixuan, et al.
Published: (2025)
by: Song, Zixuan, et al.
Published: (2025)
Hearing the Room Through the Shape of the Drum: Modal-Guided Sound Recovery from Multi-Point Surface Vibrations
by: Bagon, Shai, et al.
Published: (2026)
by: Bagon, Shai, et al.
Published: (2026)
Semi-Supervised Multi-View Crowd Counting by Ranking Multi-View Fusion Models
by: Zhang, Qi, et al.
Published: (2025)
by: Zhang, Qi, et al.
Published: (2025)
Adapting Foundation Model for Dental Caries Detection with Dual-View Co-Training
by: Luo, Tao, et al.
Published: (2025)
by: Luo, Tao, et al.
Published: (2025)
Probing Fine-Grained Action Understanding and Cross-View Generalization of Foundation Models
by: Ponbagavathi, Thinesh Thiyakesan, et al.
Published: (2024)
by: Ponbagavathi, Thinesh Thiyakesan, et al.
Published: (2024)
T-MASK: Temporal Masking for Probing Foundation Models across Camera Views in Driver Monitoring
by: Ponbagavathi, Thinesh Thiyakesan, et al.
Published: (2025)
by: Ponbagavathi, Thinesh Thiyakesan, et al.
Published: (2025)
Curia: A Multi-Modal Foundation Model for Radiology
by: Dancette, Corentin, et al.
Published: (2025)
by: Dancette, Corentin, et al.
Published: (2025)
MultiWorld: Scalable Multi-Agent Multi-View Video World Models
by: Wu, Haoyu, et al.
Published: (2026)
by: Wu, Haoyu, et al.
Published: (2026)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
by: Li, Jinke, et al.
Published: (2024)
by: Li, Jinke, et al.
Published: (2024)
RangeSAM: On the Potential of Visual Foundation Models for Range-View represented LiDAR segmentation
by: Kühn, Paul Julius, et al.
Published: (2025)
by: Kühn, Paul Julius, et al.
Published: (2025)
MVSMamba: Multi-View Stereo with State Space Model
by: Jiang, Jianfei, et al.
Published: (2025)
by: Jiang, Jianfei, et al.
Published: (2025)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
Evaluating Foundation Models' 3D Understanding Through Multi-View Correspondence Analysis
by: Lilova, Valentina, et al.
Published: (2025)
by: Lilova, Valentina, et al.
Published: (2025)
InstructMix2Mix: Consistent Sparse-View Editing Through Multi-View Model Personalization
by: Gilo, Daniel, et al.
Published: (2025)
by: Gilo, Daniel, et al.
Published: (2025)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
by: Zhu, Ruishu, et al.
Published: (2025)
by: Zhu, Ruishu, et al.
Published: (2025)
Revisiting Birds Eye View Perception Models with Frozen Foundation Models: DINOv2 and Metric3Dv2
by: Hayes, Seamie, et al.
Published: (2025)
by: Hayes, Seamie, et al.
Published: (2025)
Multi-View Factorizing and Disentangling: A Novel Framework for Incomplete Multi-View Multi-Label Classification
by: Xie, Wulin, et al.
Published: (2025)
by: Xie, Wulin, et al.
Published: (2025)
Repurposing Geometric Foundation Models for Multi-view Diffusion
by: Jang, Wooseok, et al.
Published: (2026)
by: Jang, Wooseok, et al.
Published: (2026)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025)
by: Kwon, Minkyung, et al.
Published: (2025)
Similar Items
-
A Graph-Based Approach for Category-Agnostic Pose Estimation
by: Hirschorn, Or, et al.
Published: (2023) -
Edge Weight Prediction For Category-Agnostic Pose Estimation
by: Hirschorn, Or, et al.
Published: (2024) -
Optimize the Unseen -- Fast NeRF Cleanup with Free Space Prior
by: Segre, Leo, et al.
Published: (2024) -
VF-NeRF: Viewshed Fields for Rigid NeRF Registration
by: Segre, Leo, et al.
Published: (2024) -
CapeX: Category-Agnostic Pose Estimation from Textual Point Explanation
by: Rusanovsky, Matan, et al.
Published: (2024)