Do Foundation Models Know Geometry? Probing Frozen Features for Continuous Physical Measurement
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Shkolnikov, Yakov Pyotr |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bit-Identical Medical Deep Learning via Structured Orthogonal Initialization
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
Agent Memory Below the Prompt: Persistent Q4 KV Cache for Multi-Agent LLM Inference on Edge Devices
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
von: Du, Xiaodan, et al.
Veröffentlicht: (2023)
von: Du, Xiaodan, et al.
Veröffentlicht: (2023)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
Do Vision-Language Models Measure Up? Benchmarking Visual Measurement Reading with MeasureBench
von: Lin, Fenfen, et al.
Veröffentlicht: (2025)
von: Lin, Fenfen, et al.
Veröffentlicht: (2025)
How Much 3D Do Video Foundation Models Encode?
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks
von: Guo, Yachan, et al.
Veröffentlicht: (2026)
von: Guo, Yachan, et al.
Veröffentlicht: (2026)
Beyond Pixels: Enhancing LIME with Hierarchical Features and Segmentation Foundation Models
von: Knab, Patrick, et al.
Veröffentlicht: (2024)
von: Knab, Patrick, et al.
Veröffentlicht: (2024)
Training Frozen Feature Pyramid DINOv2 for Eyelid Measurements with Infinite Encoding and Orthogonal Regularization
von: Chen, Chun-Hung
Veröffentlicht: (2025)
von: Chen, Chun-Hung
Veröffentlicht: (2025)
Foundation Models and Adaptive Feature Selection: A Synergistic Approach to Video Question Answering
von: Rongali, Sai Bhargav, et al.
Veröffentlicht: (2024)
von: Rongali, Sai Bhargav, et al.
Veröffentlicht: (2024)
Feature Quality and Adaptability of Medical Foundation Models: A Comparative Evaluation for Radiographic Classification and Segmentation
von: Li, Frank, et al.
Veröffentlicht: (2025)
von: Li, Frank, et al.
Veröffentlicht: (2025)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
Your Pre-trained Diffusion Model Secretly Knows Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
Mining and Transferring Feature-Geometry Coherence for Unsupervised Point Cloud Registration
von: Xiong, Kezheng, et al.
Veröffentlicht: (2024)
von: Xiong, Kezheng, et al.
Veröffentlicht: (2024)
From Linear Probing to Joint-Weighted Token Hierarchy: A Foundation Model Bridging Global and Cellular Representations in Biomarker Detection
von: Liu, Jingsong, et al.
Veröffentlicht: (2025)
von: Liu, Jingsong, et al.
Veröffentlicht: (2025)
DINOv3 as a Frozen Encoder for CRPS-Oriented Probabilistic Rainfall Nowcasting
von: Filho, Luciano Araujo Dourado, et al.
Veröffentlicht: (2025)
von: Filho, Luciano Araujo Dourado, et al.
Veröffentlicht: (2025)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
ConDo: Continual Domain Expansion for Absolute Pose Regression
von: Li, Zijun, et al.
Veröffentlicht: (2024)
von: Li, Zijun, et al.
Veröffentlicht: (2024)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
von: Boduljak, Gabrijel, et al.
Veröffentlicht: (2025)
von: Boduljak, Gabrijel, et al.
Veröffentlicht: (2025)
Proximal Vision Transformer: Enhancing Feature Representation through Two-Stage Manifold Geometry
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
PhysicsMind: Sim and Real Mechanics Benchmarking for Physical Reasoning and Prediction in Foundational VLMs and World Models
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models
von: Ortu, Francesco, et al.
Veröffentlicht: (2025)
von: Ortu, Francesco, et al.
Veröffentlicht: (2025)
UniSino: Physics-Driven Foundational Model for Universal CT Sinogram Standardization
von: Ai, Xingyu, et al.
Veröffentlicht: (2025)
von: Ai, Xingyu, et al.
Veröffentlicht: (2025)
FROST-Drive: Scalable and Efficient End-to-End Driving with a Frozen Vision Encoder
von: Dong, Zeyu, et al.
Veröffentlicht: (2026)
von: Dong, Zeyu, et al.
Veröffentlicht: (2026)
Are Compact Rationales Free? Measuring Tile Selection Headroom in Frozen WSI-MIL
von: Jung, Hyun Do, et al.
Veröffentlicht: (2026)
von: Jung, Hyun Do, et al.
Veröffentlicht: (2026)
Artificial Intelligence for Geometry-Based Feature Extraction, Analysis and Synthesis in Artistic Images: A Survey
von: Vijendran, Mridula, et al.
Veröffentlicht: (2024)
von: Vijendran, Mridula, et al.
Veröffentlicht: (2024)
Action Without Interaction: Probing the Physical Foundations of Video LMMs via Contact-Release Detection
von: Harari, Daniel, et al.
Veröffentlicht: (2025)
von: Harari, Daniel, et al.
Veröffentlicht: (2025)
Continual Novel Class Discovery via Feature Enhancement and Adaptation
von: Yu, Yifan, et al.
Veröffentlicht: (2024)
von: Yu, Yifan, et al.
Veröffentlicht: (2024)
UP-Fuse: Uncertainty-guided LiDAR-Camera Fusion for 3D Panoptic Segmentation
von: Mohan, Rohit, et al.
Veröffentlicht: (2026)
von: Mohan, Rohit, et al.
Veröffentlicht: (2026)
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
Frozen Forecasting: A Unified Evaluation
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
Detached Skip-Links and $R$-Probe: Decoupling Feature Aggregation from Gradient Propagation for MLLM OCR
von: Yuan, Ziye, et al.
Veröffentlicht: (2026)
von: Yuan, Ziye, et al.
Veröffentlicht: (2026)
DINO-VO: A Feature-based Visual Odometry Leveraging a Visual Foundation Model
von: Azhari, Maulana Bisyir, et al.
Veröffentlicht: (2025)
von: Azhari, Maulana Bisyir, et al.
Veröffentlicht: (2025)
Three Things to Know about Deep Metric Learning
von: Patel, Yash, et al.
Veröffentlicht: (2024)
von: Patel, Yash, et al.
Veröffentlicht: (2024)
Continual Gesture Learning without Data via Synthetic Feature Sampling
von: Lu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Lu, Zhenyu, et al.
Veröffentlicht: (2024)
Few-Shot Segmentation of Historical Maps via Linear Probing of Vision Foundation Models
von: Sterzinger, Rafael, et al.
Veröffentlicht: (2025)
von: Sterzinger, Rafael, et al.
Veröffentlicht: (2025)
The Geometry of Representational Failures in Vision Language Models
von: Savietto, Daniele, et al.
Veröffentlicht: (2026)
von: Savietto, Daniele, et al.
Veröffentlicht: (2026)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
Probing Visual Planning in Image Editing Models
von: Zhou, Zhimu, et al.
Veröffentlicht: (2026)
von: Zhou, Zhimu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Bit-Identical Medical Deep Learning via Structured Orthogonal Initialization
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026) -
Agent Memory Below the Prompt: Persistent Q4 KV Cache for Multi-Agent LLM Inference on Edge Devices
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026) -
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
von: Du, Xiaodan, et al.
Veröffentlicht: (2023) -
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
von: Zhang, Yue, et al.
Veröffentlicht: (2026) -
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)