FlatVPR: Plug-and-play Geo-linear Residual Adapter for Geometric Rectification of Foundation Model Feature Manifolds
Fuente:
arXiv
Saved in:
| Main Authors: | Hisada, Rai, Tanaka, Kanji |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Recursive Distillation for Open-Set Distributed Robot Localization
by: Tsukahara, Kenta, et al.
Published: (2023)
by: Tsukahara, Kenta, et al.
Published: (2023)
CON: Continual Object Navigation via Data-Free Inter-Agent Knowledge Transfer in Unseen and Unfamiliar Places
by: Terashima, Kouki, et al.
Published: (2024)
by: Terashima, Kouki, et al.
Published: (2024)
Training Self-localization Models for Unseen Unfamiliar Places via Teacher-to-Student Data-Free Knowledge Transfer
by: Tsukahara, Kenta, et al.
Published: (2024)
by: Tsukahara, Kenta, et al.
Published: (2024)
Hyperspectral Adapter for Semantic Segmentation with Vision Foundation Models
by: Hurtado, Juana Valeria, et al.
Published: (2025)
by: Hurtado, Juana Valeria, et al.
Published: (2025)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
by: Zhang, Zezhou, et al.
Published: (2026)
by: Zhang, Zezhou, et al.
Published: (2026)
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
by: Lu, Ziqi, et al.
Published: (2024)
by: Lu, Ziqi, et al.
Published: (2024)
A Dual-Stream Transformer Architecture for Illumination-Invariant TIR-LiDAR Person Tracking
by: Minase, Yuki, et al.
Published: (2026)
by: Minase, Yuki, et al.
Published: (2026)
Learning Visual Feature-Based World Models via Residual Latent Action
by: Zhang, Xinyu, et al.
Published: (2026)
by: Zhang, Xinyu, et al.
Published: (2026)
A Survey for Foundation Models in Autonomous Driving
by: Gao, Haoxiang, et al.
Published: (2024)
by: Gao, Haoxiang, et al.
Published: (2024)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
by: Zhang, Junjie, et al.
Published: (2024)
by: Zhang, Junjie, et al.
Published: (2024)
GAGrasp: Geometric Algebra Diffusion for Dexterous Grasping
by: Zhong, Tao, et al.
Published: (2025)
by: Zhong, Tao, et al.
Published: (2025)
FoundationStereo: Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark
by: Dang, Ronghao, et al.
Published: (2025)
by: Dang, Ronghao, et al.
Published: (2025)
Deep SE(3)-Equivariant Geometric Reasoning for Precise Placement Tasks
by: Eisner, Ben, et al.
Published: (2024)
by: Eisner, Ben, et al.
Published: (2024)
See Less, Drive Better: Generalizable End-to-End Autonomous Driving via Foundation Models Stochastic Patch Selection
by: Mallak, Amir, et al.
Published: (2026)
by: Mallak, Amir, et al.
Published: (2026)
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
by: Ni, Zehao, et al.
Published: (2025)
by: Ni, Zehao, et al.
Published: (2025)
EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
Learning Generalizable Feature Fields for Mobile Manipulation
by: Qiu, Ri-Zhao, et al.
Published: (2024)
by: Qiu, Ri-Zhao, et al.
Published: (2024)
Aether: Geometric-Aware Unified World Modeling
by: Aether Team, et al.
Published: (2025)
by: Aether Team, et al.
Published: (2025)
SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision
by: Rai, Utsav, et al.
Published: (2025)
by: Rai, Utsav, et al.
Published: (2025)
Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments
by: Shao, Xingyu, et al.
Published: (2026)
by: Shao, Xingyu, et al.
Published: (2026)
GraspSplats: Efficient Manipulation with 3D Feature Splatting
by: Ji, Mazeyu, et al.
Published: (2024)
by: Ji, Mazeyu, et al.
Published: (2024)
AirIO: Learning Inertial Odometry with Enhanced IMU Feature Observability
by: Qiu, Yuheng, et al.
Published: (2025)
by: Qiu, Yuheng, et al.
Published: (2025)
Nothing Stands Still: A Spatiotemporal Benchmark on 3D Point Cloud Registration Under Large Geometric and Temporal Change
by: Sun, Tao, et al.
Published: (2023)
by: Sun, Tao, et al.
Published: (2023)
GNFactor: Multi-Task Real Robot Learning with Generalizable Neural Feature Fields
by: Ze, Yanjie, et al.
Published: (2023)
by: Ze, Yanjie, et al.
Published: (2023)
Exploring Transformer-Augmented LSTM for Temporal and Spatial Feature Learning in Trajectory Prediction
by: Raskoti, Chandra, et al.
Published: (2024)
by: Raskoti, Chandra, et al.
Published: (2024)
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
by: Gu, Xunjiang, et al.
Published: (2024)
by: Gu, Xunjiang, et al.
Published: (2024)
SpeedUpNet: A Plug-and-Play Adapter Network for Accelerating Text-to-Image Diffusion Models
by: Chai, Weilong, et al.
Published: (2023)
by: Chai, Weilong, et al.
Published: (2023)
How to Benchmark Vision Foundation Models for Semantic Segmentation?
by: Kerssies, Tommie, et al.
Published: (2024)
by: Kerssies, Tommie, et al.
Published: (2024)
Cosmos World Foundation Model Platform for Physical AI
by: NVIDIA, et al.
Published: (2025)
by: NVIDIA, et al.
Published: (2025)
World Simulation with Video Foundation Models for Physical AI
by: NVIDIA, et al.
Published: (2025)
by: NVIDIA, et al.
Published: (2025)
Learning Part-Aware Dense 3D Feature Field for Generalizable Articulated Object Manipulation
by: Chen, Yue, et al.
Published: (2026)
by: Chen, Yue, et al.
Published: (2026)
TK-Planes: Tiered K-Planes with High Dimensional Feature Vectors for Dynamic UAV-based Scenes
by: Maxey, Christopher, et al.
Published: (2024)
by: Maxey, Christopher, et al.
Published: (2024)
Real-World Robot Applications of Foundation Models: A Review
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
Towards Natural Language-Driven Assembly Using Foundation Models
by: Joglekar, Omkar, et al.
Published: (2024)
by: Joglekar, Omkar, et al.
Published: (2024)
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
by: Shang, Jinghuan, et al.
Published: (2024)
by: Shang, Jinghuan, et al.
Published: (2024)
UniFField: A Generalizable Unified Neural Feature Field for Visual, Semantic, and Spatial Uncertainties in Any Scene
by: Maurer, Christian, et al.
Published: (2025)
by: Maurer, Christian, et al.
Published: (2025)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
by: Liu, Songming, et al.
Published: (2024)
by: Liu, Songming, et al.
Published: (2024)
Advances in Multimodal Adaptation and Generalization: From Traditional Approaches to Foundation Models
by: Dong, Hao, et al.
Published: (2025)
by: Dong, Hao, et al.
Published: (2025)
TraIL-Det: Transformation-Invariant Local Feature Networks for 3D LiDAR Object Detection with Unsupervised Pre-Training
by: Li, Li, et al.
Published: (2024)
by: Li, Li, et al.
Published: (2024)
Similar Items
-
Recursive Distillation for Open-Set Distributed Robot Localization
by: Tsukahara, Kenta, et al.
Published: (2023) -
CON: Continual Object Navigation via Data-Free Inter-Agent Knowledge Transfer in Unseen and Unfamiliar Places
by: Terashima, Kouki, et al.
Published: (2024) -
Training Self-localization Models for Unseen Unfamiliar Places via Teacher-to-Student Data-Free Knowledge Transfer
by: Tsukahara, Kenta, et al.
Published: (2024) -
Hyperspectral Adapter for Semantic Segmentation with Vision Foundation Models
by: Hurtado, Juana Valeria, et al.
Published: (2025) -
PVI: Plug-in Visual Injection for Vision-Language-Action Models
by: Zhang, Zezhou, et al.
Published: (2026)