C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Kuan Wei, Li, Brandon, Hariharan, Bharath, Snavely, Noah |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
G3T Up! Gravity Aligned Coordinate Frames Simplify Pointmap Processing
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024)
by: Tung, Joseph, et al.
Published: (2024)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
by: Chou, Gene, et al.
Published: (2024)
by: Chou, Gene, et al.
Published: (2024)
Pointmap-Conditioned Diffusion for Consistent Novel View Synthesis
by: Nguyen, Thang-Anh-Quan, et al.
Published: (2025)
by: Nguyen, Thang-Anh-Quan, et al.
Published: (2025)
FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution
by: Chou, Gene, et al.
Published: (2025)
by: Chou, Gene, et al.
Published: (2025)
Learning Cross-View Object Correspondence via Cycle-Consistent Mask Prediction
by: Yan, Shannan, et al.
Published: (2026)
by: Yan, Shannan, et al.
Published: (2026)
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
by: Kim, Seungwook, et al.
Published: (2024)
by: Kim, Seungwook, et al.
Published: (2024)
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
by: Chetan, Aditya, et al.
Published: (2026)
by: Chetan, Aditya, et al.
Published: (2026)
MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
by: Chou, Gene, et al.
Published: (2026)
by: Chou, Gene, et al.
Published: (2026)
Towards Cross-View Point Correspondence in Vision-Language Models
by: Wang, Yipu, et al.
Published: (2025)
by: Wang, Yipu, et al.
Published: (2025)
Cross-View Completion Models are Zero-shot Correspondence Estimators
by: An, Honggyu, et al.
Published: (2024)
by: An, Honggyu, et al.
Published: (2024)
CLNet: Cross-View Correspondence Makes a Stronger Geo-Localizationer
by: Cao, Xianwei, et al.
Published: (2025)
by: Cao, Xianwei, et al.
Published: (2025)
Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality Signals
by: Fang, Shaoheng, et al.
Published: (2024)
by: Fang, Shaoheng, et al.
Published: (2024)
Wide-Baseline Relative Camera Pose Estimation with Directional Learning
by: Chen, Kefan, et al.
Published: (2021)
by: Chen, Kefan, et al.
Published: (2021)
Range and Bird's Eye View Fused Cross-Modal Visual Place Recognition
by: Peng, Jianyi, et al.
Published: (2025)
by: Peng, Jianyi, et al.
Published: (2025)
ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild
by: Chen, Hanyu, et al.
Published: (2026)
by: Chen, Hanyu, et al.
Published: (2026)
UniABG: Unified Adversarial View Bridging and Graph Correspondence for Unsupervised Cross-View Geo-Localization
by: Chen, Cuiqun, et al.
Published: (2025)
by: Chen, Cuiqun, et al.
Published: (2025)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
by: Li, Weijia, et al.
Published: (2024)
by: Li, Weijia, et al.
Published: (2024)
Generative Image Dynamics
by: Li, Zhengqi, et al.
Published: (2023)
by: Li, Zhengqi, et al.
Published: (2023)
ACIT: Attention-Guided Cross-Modal Interaction Transformer for Pedestrian Crossing Intention Prediction
by: Li, Yuanzhe, et al.
Published: (2025)
by: Li, Yuanzhe, et al.
Published: (2025)
PAUL: Uncertainty-Guided Partition and Augmentation for Robust Cross-View Geo-Localization under Noisy Correspondence
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Outdoor Monocular SLAM with Global Scale-Consistent 3D Gaussian Pointmaps
by: Cheng, Chong, et al.
Published: (2025)
by: Cheng, Chong, et al.
Published: (2025)
CVVNet: A Cross-Vertical-View Network for Gait Recognition
by: Li, Xiangru, et al.
Published: (2025)
by: Li, Xiangru, et al.
Published: (2025)
V$^{2}$-SAM: Marrying SAM2 with Multi-Prompt Experts for Cross-View Object Correspondence
by: Pan, Jiancheng, et al.
Published: (2025)
by: Pan, Jiancheng, et al.
Published: (2025)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
Contrastive Language-Colored Pointmap Pretraining for Unified 3D Scene Understanding
by: Mao, Ye, et al.
Published: (2026)
by: Mao, Ye, et al.
Published: (2026)
Dual-Level Cross-Modal Contrastive Clustering
by: Zhang, Haixin, et al.
Published: (2024)
by: Zhang, Haixin, et al.
Published: (2024)
Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features
by: Xiangli, Yuanbo, et al.
Published: (2024)
by: Xiangli, Yuanbo, et al.
Published: (2024)
Cross-View Cross-Modal Unsupervised Domain Adaptation for Driver Monitoring System
by: Bhalla, Aditi, et al.
Published: (2025)
by: Bhalla, Aditi, et al.
Published: (2025)
Continual Cross-Modal Generalization
by: Xia, Yan, et al.
Published: (2025)
by: Xia, Yan, et al.
Published: (2025)
Seeing a Rose in Five Thousand Ways
by: Zhang, Yunzhi, et al.
Published: (2022)
by: Zhang, Yunzhi, et al.
Published: (2022)
Honey, I Shrunk the Arc de Triomphe!
by: Xiangli, Yuanbo, et al.
Published: (2026)
by: Xiangli, Yuanbo, et al.
Published: (2026)
MV-SAM: Multi-view Promptable Segmentation using Pointmap Guidance
by: Jeong, Yoonwoo, et al.
Published: (2026)
by: Jeong, Yoonwoo, et al.
Published: (2026)
ShadowDraw: From Any Object to Shadow-Drawing Compositional Art
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
VAGeo: View-specific Attention for Cross-View Object Geo-Localization
by: Li, Zhongyang, et al.
Published: (2025)
by: Li, Zhongyang, et al.
Published: (2025)
Similar Items
-
G3T Up! Gravity Aligned Coordinate Frames Simplify Pointmap Processing
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026) -
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020) -
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024) -
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024) -
COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
by: Wu, Hao, et al.
Published: (2024)