Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features
Fuente:
arXiv
Saved in:
| Main Authors: | Xiangli, Yuanbo, Cai, Ruojin, Chen, Hanyu, Byrne, Jeffrey, Snavely, Noah |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild
by: Chen, Hanyu, et al.
Published: (2026)
by: Chen, Hanyu, et al.
Published: (2026)
Long-tail Internet photo reconstruction
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
Honey, I Shrunk the Arc de Triomphe!
by: Xiangli, Yuanbo, et al.
Published: (2026)
by: Xiangli, Yuanbo, et al.
Published: (2026)
Can Generative Video Models Help Pose Estimation?
by: Cai, Ruojin, et al.
Published: (2024)
by: Cai, Ruojin, et al.
Published: (2024)
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024)
by: Tung, Joseph, et al.
Published: (2024)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
GSDF: 3DGS Meets SDF for Improved Rendering and Reconstruction
by: Yu, Mulin, et al.
Published: (2024)
by: Yu, Mulin, et al.
Published: (2024)
G3T Up! Gravity Aligned Coordinate Frames Simplify Pointmap Processing
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
Wide-Baseline Relative Camera Pose Estimation with Directional Learning
by: Chen, Kefan, et al.
Published: (2021)
by: Chen, Kefan, et al.
Published: (2021)
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
by: Lee, Miso, et al.
Published: (2026)
by: Lee, Miso, et al.
Published: (2026)
Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos
by: Jin, Linyi, et al.
Published: (2024)
by: Jin, Linyi, et al.
Published: (2024)
Emergent Extreme-View Geometry in 3D Foundation Models
by: Zhang, Yiwen, et al.
Published: (2025)
by: Zhang, Yiwen, et al.
Published: (2025)
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Generative Image Dynamics
by: Li, Zhengqi, et al.
Published: (2023)
by: Li, Zhengqi, et al.
Published: (2023)
Seeing a Rose in Five Thousand Ways
by: Zhang, Yunzhi, et al.
Published: (2022)
by: Zhang, Yunzhi, et al.
Published: (2022)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
Turbo-GS: Accelerating 3D Gaussian Fitting for High-Quality Radiance Fields
by: Dhiman, Ankit, et al.
Published: (2024)
by: Dhiman, Ankit, et al.
Published: (2024)
3D Part Segmentation via Geometric Aggregation of 2D Visual Features
by: Garosi, Marco, et al.
Published: (2024)
by: Garosi, Marco, et al.
Published: (2024)
Cog2Gen3D: Sculpturing 3D Semantic-Geometric Cognition for 3D Generation
by: Wang, Haonan, et al.
Published: (2026)
by: Wang, Haonan, et al.
Published: (2026)
Proc-GS: Procedural Building Generation for City Assembly with 3D Gaussians
by: Li, Yixuan, et al.
Published: (2024)
by: Li, Yixuan, et al.
Published: (2024)
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
by: Chou, Gene, et al.
Published: (2024)
by: Chou, Gene, et al.
Published: (2024)
Extreme Rotation Estimation in the Wild
by: Bezalel, Hana, et al.
Published: (2024)
by: Bezalel, Hana, et al.
Published: (2024)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
by: Lei, Jiahui, et al.
Published: (2025)
by: Lei, Jiahui, et al.
Published: (2025)
VGLD: Visually-Guided Linguistic Disambiguation for Monocular Depth Scale Recovery
by: Wu, Bojin, et al.
Published: (2025)
by: Wu, Bojin, et al.
Published: (2025)
3D Spatial Understanding in MLLMs: Disambiguation and Evaluation
by: Chang, Chun-Peng, et al.
Published: (2024)
by: Chang, Chun-Peng, et al.
Published: (2024)
ShadowDraw: From Any Object to Shadow-Drawing Compositional Art
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
FUSELOC: Fusing Global and Local Descriptors to Disambiguate 2D-3D Matching in Visual Localization
by: Nguyen, Son Tung, et al.
Published: (2024)
by: Nguyen, Son Tung, et al.
Published: (2024)
Beyond the Frame: Generating 360 Panoramic Videos from Perspective Videos
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
dopanim: A Dataset of Doppelganger Animals with Noisy Annotations from Multiple Humans
by: Herde, Marek, et al.
Published: (2024)
by: Herde, Marek, et al.
Published: (2024)
Learning 3D-Aware GANs from Unposed Images with Template Feature Field
by: Chen, Xinya, et al.
Published: (2024)
by: Chen, Xinya, et al.
Published: (2024)
Selfi: Self Improving Reconstruction Engine via 3D Geometric Feature Alignment
by: Deng, Youming, et al.
Published: (2025)
by: Deng, Youming, et al.
Published: (2025)
Golyadkin's Torment: Doppelgängers and Adversarial Vulnerability
by: Kamberov, George I.
Published: (2024)
by: Kamberov, George I.
Published: (2024)
Visual Chronicles: Using Multimodal LLMs to Analyze Massive Collections of Images
by: Deng, Boyang, et al.
Published: (2025)
by: Deng, Boyang, et al.
Published: (2025)
Disambiguating Monocular Reconstruction of 3D Clothed Human with Spatial-Temporal Transformer
by: Deng, Yong, et al.
Published: (2024)
by: Deng, Yong, et al.
Published: (2024)
Ukrainian Visual Word Sense Disambiguation Benchmark
by: Laba, Yurii, et al.
Published: (2026)
by: Laba, Yurii, et al.
Published: (2026)
3D Reconstruction with Fast Dipole Sums
by: Chen, Hanyu, et al.
Published: (2024)
by: Chen, Hanyu, et al.
Published: (2024)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
Self-Supervised Ultrasound-Video Segmentation with Feature Prediction and 3D Localised Loss
by: Ellis, Edward, et al.
Published: (2025)
by: Ellis, Edward, et al.
Published: (2025)
Similar Items
-
ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild
by: Chen, Hanyu, et al.
Published: (2026) -
Long-tail Internet photo reconstruction
by: Li, Yuan, et al.
Published: (2026) -
Honey, I Shrunk the Arc de Triomphe!
by: Xiangli, Yuanbo, et al.
Published: (2026) -
Can Generative Video Models Help Pose Estimation?
by: Cai, Ruojin, et al.
Published: (2024) -
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024)