Dual-Domain Representation Alignment: Bridging 2D and 3D Vision via Geometry-Aware Architecture Search
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Haoyu, Yu, Zhihao, Wang, Rui, Jin, Yaochu, Liu, Qiqi, Cheng, Ran |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DoReMi: Bridging 3D Domains via Topology-Aware Domain-Representation Mixture of Experts
di: Xing, Mingwei, et al.
Pubblicazione: (2025)
di: Xing, Mingwei, et al.
Pubblicazione: (2025)
Bridge 2D-3D: Uncertainty-aware Hierarchical Registration Network with Domain Alignment
di: Cheng, Zhixin, et al.
Pubblicazione: (2025)
di: Cheng, Zhixin, et al.
Pubblicazione: (2025)
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
di: Kim, Jin Hyeon, et al.
Pubblicazione: (2026)
di: Kim, Jin Hyeon, et al.
Pubblicazione: (2026)
Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
di: Jiang, Jerry, et al.
Pubblicazione: (2026)
di: Jiang, Jerry, et al.
Pubblicazione: (2026)
Robust 3D Face Alignment with Multi-Path Neural Architecture Search
di: Jiang, Zhichao, et al.
Pubblicazione: (2024)
di: Jiang, Zhichao, et al.
Pubblicazione: (2024)
Direction-Aware Hybrid Representation Learning for 3D Hand Pose and Shape Estimation
di: Liu, Shiyong, et al.
Pubblicazione: (2025)
di: Liu, Shiyong, et al.
Pubblicazione: (2025)
Point-SRA: Self-Representation Alignment for 3D Representation Learning
di: Wei, Lintong, et al.
Pubblicazione: (2026)
di: Wei, Lintong, et al.
Pubblicazione: (2026)
GAPrompt: Geometry-Aware Point Cloud Prompt for 3D Vision Model
di: Ai, Zixiang, et al.
Pubblicazione: (2025)
di: Ai, Zixiang, et al.
Pubblicazione: (2025)
NeuralGS: Bridging Neural Fields and 3D Gaussian Splatting for Compact 3D Representations
di: Tang, Zhenyu, et al.
Pubblicazione: (2025)
di: Tang, Zhenyu, et al.
Pubblicazione: (2025)
GAMA: Geometry-Aware Manifold Alignment via Structured Adversarial Perturbations for Robust Domain Adaptation
di: Satou, Hana, et al.
Pubblicazione: (2025)
di: Satou, Hana, et al.
Pubblicazione: (2025)
Geometry-Aware Score Distillation via 3D Consistent Noising and Gradient Consistency Modeling
di: Kwak, Min-Seop, et al.
Pubblicazione: (2024)
di: Kwak, Min-Seop, et al.
Pubblicazione: (2024)
IMFine: 3D Inpainting via Geometry-guided Multi-view Refinement
di: Shi, Zhihao, et al.
Pubblicazione: (2025)
di: Shi, Zhihao, et al.
Pubblicazione: (2025)
When Alignment Fails: Multimodal Adversarial Attacks on Vision-Language-Action Models
di: Yan, Yuping, et al.
Pubblicazione: (2025)
di: Yan, Yuping, et al.
Pubblicazione: (2025)
D3GU: Multi-Target Active Domain Adaptation via Enhancing Domain Alignment
di: Zhang, Lin, et al.
Pubblicazione: (2024)
di: Zhang, Lin, et al.
Pubblicazione: (2024)
Proximal Vision Transformer: Enhancing Feature Representation through Two-Stage Manifold Geometry
di: Yun, Haoyu, et al.
Pubblicazione: (2025)
di: Yun, Haoyu, et al.
Pubblicazione: (2025)
3D Congealing: 3D-Aware Image Alignment in the Wild
di: Zhang, Yunzhi, et al.
Pubblicazione: (2024)
di: Zhang, Yunzhi, et al.
Pubblicazione: (2024)
Towards Intrinsic-Aware Monocular 3D Object Detection
di: Zhang, Zhihao, et al.
Pubblicazione: (2026)
di: Zhang, Zhihao, et al.
Pubblicazione: (2026)
Geo-Align: Video Generation Alignment via Metric Geometry Reward
di: Li, Zizun, et al.
Pubblicazione: (2026)
di: Li, Zizun, et al.
Pubblicazione: (2026)
3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning
di: Hu, Naiwen, et al.
Pubblicazione: (2024)
di: Hu, Naiwen, et al.
Pubblicazione: (2024)
Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling
di: Li, Zhihao, et al.
Pubblicazione: (2025)
di: Li, Zhihao, et al.
Pubblicazione: (2025)
Moving Light Adaptive Colonoscopy Reconstruction via Illumination-Attenuation-Aware 3D Gaussian Splatting
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
di: Huang, Hai, et al.
Pubblicazione: (2025)
di: Huang, Hai, et al.
Pubblicazione: (2025)
Spherical Geometry Diffusion: Generating High-quality 3D Face Geometry via Sphere-anchored Representations
di: Zhang, Junyi, et al.
Pubblicazione: (2026)
di: Zhang, Junyi, et al.
Pubblicazione: (2026)
LiftFeat: 3D Geometry-Aware Local Feature Matching
di: Liu, Yepeng, et al.
Pubblicazione: (2025)
di: Liu, Yepeng, et al.
Pubblicazione: (2025)
Unsupervised Part Discovery via Dual Representation Alignment
di: Xia, Jiahao, et al.
Pubblicazione: (2024)
di: Xia, Jiahao, et al.
Pubblicazione: (2024)
Information Maximization Clustering via Multi-View Self-Labelling
di: Ntelemis, Foivos, et al.
Pubblicazione: (2021)
di: Ntelemis, Foivos, et al.
Pubblicazione: (2021)
Domain-Adaptive 2D Human Pose Estimation via Dual Teachers in Extremely Low-Light Conditions
di: Ai, Yihao, et al.
Pubblicazione: (2024)
di: Ai, Yihao, et al.
Pubblicazione: (2024)
PromptSync: Bridging Domain Gaps in Vision-Language Models through Class-Aware Prototype Alignment and Discrimination
di: Khandelwal, Anant
Pubblicazione: (2024)
di: Khandelwal, Anant
Pubblicazione: (2024)
Learning Robust 3D Representation from CLIP via Dual Denoising
di: Luo, Shuqing, et al.
Pubblicazione: (2024)
di: Luo, Shuqing, et al.
Pubblicazione: (2024)
Vision to Geometry: 3D Spatial Memory for Sequential Embodied MLLM Reasoning and Exploration
di: Cai, Zhongyi, et al.
Pubblicazione: (2025)
di: Cai, Zhongyi, et al.
Pubblicazione: (2025)
Geometry-Aware 3D Salient Object Detection Network
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction
di: Chen, Xun, et al.
Pubblicazione: (2026)
di: Chen, Xun, et al.
Pubblicazione: (2026)
DualNeRF: Text-Driven 3D Scene Editing via Dual-Field Representation
di: Xiong, Yuxuan, et al.
Pubblicazione: (2025)
di: Xiong, Yuxuan, et al.
Pubblicazione: (2025)
3D Aware Region Prompted Vision Language Model
di: Cheng, An-Chieh, et al.
Pubblicazione: (2025)
di: Cheng, An-Chieh, et al.
Pubblicazione: (2025)
REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment
di: Han, Haonan, et al.
Pubblicazione: (2024)
di: Han, Haonan, et al.
Pubblicazione: (2024)
GLASS: Geometry-aware Local Alignment and Structure Synchronization Network for 2D-3D Registration
di: Cheng, Zhixin, et al.
Pubblicazione: (2026)
di: Cheng, Zhixin, et al.
Pubblicazione: (2026)
Generalized Robot 3D Vision-Language Model with Fast Rendering and Pre-Training Vision-Language Alignment
di: Liu, Kangcheng, et al.
Pubblicazione: (2023)
di: Liu, Kangcheng, et al.
Pubblicazione: (2023)
Bridging Domain Gap of Point Cloud Representations via Self-Supervised Geometric Augmentation
di: Yu, Li, et al.
Pubblicazione: (2024)
di: Yu, Li, et al.
Pubblicazione: (2024)
Grounded 3D-Aware Spatial Vision-Language Modeling
di: Cheng, An-Chieh, et al.
Pubblicazione: (2026)
di: Cheng, An-Chieh, et al.
Pubblicazione: (2026)
Documenti analoghi
-
DoReMi: Bridging 3D Domains via Topology-Aware Domain-Representation Mixture of Experts
di: Xing, Mingwei, et al.
Pubblicazione: (2025) -
Bridge 2D-3D: Uncertainty-aware Hierarchical Registration Network with Domain Alignment
di: Cheng, Zhixin, et al.
Pubblicazione: (2025) -
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
di: Kim, Jin Hyeon, et al.
Pubblicazione: (2026) -
Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
di: Jiang, Jerry, et al.
Pubblicazione: (2026) -
Robust 3D Face Alignment with Multi-Path Neural Architecture Search
di: Jiang, Zhichao, et al.
Pubblicazione: (2024)