VSFormer: Mining Correlations in Flexible View Set for Multi-view 3D Shape Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Hongyu, Wang, Yongcai, Wang, Peng, Deng, Haoran, Cai, Xudong, Li, Deying |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Parameter-efficient Prompt Learning for 3D Point Cloud Understanding
by: Sun, Hongyu, et al.
Published: (2024)
by: Sun, Hongyu, et al.
Published: (2024)
DroneMOT: Drone-based Multi-Object Tracking Considering Detection Difficulties and Simultaneous Moving of Drones and Objects
by: Wang, Peng, et al.
Published: (2024)
by: Wang, Peng, et al.
Published: (2024)
VOLoc: Visual Place Recognition by Querying Compressed Lidar Map
by: Cai, Xudong, et al.
Published: (2024)
by: Cai, Xudong, et al.
Published: (2024)
Point-PRC: A Prompt Learning Based Regulation Framework for Generalizable Point Cloud Analysis
by: Sun, Hongyu, et al.
Published: (2024)
by: Sun, Hongyu, et al.
Published: (2024)
Dust to Tower: Coarse-to-Fine Photo-Realistic Scene Reconstruction from Sparse Uncalibrated Images
by: Cai, Xudong, et al.
Published: (2024)
by: Cai, Xudong, et al.
Published: (2024)
Point-Cache: Test-time Dynamic and Hierarchical Cache for Robust and Generalizable Point Cloud Analysis
by: Sun, Hongyu, et al.
Published: (2025)
by: Sun, Hongyu, et al.
Published: (2025)
EIMC: Efficient Instance-aware Multi-modal Collaborative Perception
by: Yang, Kang, et al.
Published: (2026)
by: Yang, Kang, et al.
Published: (2026)
Mem4D: Decoupling Static and Dynamic Memory for Dynamic Scene Reconstruction
by: Cai, Xudong, et al.
Published: (2025)
by: Cai, Xudong, et al.
Published: (2025)
BOLT: Online Lightweight Adaptation for Preparation-Free Heterogeneous Cooperative Perception
by: Yang, Kang, et al.
Published: (2026)
by: Yang, Kang, et al.
Published: (2026)
MambaVO: Deep Visual Odometry Based on Sequential Matching Refinement and Training Smoothing
by: Wang, Shuo, et al.
Published: (2024)
by: Wang, Shuo, et al.
Published: (2024)
PRISM: PRogressive dependency maxImization for Scale-invariant image Matching
by: Cai, Xudong, et al.
Published: (2024)
by: Cai, Xudong, et al.
Published: (2024)
VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning
by: Liao, Tangfei, et al.
Published: (2023)
by: Liao, Tangfei, et al.
Published: (2023)
Shape from Semantics: 3D Shape Generation from Multi-View Semantics
by: Li, Liangchen, et al.
Published: (2025)
by: Li, Liangchen, et al.
Published: (2025)
Is Discretization Fusion All You Need for Collaborative Perception?
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
IDLS: Inverse Depth Line based Visual-Inertial SLAM
by: Li, Wanting, et al.
Published: (2023)
by: Li, Wanting, et al.
Published: (2023)
CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis
by: Kong, Xin, et al.
Published: (2025)
by: Kong, Xin, et al.
Published: (2025)
3D-Aware Multi-Task Learning with Cross-View Correlations for Dense Scene Understanding
by: Wang, Xiaoye, et al.
Published: (2025)
by: Wang, Xiaoye, et al.
Published: (2025)
MonoDream: Monocular Vision-Language Navigation with Panoramic Dreaming
by: Wang, Shuo, et al.
Published: (2025)
by: Wang, Shuo, et al.
Published: (2025)
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
by: Kim, Seungwook, et al.
Published: (2024)
by: Kim, Seungwook, et al.
Published: (2024)
GSLAMOT: A Tracklet and Query Graph-based Simultaneous Locating, Mapping, and Multiple Object Tracking System
by: Wang, Shuo, et al.
Published: (2024)
by: Wang, Shuo, et al.
Published: (2024)
Exploring Recurrent Long-term Temporal Fusion for Multi-view 3D Perception
by: Han, Chunrui, et al.
Published: (2023)
by: Han, Chunrui, et al.
Published: (2023)
MVDream: Multi-view Diffusion for 3D Generation
by: Shi, Yichun, et al.
Published: (2023)
by: Shi, Yichun, et al.
Published: (2023)
Transformed Multi-view 3D Shape Features with Contrastive Learning
by: Costa, Márcus Vinícius Lobo, et al.
Published: (2025)
by: Costa, Márcus Vinícius Lobo, et al.
Published: (2025)
Make Your MoVe: Make Your 3D Contents by Adapting Multi-View Diffusion Models to External Editing
by: Wang, Weitao, et al.
Published: (2025)
by: Wang, Weitao, et al.
Published: (2025)
TAMM: TriAdapter Multi-Modal Learning for 3D Shape Understanding
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
CoDiff: Conditional Diffusion Model for Collaborative 3D Object Detection
by: Huang, Zhe, et al.
Published: (2025)
by: Huang, Zhe, et al.
Published: (2025)
Robust Fuzzy Multi-view Learning under View Conflict
by: Duan, Siyuan, et al.
Published: (2026)
by: Duan, Siyuan, et al.
Published: (2026)
COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
MVPGS: Excavating Multi-view Priors for Gaussian Splatting from Sparse Input Views
by: Xu, Wangze, et al.
Published: (2024)
by: Xu, Wangze, et al.
Published: (2024)
EVA-Gaussian: 3D Gaussian-based Real-time Human Novel View Synthesis under Diverse Multi-view Camera Settings
by: Hu, Yingdong, et al.
Published: (2024)
by: Hu, Yingdong, et al.
Published: (2024)
ShapeLLM: Universal 3D Object Understanding for Embodied Interaction
by: Qi, Zekun, et al.
Published: (2024)
by: Qi, Zekun, et al.
Published: (2024)
Unpaired Multi-view Clustering via Reliable View Guidance
by: Xin, Like, et al.
Published: (2024)
by: Xin, Like, et al.
Published: (2024)
MVGS: Multi-view Regulated Gaussian Splatting for Novel View Synthesis
by: Du, Xiaobiao, et al.
Published: (2024)
by: Du, Xiaobiao, et al.
Published: (2024)
Reliable Disentanglement Multi-view Learning Against View Adversarial Attacks
by: Wang, Xuyang, et al.
Published: (2025)
by: Wang, Xuyang, et al.
Published: (2025)
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
by: Kim, Seungwook, et al.
Published: (2024)
by: Kim, Seungwook, et al.
Published: (2024)
Flexible 3D Lane Detection by Hierarchical Shape MatchingFlexible 3D Lane Detection by Hierarchical Shape Matching
by: Guan, Zhihao, et al.
Published: (2024)
by: Guan, Zhihao, et al.
Published: (2024)
Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer
by: Deng, Yu, et al.
Published: (2024)
by: Deng, Yu, et al.
Published: (2024)
Unleashing the Multi-View Fusion Potential: Noise Correction in VLM for Open-Vocabulary 3D Scene Understanding
by: Yin, Xingyilang, et al.
Published: (2025)
by: Yin, Xingyilang, et al.
Published: (2025)
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images
by: Lee, Han-Hung, et al.
Published: (2024)
by: Lee, Han-Hung, et al.
Published: (2024)
Similar Items
-
Parameter-efficient Prompt Learning for 3D Point Cloud Understanding
by: Sun, Hongyu, et al.
Published: (2024) -
DroneMOT: Drone-based Multi-Object Tracking Considering Detection Difficulties and Simultaneous Moving of Drones and Objects
by: Wang, Peng, et al.
Published: (2024) -
VOLoc: Visual Place Recognition by Querying Compressed Lidar Map
by: Cai, Xudong, et al.
Published: (2024) -
Point-PRC: A Prompt Learning Based Regulation Framework for Generalizable Point Cloud Analysis
by: Sun, Hongyu, et al.
Published: (2024) -
Dust to Tower: Coarse-to-Fine Photo-Realistic Scene Reconstruction from Sparse Uncalibrated Images
by: Cai, Xudong, et al.
Published: (2024)