ViewSRD: 3D Visual Grounding via Structured Multi-View Decomposition
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Ronggang, Yang, Haoxin, Cai, Yan, Xu, Xuemiao, Zhang, Huaidong, He, Shengfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SITA: Structurally Imperceptible and Transferable Adversarial Attacks for Stylized Image Generation
by: Kang, Jingdan, et al.
Published: (2025)
by: Kang, Jingdan, et al.
Published: (2025)
G2Face: High-Fidelity Reversible Face Anonymization via Generative and Geometric Priors
by: Yang, Haoxin, et al.
Published: (2024)
by: Yang, Haoxin, et al.
Published: (2024)
StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion
by: Yang, Haoxin, et al.
Published: (2025)
by: Yang, Haoxin, et al.
Published: (2025)
VrdONE: One-stage Video Visual Relation Detection
by: Jiang, Xinjie, et al.
Published: (2024)
by: Jiang, Xinjie, et al.
Published: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
by: Liao, Liwei, et al.
Published: (2025)
by: Liao, Liwei, et al.
Published: (2025)
NexusGS: Sparse View Synthesis with Epipolar Depth Priors in 3D Gaussian Splatting
by: Zheng, Yulong, et al.
Published: (2025)
by: Zheng, Yulong, et al.
Published: (2025)
Registration is a Powerful Rotation-Invariance Learner for 3D Anomaly Detection
by: Yu, Yuyang, et al.
Published: (2025)
by: Yu, Yuyang, et al.
Published: (2025)
SCJD: Sparse Correlation and Joint Distillation for Efficient 3D Human Pose Estimation
by: Chen, Weihong, et al.
Published: (2025)
by: Chen, Weihong, et al.
Published: (2025)
Rotation-Adaptive Point Cloud Domain Generalization via Intricate Orientation Learning
by: Liu, Bangzhen, et al.
Published: (2025)
by: Liu, Bangzhen, et al.
Published: (2025)
Beyond Inference Intervention: Identity-Decoupled Diffusion for Face Anonymization
by: Yang, Haoxin, et al.
Published: (2025)
by: Yang, Haoxin, et al.
Published: (2025)
StableGuard: Towards Unified Copyright Protection and Tamper Localization in Latent Diffusion Models
by: Yang, Haoxin, et al.
Published: (2025)
by: Yang, Haoxin, et al.
Published: (2025)
Seeing 3D Through 2D Lenses: 3D Few-Shot Class-Incremental Learning via Cross-Modal Geometric Rectification
by: Xiang, Tuo, et al.
Published: (2025)
by: Xiang, Tuo, et al.
Published: (2025)
RecDreamer: Consistent Text-to-3D Generation via Uniform Score Distillation
by: Zheng, Chenxi, et al.
Published: (2025)
by: Zheng, Chenxi, et al.
Published: (2025)
Multimodal-Prior-Guided Importance Sampling for Hierarchical Gaussian Splatting in Sparse-View Novel View Synthesis
by: Xiong, Kaiqiang, et al.
Published: (2026)
by: Xiong, Kaiqiang, et al.
Published: (2026)
NuGrounding: A Multi-View 3D Visual Grounding Framework in Autonomous Driving
by: Li, Fuhao, et al.
Published: (2025)
by: Li, Fuhao, et al.
Published: (2025)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
by: Li, Jinke, et al.
Published: (2024)
by: Li, Jinke, et al.
Published: (2024)
MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics
by: Liu, Haofeng, et al.
Published: (2026)
by: Liu, Haofeng, et al.
Published: (2026)
PlaneMVS: 3D Plane Reconstruction from Multi-View Stereo
by: Liu, Jiachen, et al.
Published: (2022)
by: Liu, Jiachen, et al.
Published: (2022)
Structure Consistent Gaussian Splatting with Matching Prior for Few-shot Novel View Synthesis
by: Peng, Rui, et al.
Published: (2024)
by: Peng, Rui, et al.
Published: (2024)
MVPGS: Excavating Multi-view Priors for Gaussian Splatting from Sparse Input Views
by: Xu, Wangze, et al.
Published: (2024)
by: Xu, Wangze, et al.
Published: (2024)
GenS: Generalizable Neural Surface Reconstruction from Multi-View Images
by: Peng, Rui, et al.
Published: (2024)
by: Peng, Rui, et al.
Published: (2024)
View-on-Graph: Zero-shot 3D Visual Grounding via Vision-Language Reasoning on Scene Graphs
by: Liu, Yuanyuan, et al.
Published: (2025)
by: Liu, Yuanyuan, et al.
Published: (2025)
Déjà View: Looping Transformers for Multi-View 3D Reconstruction
by: Burzio, Alessandro, et al.
Published: (2026)
by: Burzio, Alessandro, et al.
Published: (2026)
IDT: A Physically Grounded Transformer for Feed-Forward Multi-View Intrinsic Decomposition
by: Du, Kang, et al.
Published: (2025)
by: Du, Kang, et al.
Published: (2025)
View Transformation Robustness for Multi-View 3D Object Reconstruction with Reconstruction Error-Guided View Selection
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
by: Yang, Xianghui, et al.
Published: (2024)
by: Yang, Xianghui, et al.
Published: (2024)
Satellite to GroundScape -- Large-scale Consistent Ground View Generation from Satellite Views
by: Xu, Ningli, et al.
Published: (2025)
by: Xu, Ningli, et al.
Published: (2025)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
Enhancing Novel View Synthesis via Geometry Grounded Set Diffusion
by: Zanjani, Farhad G., et al.
Published: (2026)
by: Zanjani, Farhad G., et al.
Published: (2026)
PersonaMagic: Stage-Regulated High-Fidelity Face Customization with Tandem Equilibrium
by: Li, Xinzhe, et al.
Published: (2024)
by: Li, Xinzhe, et al.
Published: (2024)
Geospecific View Generation -- Geometry-Context Aware High-resolution Ground View Inference from Satellite Views
by: Xu, Ningli, et al.
Published: (2024)
by: Xu, Ningli, et al.
Published: (2024)
FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation
by: Shen, Tingrui, et al.
Published: (2025)
by: Shen, Tingrui, et al.
Published: (2025)
Multi-View 3D Point Tracking
by: Rajič, Frano, et al.
Published: (2025)
by: Rajič, Frano, et al.
Published: (2025)
ViewFusion: Structured Spatial Thinking Chains for Multi-View Reasoning
by: Tao, Xingjian, et al.
Published: (2026)
by: Tao, Xingjian, et al.
Published: (2026)
RIAV-MVS: Recurrent-Indexing an Asymmetric Volume for Multi-View Stereo
by: Cai, Changjiang, et al.
Published: (2022)
by: Cai, Changjiang, et al.
Published: (2022)
MagicView: Multi-View Consistent Identity Customization via Priors-Guided In-Context Learning
by: Li, Hengjia, et al.
Published: (2025)
by: Li, Hengjia, et al.
Published: (2025)
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
by: Lin, Jiawen, et al.
Published: (2025)
by: Lin, Jiawen, et al.
Published: (2025)
ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction
by: Liang, Jie, et al.
Published: (2026)
by: Liang, Jie, et al.
Published: (2026)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
by: Wang, Jiasen, et al.
Published: (2024)
by: Wang, Jiasen, et al.
Published: (2024)
Multi-Person 3D Pose Estimation from Multi-View Uncalibrated Depth Cameras
by: Li, Yu-Jhe, et al.
Published: (2024)
by: Li, Yu-Jhe, et al.
Published: (2024)
Similar Items
-
SITA: Structurally Imperceptible and Transferable Adversarial Attacks for Stylized Image Generation
by: Kang, Jingdan, et al.
Published: (2025) -
G2Face: High-Fidelity Reversible Face Anonymization via Generative and Geometric Priors
by: Yang, Haoxin, et al.
Published: (2024) -
StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion
by: Yang, Haoxin, et al.
Published: (2025) -
VrdONE: One-stage Video Visual Relation Detection
by: Jiang, Xinjie, et al.
Published: (2024) -
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
by: Liao, Liwei, et al.
Published: (2025)