B2N3D: Progressive Learning from Binary to N-ary Relationships for 3D Object Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Feng, Xu, Hongbin, Ci, Hai, Kang, Wenxiong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LSVG: Language-Guided Scene Graphs with 2D-Assisted Multi-Modal Encoding for 3D Visual Grounding
by: Xiao, Feng, et al.
Published: (2025)
by: Xiao, Feng, et al.
Published: (2025)
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
by: Xiao, Feng, et al.
Published: (2024)
by: Xiao, Feng, et al.
Published: (2024)
4DStyleGaussian: Zero-shot 4D Style Transfer with Gaussian Splatting
by: Liang, Wanlin, et al.
Published: (2024)
by: Liang, Wanlin, et al.
Published: (2024)
Cyc3D: Fine-grained Controllable 3D Generation via Cycle Consistency Regularization
by: Xu, Hongbin, et al.
Published: (2025)
by: Xu, Hongbin, et al.
Published: (2025)
StyleDyRF: Zero-shot 4D Style Transfer for Dynamic Neural Radiance Fields
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
EA-3DGS: Efficient and Adaptive 3D Gaussians with Highly Enhanced Quality for outdoor scenes
by: Guo, Jianlin, et al.
Published: (2025)
by: Guo, Jianlin, et al.
Published: (2025)
Improving 3D Finger Traits Recognition via Generalizable Neural Rendering
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
PointDC:Unsupervised Semantic Segmentation of 3D Point Clouds via Cross-modal Distillation and Super-Voxel Clustering
by: Chen, Zisheng, et al.
Published: (2023)
by: Chen, Zisheng, et al.
Published: (2023)
Gesplat: Robust Pose-Free 3D Reconstruction via Geometry-Guided Gaussian Splatting
by: Lu, Jiahui, et al.
Published: (2025)
by: Lu, Jiahui, et al.
Published: (2025)
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
by: Baik, Sangwon, et al.
Published: (2025)
by: Baik, Sangwon, et al.
Published: (2025)
N3D-VLM: Native 3D Grounding Enables Accurate Spatial Reasoning in Vision-Language Models
by: Wang, Yuxin, et al.
Published: (2025)
by: Wang, Yuxin, et al.
Published: (2025)
RobustMVS: Single Domain Generalized Deep Multi-view Stereo
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
by: Tian, Tongxuan, et al.
Published: (2025)
by: Tian, Tongxuan, et al.
Published: (2025)
M3DHMR: Monocular 3D Hand Mesh Recovery
by: Lin, Yihong, et al.
Published: (2025)
by: Lin, Yihong, et al.
Published: (2025)
GraphRelate3D: Context-Dependent 3D Object Detection with Inter-Object Relationship Graphs
by: Liu, Mingyu, et al.
Published: (2024)
by: Liu, Mingyu, et al.
Published: (2024)
Fully Test-Time Adaptation for Monocular 3D Object Detection
by: Lin, Hongbin, et al.
Published: (2024)
by: Lin, Hongbin, et al.
Published: (2024)
Binary-Gaussian: Compact and Progressive Representation for 3D Gaussian Segmentation
by: Yang, An, et al.
Published: (2025)
by: Yang, An, et al.
Published: (2025)
Progressive Inertial Poser: Progressive Real-Time Kinematic Chain Estimation for 3D Full-Body Pose from Three IMU Sensors
by: Zhu, Zunjie, et al.
Published: (2025)
by: Zhu, Zunjie, et al.
Published: (2025)
SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding
by: Jin, Zhao, et al.
Published: (2025)
by: Jin, Zhao, et al.
Published: (2025)
Seg2Box: 3D Object Detection by Point-Wise Semantics Supervision
by: Zheng, Maoji, et al.
Published: (2025)
by: Zheng, Maoji, et al.
Published: (2025)
GO-N3RDet: Geometry Optimized NeRF-enhanced 3D Object Detector
by: Li, Zechuan, et al.
Published: (2025)
by: Li, Zechuan, et al.
Published: (2025)
CLHOP: Combined Audio-Video Learning for Horse 3D Pose and Shape Estimation
by: Li, Ci, et al.
Published: (2024)
by: Li, Ci, et al.
Published: (2024)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
by: Mao, Aihua, et al.
Published: (2026)
by: Mao, Aihua, et al.
Published: (2026)
Resolving Long-Tail Ambiguity in Unsupervised 3D Point Cloud Segmentation with Language Priors
by: Wei, Siqi, et al.
Published: (2026)
by: Wei, Siqi, et al.
Published: (2026)
Brain3D: Generating 3D Objects from fMRI
by: Yang, Yuankun, et al.
Published: (2024)
by: Yang, Yuankun, et al.
Published: (2024)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
by: Gao, Xianqiang, et al.
Published: (2024)
by: Gao, Xianqiang, et al.
Published: (2024)
Multi-Task Domain Adaptation for Language Grounding with 3D Objects
by: Sun, Penglei, et al.
Published: (2024)
by: Sun, Penglei, et al.
Published: (2024)
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
by: Xie, Junyu, et al.
Published: (2026)
by: Xie, Junyu, et al.
Published: (2026)
Progressive Multi-Modal Fusion for Robust 3D Object Detection
by: Mohan, Rohit, et al.
Published: (2024)
by: Mohan, Rohit, et al.
Published: (2024)
Intent3D: 3D Object Detection in RGB-D Scans Based on Human Intention
by: Kang, Weitai, et al.
Published: (2024)
by: Kang, Weitai, et al.
Published: (2024)
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
by: Lin, Yihong, et al.
Published: (2024)
by: Lin, Yihong, et al.
Published: (2024)
Fusion4CA: Boosting 3D Object Detection via Comprehensive Image Exploitation
by: Luo, Kang, et al.
Published: (2026)
by: Luo, Kang, et al.
Published: (2026)
DM3D: Distortion-Minimized Weight Pruning for Lossless 3D Object Detection
by: Xu, Kaixin, et al.
Published: (2024)
by: Xu, Kaixin, et al.
Published: (2024)
Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph
by: Linok, Sergey, et al.
Published: (2025)
by: Linok, Sergey, et al.
Published: (2025)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
by: Yang, Yuhang, et al.
Published: (2023)
by: Yang, Yuhang, et al.
Published: (2023)
ChangingGrounding: 3D Visual Grounding in Changing Scenes
by: Hu, Miao, et al.
Published: (2025)
by: Hu, Miao, et al.
Published: (2025)
SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Commonsense Prototype for Outdoor Unsupervised 3D Object Detection
by: Wu, Hai, et al.
Published: (2024)
by: Wu, Hai, et al.
Published: (2024)
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
by: Lin, Yihong, et al.
Published: (2024)
by: Lin, Yihong, et al.
Published: (2024)
Similar Items
-
LSVG: Language-Guided Scene Graphs with 2D-Assisted Multi-Modal Encoding for 3D Visual Grounding
by: Xiao, Feng, et al.
Published: (2025) -
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
by: Xiao, Feng, et al.
Published: (2024) -
4DStyleGaussian: Zero-shot 4D Style Transfer with Gaussian Splatting
by: Liang, Wanlin, et al.
Published: (2024) -
Cyc3D: Fine-grained Controllable 3D Generation via Cycle Consistency Regularization
by: Xu, Hongbin, et al.
Published: (2025) -
StyleDyRF: Zero-shot 4D Style Transfer for Dynamic Neural Radiance Fields
by: Xu, Hongbin, et al.
Published: (2024)