3DBench: A Scalable 3D Benchmark and Instruction-Tuning Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Junjie, Hu, Tianci, Huang, Xiaoshui, Gong, Yongshun, Zeng, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RS3DBench: A Comprehensive Benchmark for 3D Spatial Perception in Remote Sensing
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models
von: Zhang, Jiyao, et al.
Veröffentlicht: (2026)
von: Zhang, Jiyao, et al.
Veröffentlicht: (2026)
A Self-Conditioned Representation Guided Diffusion Model for Realistic Text-to-LiDAR Scene Generation
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
Diverse Teacher-Students for Deep Safe Semi-Supervised Learning under Class Mismatch
von: Wang, Qikai, et al.
Veröffentlicht: (2024)
von: Wang, Qikai, et al.
Veröffentlicht: (2024)
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration
von: Huang, Xiaoshui, et al.
Veröffentlicht: (2025)
von: Huang, Xiaoshui, et al.
Veröffentlicht: (2025)
COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
von: Sun, Shaorong, et al.
Veröffentlicht: (2024)
von: Sun, Shaorong, et al.
Veröffentlicht: (2024)
Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
CanoVerse: 3D Object Scalable Canonicalization and Dataset for Generation and Pose
von: Jin, Li, et al.
Veröffentlicht: (2026)
von: Jin, Li, et al.
Veröffentlicht: (2026)
Cross3DReg: Towards a Large-scale Real-world Cross-source Point Cloud Registration Benchmark
von: Xu, Zongyi, et al.
Veröffentlicht: (2025)
von: Xu, Zongyi, et al.
Veröffentlicht: (2025)
On the Evaluation and Refinement of Vision-Language Instruction Tuning Datasets
von: Liao, Ning, et al.
Veröffentlicht: (2023)
von: Liao, Ning, et al.
Veröffentlicht: (2023)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
von: Li, Zeju, et al.
Veröffentlicht: (2024)
von: Li, Zeju, et al.
Veröffentlicht: (2024)
GVGEN: Text-to-3D Generation with Volumetric Representation
von: He, Xianglong, et al.
Veröffentlicht: (2024)
von: He, Xianglong, et al.
Veröffentlicht: (2024)
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
GaussianEditor: Editing 3D Gaussians Delicately with Text Instructions
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
Benchmarking Micro-action Recognition: Dataset, Methods, and Applications
von: Guo, Dan, et al.
Veröffentlicht: (2024)
von: Guo, Dan, et al.
Veröffentlicht: (2024)
NeRF-Det++: Incorporating Semantic Cues and Perspective-aware Depth Supervision for Indoor Multi-View 3D Detection
von: Huang, Chenxi, et al.
Veröffentlicht: (2024)
von: Huang, Chenxi, et al.
Veröffentlicht: (2024)
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
M3D: Dataset Condensation by Minimizing Maximum Mean Discrepancy
von: Zhang, Hansong, et al.
Veröffentlicht: (2023)
von: Zhang, Hansong, et al.
Veröffentlicht: (2023)
LAHNet: Local Attentive Hashing Network for Point Cloud Registration
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
TSRE: Channel-Aware Typical Set Refinement for Out-of-Distribution Detection
von: Gao, Weijun, et al.
Veröffentlicht: (2025)
von: Gao, Weijun, et al.
Veröffentlicht: (2025)
An End-to-End Robust Point Cloud Semantic Segmentation Network with Single-Step Conditional Diffusion Models
von: Qu, Wentao, et al.
Veröffentlicht: (2024)
von: Qu, Wentao, et al.
Veröffentlicht: (2024)
SS3DM: Benchmarking Street-View Surface Reconstruction with a Synthetic 3D Mesh Dataset
von: Hu, Yubin, et al.
Veröffentlicht: (2024)
von: Hu, Yubin, et al.
Veröffentlicht: (2024)
Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
Space-Aware Instruction Tuning: Dataset and Benchmark for Guide Dog Robots Assisting the Visually Impaired
von: Han, ByungOk, et al.
Veröffentlicht: (2025)
von: Han, ByungOk, et al.
Veröffentlicht: (2025)
Real-Time 3D Occupancy Prediction via Geometric-Semantic Disentanglement
von: He, Yulin, et al.
Veröffentlicht: (2024)
von: He, Yulin, et al.
Veröffentlicht: (2024)
OB3D: A New Dataset for Benchmarking Omnidirectional 3D Reconstruction Using Blender
von: Ito, Shintaro, et al.
Veröffentlicht: (2025)
von: Ito, Shintaro, et al.
Veröffentlicht: (2025)
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
von: Yu, Zhengdi, et al.
Veröffentlicht: (2023)
von: Yu, Zhengdi, et al.
Veröffentlicht: (2023)
Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation
von: He, Liu, et al.
Veröffentlicht: (2025)
von: He, Liu, et al.
Veröffentlicht: (2025)
A High-Quality Text-Rich Image Instruction Tuning Dataset via Hybrid Instruction Generation
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
IL3D: A Large-Scale Indoor Layout Dataset for LLM-Driven 3D Scene Generation
von: Zhou, Wenxu, et al.
Veröffentlicht: (2025)
von: Zhou, Wenxu, et al.
Veröffentlicht: (2025)
Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
Semi-supervised 3D Object Detection with PatchTeacher and PillarMix
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
Realistic Surgical Image Dataset Generation Based On 3D Gaussian Splatting
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
Fine-Grained Urban Flow Inference with Multi-scale Representation Learning
von: Yuan, Shilu, et al.
Veröffentlicht: (2024)
von: Yuan, Shilu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RS3DBench: A Comprehensive Benchmark for 3D Spatial Perception in Remote Sensing
von: Wang, Jiayu, et al.
Veröffentlicht: (2025) -
Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models
von: Zhang, Jiyao, et al.
Veröffentlicht: (2026) -
A Self-Conditioned Representation Guided Diffusion Model for Realistic Text-to-LiDAR Scene Generation
von: Qu, Wentao, et al.
Veröffentlicht: (2025) -
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2025) -
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
von: Liu, Dingning, et al.
Veröffentlicht: (2024)