ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yeshwanth, Chandan, Rozenberszki, David, Dai, Angela |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
von: Rozenberszki, David, et al.
Veröffentlicht: (2023)
von: Rozenberszki, David, et al.
Veröffentlicht: (2023)
Lookalike3D: Seeing Double in 3D
von: Yeshwanth, Chandan, et al.
Veröffentlicht: (2026)
von: Yeshwanth, Chandan, et al.
Veröffentlicht: (2026)
TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes
von: Jin, Bu, et al.
Veröffentlicht: (2024)
von: Jin, Bu, et al.
Veröffentlicht: (2024)
DCSEG: Decoupled 3D Open-Set Segmentation using Gaussian Splatting
von: Wiedmann, Luis, et al.
Veröffentlicht: (2024)
von: Wiedmann, Luis, et al.
Veröffentlicht: (2024)
DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB Image
von: Gao, Daoyi, et al.
Veröffentlicht: (2023)
von: Gao, Daoyi, et al.
Veröffentlicht: (2023)
A Comprehensive Survey of 3D Dense Captioning: Localizing and Describing Objects in 3D Scenes
von: Yu, Ting, et al.
Veröffentlicht: (2024)
von: Yu, Ting, et al.
Veröffentlicht: (2024)
SceneFactor: Factored Latent 3D Diffusion for Controllable 3D Scene Generation
von: Bokhovkin, Alexey, et al.
Veröffentlicht: (2024)
von: Bokhovkin, Alexey, et al.
Veröffentlicht: (2024)
OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward
von: Zhong, Chunlin, et al.
Veröffentlicht: (2025)
von: Zhong, Chunlin, et al.
Veröffentlicht: (2025)
AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
von: Chai, Wenhao, et al.
Veröffentlicht: (2024)
von: Chai, Wenhao, et al.
Veröffentlicht: (2024)
CapArena: Benchmarking and Analyzing Detailed Image Captioning in the LLM Era
von: Cheng, Kanzhi, et al.
Veröffentlicht: (2025)
von: Cheng, Kanzhi, et al.
Veröffentlicht: (2025)
OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding
von: Huang, Sheng-Yu, et al.
Veröffentlicht: (2026)
von: Huang, Sheng-Yu, et al.
Veröffentlicht: (2026)
WorldMesh: Generating Navigable Multi-Room 3D Scenes via Mesh-Conditioned Image Diffusion
von: Schneider, Manuel-Andreas, et al.
Veröffentlicht: (2026)
von: Schneider, Manuel-Andreas, et al.
Veröffentlicht: (2026)
VoCap: Video Object Captioning and Segmentation from Any Prompt
von: Uijlings, Jasper, et al.
Veröffentlicht: (2025)
von: Uijlings, Jasper, et al.
Veröffentlicht: (2025)
Seen2Scene: Completing Realistic 3D Scenes with Visibility-Guided Flow
von: Meng, Quan, et al.
Veröffentlicht: (2026)
von: Meng, Quan, et al.
Veröffentlicht: (2026)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
von: Chu, Jisheng, et al.
Veröffentlicht: (2026)
von: Chu, Jisheng, et al.
Veröffentlicht: (2026)
Curvature-Aware Captioning:Leveraging Geodesic Attention for 3D Scene Understanding
von: He, Ziyao, et al.
Veröffentlicht: (2026)
von: He, Ziyao, et al.
Veröffentlicht: (2026)
Coherent 3D Scene Diffusion From a Single RGB Image
von: Dahnert, Manuel, et al.
Veröffentlicht: (2024)
von: Dahnert, Manuel, et al.
Veröffentlicht: (2024)
Beyond Existance: Fulfill 3D Reconstructed Scenes with Pseudo Details
von: Gao, Yifei, et al.
Veröffentlicht: (2025)
von: Gao, Yifei, et al.
Veröffentlicht: (2025)
SuperGS: Consistent and Detailed 3D Super-Resolution Scene Reconstruction via Gaussian Splatting
von: Xie, Shiyun, et al.
Veröffentlicht: (2025)
von: Xie, Shiyun, et al.
Veröffentlicht: (2025)
Descrip3D: Enhancing Large Language Model-based 3D Scene Understanding with Object-Level Text Descriptions
von: Xue, Jintang, et al.
Veröffentlicht: (2025)
von: Xue, Jintang, et al.
Veröffentlicht: (2025)
LT3SD: Latent Trees for 3D Scene Diffusion
von: Meng, Quan, et al.
Veröffentlicht: (2024)
von: Meng, Quan, et al.
Veröffentlicht: (2024)
Articulate3D: Holistic Understanding of 3D Scenes as Universal Scene Description
von: Halacheva, Anna-Maria, et al.
Veröffentlicht: (2024)
von: Halacheva, Anna-Maria, et al.
Veröffentlicht: (2024)
POMA-3D: The Point Map Way to 3D Scene Understanding
von: Mao, Ye, et al.
Veröffentlicht: (2025)
von: Mao, Ye, et al.
Veröffentlicht: (2025)
SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE
von: Chen, Yongwei, et al.
Veröffentlicht: (2024)
von: Chen, Yongwei, et al.
Veröffentlicht: (2024)
Scene-Conditional 3D Object Stylization and Composition
von: Zhou, Jinghao, et al.
Veröffentlicht: (2023)
von: Zhou, Jinghao, et al.
Veröffentlicht: (2023)
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
von: Lemeshko, Andrey, et al.
Veröffentlicht: (2025)
von: Lemeshko, Andrey, et al.
Veröffentlicht: (2025)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
von: Huang, Ting, et al.
Veröffentlicht: (2025)
von: Huang, Ting, et al.
Veröffentlicht: (2025)
SceneGPT: A Language Model for 3D Scene Understanding
von: Chandhok, Shivam
Veröffentlicht: (2024)
von: Chandhok, Shivam
Veröffentlicht: (2024)
R3DS: Reality-linked 3D Scenes for Panoramic Scene Understanding
von: Wu, Qirui, et al.
Veröffentlicht: (2024)
von: Wu, Qirui, et al.
Veröffentlicht: (2024)
Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
Fed3D: Federated 3D Object Detection
von: Dai, Suyan, et al.
Veröffentlicht: (2026)
von: Dai, Suyan, et al.
Veröffentlicht: (2026)
LangSurf: Language-Embedded Surface Gaussians for 3D Scene Understanding
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
Unified Semantic Transformer for 3D Scene Understanding
von: Koch, Sebastian, et al.
Veröffentlicht: (2025)
von: Koch, Sebastian, et al.
Veröffentlicht: (2025)
A Unified Framework for 3D Scene Understanding
von: Xu, Wei, et al.
Veröffentlicht: (2024)
von: Xu, Wei, et al.
Veröffentlicht: (2024)
3D Question Answering for City Scene Understanding
von: Sun, Penglei, et al.
Veröffentlicht: (2024)
von: Sun, Penglei, et al.
Veröffentlicht: (2024)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
von: Zhou, Hongyu, et al.
Veröffentlicht: (2024)
von: Zhou, Hongyu, et al.
Veröffentlicht: (2024)
Dynamic Scene 3D Reconstruction of an Uncooperative Resident Space Object
von: Gopu, Bala Prenith Reddy, et al.
Veröffentlicht: (2025)
von: Gopu, Bala Prenith Reddy, et al.
Veröffentlicht: (2025)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
von: Rim, Patrick, et al.
Veröffentlicht: (2026)
von: Rim, Patrick, et al.
Veröffentlicht: (2026)
GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction
von: Schmid, Katharina, et al.
Veröffentlicht: (2026)
von: Schmid, Katharina, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
von: Rozenberszki, David, et al.
Veröffentlicht: (2023) -
Lookalike3D: Seeing Double in 3D
von: Yeshwanth, Chandan, et al.
Veröffentlicht: (2026) -
TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes
von: Jin, Bu, et al.
Veröffentlicht: (2024) -
DCSEG: Decoupled 3D Open-Set Segmentation using Gaussian Splatting
von: Wiedmann, Luis, et al.
Veröffentlicht: (2024) -
DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB Image
von: Gao, Daoyi, et al.
Veröffentlicht: (2023)