ESGNN: Towards Equivariant Scene Graph Neural Network for 3D Scene Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pham, Quang P. M., Nguyen, Khoi T. N., Ngo, Lan C., Do, Truong, Hy, Truong Son |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding
von: Pham, Quang P. M., et al.
Veröffentlicht: (2024)
von: Pham, Quang P. M., et al.
Veröffentlicht: (2024)
Stable Messenger: Steganography for Message-Concealed Image Generation
von: Nguyen, Quang, et al.
Veröffentlicht: (2023)
von: Nguyen, Quang, et al.
Veröffentlicht: (2023)
IQBench: How "Smart'' Are Vision-Language Models? A Study with Human IQ Tests
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
LP-OVOD: Open-Vocabulary Object Detection by Linear Probing
von: Pham, Chau, et al.
Veröffentlicht: (2023)
von: Pham, Chau, et al.
Veröffentlicht: (2023)
SilVar-Med: A Speech-Driven Visual Language Model for Explainable Abnormality Detection in Medical Imaging
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
MOB-GCN: A Novel Multiscale Object-Based Graph Neural Network for Hyperspectral Image Classification
von: Yang, Tuan-Anh, et al.
Veröffentlicht: (2025)
von: Yang, Tuan-Anh, et al.
Veröffentlicht: (2025)
SilVar: Speech Driven Multimodal Model for Reasoning Visual Question Answering and Object Localization
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2024)
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2024)
A Novel Framework for Automated Explain Vision Model Using Vision-Language Models
von: Nguyen, Phu-Vinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phu-Vinh, et al.
Veröffentlicht: (2025)
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
von: Pham, Duc-Hai, et al.
Veröffentlicht: (2024)
von: Pham, Duc-Hai, et al.
Veröffentlicht: (2024)
Schrodinger AI: A Unified Spectral-Dynamical Framework for Classification, Reasoning, and Operator-Based Generalization
von: Nguyen, Truong Son
Veröffentlicht: (2025)
von: Nguyen, Truong Son
Veröffentlicht: (2025)
Halfway to 3D: Ensembling 2.5D and 3D Models for Robust COVID-19 CT Diagnosis
von: Yang, Tuan-Anh, et al.
Veröffentlicht: (2026)
von: Yang, Tuan-Anh, et al.
Veröffentlicht: (2026)
MonoSelfRecon: Purely Self-Supervised Explicit Generalizable 3D Reconstruction of Indoor Scenes from Monocular RGB Views
von: Li, Runfa, et al.
Veröffentlicht: (2024)
von: Li, Runfa, et al.
Veröffentlicht: (2024)
RelWitness: Open-Vocabulary 3D Scene Graph Generation with Visual-Geometric Relation Witnesses
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
Text-to-3D Generation using Jensen-Shannon Score Distillation
von: Do, Khoi, et al.
Veröffentlicht: (2025)
von: Do, Khoi, et al.
Veröffentlicht: (2025)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
von: Truong, Khang, et al.
Veröffentlicht: (2025)
von: Truong, Khang, et al.
Veröffentlicht: (2025)
A Deep-Learning Framework for Land-Sliding Classification from Remote Sensing Image
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
Toward Scene Graph and Layout Guided Complex 3D Scene Generation
von: Huang, Yu-Hsiang, et al.
Veröffentlicht: (2024)
von: Huang, Yu-Hsiang, et al.
Veröffentlicht: (2024)
DiFlow-TTS: Compact and Low-Latency Zero-Shot Text-to-Speech with Factorized Discrete Flow Matching
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2025)
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2025)
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
AffordMatcher: Affordance Learning in 3D Scenes from Visual Signifiers
von: Vu, Nghia, et al.
Veröffentlicht: (2026)
von: Vu, Nghia, et al.
Veröffentlicht: (2026)
GaussianGraph: 3D Gaussian-based Scene Graph Generation for Open-world Scene Understanding
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
Multi-Perspective Data Augmentation for Few-shot Object Detection
von: Vu, Anh-Khoa Nguyen, et al.
Veröffentlicht: (2025)
von: Vu, Anh-Khoa Nguyen, et al.
Veröffentlicht: (2025)
SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation
von: Pham, Duc-Hai, et al.
Veröffentlicht: (2024)
von: Pham, Duc-Hai, et al.
Veröffentlicht: (2024)
HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2023)
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2023)
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
Open-Vocabulary Octree-Graph for 3D Scene Understanding
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
von: Truong, Quang-Trung, et al.
Veröffentlicht: (2024)
von: Truong, Quang-Trung, et al.
Veröffentlicht: (2024)
SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
von: Huang, Ting, et al.
Veröffentlicht: (2025)
von: Huang, Ting, et al.
Veröffentlicht: (2025)
SceneGPT: A Language Model for 3D Scene Understanding
von: Chandhok, Shivam
Veröffentlicht: (2024)
von: Chandhok, Shivam
Veröffentlicht: (2024)
SIGMA: A Physics-Based Benchmark for Gas Chimney Understanding in Seismic Images
von: Truong, Bao, et al.
Veröffentlicht: (2026)
von: Truong, Bao, et al.
Veröffentlicht: (2026)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
Explainable Scene Understanding with Qualitative Representations and Graph Neural Networks
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
HIS-GPT: Towards 3D Human-In-Scene Multimodal Understanding
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
SceneProp: Combining Neural Network and Markov Random Field for Scene-Graph Grounding
von: Otani, Keita, et al.
Veröffentlicht: (2025)
von: Otani, Keita, et al.
Veröffentlicht: (2025)
DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding
von: Ge, Luzhou, et al.
Veröffentlicht: (2026)
von: Ge, Luzhou, et al.
Veröffentlicht: (2026)
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2026)
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding
von: Pham, Quang P. M., et al.
Veröffentlicht: (2024) -
Stable Messenger: Steganography for Message-Concealed Image Generation
von: Nguyen, Quang, et al.
Veröffentlicht: (2023) -
IQBench: How "Smart'' Are Vision-Language Models? A Study with Human IQ Tests
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025) -
LP-OVOD: Open-Vocabulary Object Detection by Linear Probing
von: Pham, Chau, et al.
Veröffentlicht: (2023) -
SilVar-Med: A Speech-Driven Visual Language Model for Explainable Abnormality Detection in Medical Imaging
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)