Towards Holistic Surgical Scene Graph
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Shin, Jongmin, Cho, Enki, Kim, Ka Young, Kim, Jung Yong, Kim, Seong Tae, Oh, Namkee |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?
par: Shin, Jongmin, et autres
Publié: (2026)
par: Shin, Jongmin, et autres
Publié: (2026)
SurgX: Neuron-Concept Association for Explainable Surgical Phase Recognition
par: Kim, Ka Young, et autres
Publié: (2025)
par: Kim, Ka Young, et autres
Publié: (2025)
CurConMix+: A Unified Spatio-Temporal Framework for Hierarchical Surgical Workflow Understanding
par: Jeon, Yongjun, et autres
Publié: (2026)
par: Jeon, Yongjun, et autres
Publié: (2026)
A generalizable foundation model for intraoperative understanding across surgical procedures
par: Park, Kanggil, et autres
Publié: (2026)
par: Park, Kanggil, et autres
Publié: (2026)
HDR-NSFF: High Dynamic Range Neural Scene Flow Fields
par: Dong-Yeon, Shin, et autres
Publié: (2026)
par: Dong-Yeon, Shin, et autres
Publié: (2026)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
par: Oh, Youngmin, et autres
Publié: (2024)
par: Oh, Youngmin, et autres
Publié: (2024)
GOTPR: General Outdoor Text-based Place Recognition Using Scene Graph Retrieval with OpenStreetMap
par: Jung, Donghwi, et autres
Publié: (2025)
par: Jung, Donghwi, et autres
Publié: (2025)
Environmental Change Detection: Toward a Practical Task of Scene Change Detection
par: Cho, Kyusik, et autres
Publié: (2025)
par: Cho, Kyusik, et autres
Publié: (2025)
SceneLinker: Compositional 3D Scene Generation via Semantic Scene Graph from RGB Sequences
par: Kim, Seok-Young, et autres
Publié: (2026)
par: Kim, Seok-Young, et autres
Publié: (2026)
AURA: Development and Validation of an Augmented Unplanned Removal Alert System using Synthetic ICU Videos
par: Seo, Junhyuk, et autres
Publié: (2025)
par: Seo, Junhyuk, et autres
Publié: (2025)
Towards Holistic Surgical Scene Understanding
par: Valderrama, Natalia, et autres
Publié: (2022)
par: Valderrama, Natalia, et autres
Publié: (2022)
WWW: A Unified Framework for Explaining What, Where and Why of Neural Networks by Interpretation of Neuron Concepts
par: Ahn, Yong Hyun, et autres
Publié: (2024)
par: Ahn, Yong Hyun, et autres
Publié: (2024)
Mask-Free Neuron Concept Annotation for Interpreting Neural Networks in Medical Domain
par: Kim, Hyeon Bae, et autres
Publié: (2024)
par: Kim, Hyeon Bae, et autres
Publié: (2024)
Factorized Multi-Resolution HashGrid for Efficient Neural Radiance Fields: Execution on Edge-Devices
par: Jun-Seong, Kim, et autres
Publié: (2026)
par: Jun-Seong, Kim, et autres
Publié: (2026)
Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion
par: Jang, Oh-Tae, et autres
Publié: (2025)
par: Jang, Oh-Tae, et autres
Publié: (2025)
SoundBrush: Sound as a Brush for Visual Scene Editing
par: Sung-Bin, Kim, et autres
Publié: (2024)
par: Sung-Bin, Kim, et autres
Publié: (2024)
IRASNet: Improved Feature-Level Clutter Reduction for Domain Generalized SAR-ATR
par: Jang, Oh-Tae, et autres
Publié: (2024)
par: Jang, Oh-Tae, et autres
Publié: (2024)
Balancing Efficiency and Quality: MoEISR for Arbitrary-Scale Image Super-Resolution
par: Oh, Young Jae, et autres
Publié: (2023)
par: Oh, Young Jae, et autres
Publié: (2023)
ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation
par: Kim, Shin Seong, et autres
Publié: (2025)
par: Kim, Shin Seong, et autres
Publié: (2025)
Finding 3D Scene Analogies with Multimodal Foundation Models
par: Kim, Junho, et autres
Publié: (2025)
par: Kim, Junho, et autres
Publié: (2025)
Zero-Shot Scene Change Detection
par: Cho, Kyusik, et autres
Publié: (2024)
par: Cho, Kyusik, et autres
Publié: (2024)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
par: Abdullah, Abdullah, et autres
Publié: (2025)
par: Abdullah, Abdullah, et autres
Publié: (2025)
LLaVA Needs More Knowledge: Retrieval Augmented Natural Language Generation with Knowledge Graph for Explaining Thoracic Pathologies
par: Hamza, Ameer, et autres
Publié: (2024)
par: Hamza, Ameer, et autres
Publié: (2024)
TWLV-I: Analysis and Insights from Holistic Evaluation on Video Foundation Models
par: Lee, Hyeongmin, et autres
Publié: (2024)
par: Lee, Hyeongmin, et autres
Publié: (2024)
Surgical Video Understanding with Label Interpolation
par: Kim, Garam, et autres
Publié: (2025)
par: Kim, Garam, et autres
Publié: (2025)
4D Scaffold Gaussian Splatting with Dynamic-Aware Anchor Growing for Efficient and High-Fidelity Dynamic Scene Reconstruction
par: Cho, Woong Oh, et autres
Publié: (2024)
par: Cho, Woong Oh, et autres
Publié: (2024)
Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph Prediction
par: Heo, KunHo, et autres
Publié: (2025)
par: Heo, KunHo, et autres
Publié: (2025)
Body-Hand Modality Expertized Networks with Cross-attention for Fine-grained Skeleton Action Recognition
par: Cho, Seungyeon, et autres
Publié: (2025)
par: Cho, Seungyeon, et autres
Publié: (2025)
BHaRNet: Reliability-Aware Body-Hand Modality Expertized Networks for Fine-grained Skeleton Action Recognition
par: Cho, Seungyeon, et autres
Publié: (2026)
par: Cho, Seungyeon, et autres
Publié: (2026)
Retrieval-Augmented Natural Language Reasoning for Explainable Visual Question Answering
par: Lim, Su Hyeon, et autres
Publié: (2024)
par: Lim, Su Hyeon, et autres
Publié: (2024)
Geometry-Aware Scene Configurations for Novel View Synthesis
par: Kim, Minkwan, et autres
Publié: (2025)
par: Kim, Minkwan, et autres
Publié: (2025)
UCMNet: Uncertainty-Aware Context Memory Network for Under-Display Camera Image Restoration
par: Kim, Daehyun, et autres
Publié: (2026)
par: Kim, Daehyun, et autres
Publié: (2026)
mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval
par: Kim, Kyeong Seon, et autres
Publié: (2026)
par: Kim, Kyeong Seon, et autres
Publié: (2026)
TopoOR: A Unified Topological Scene Representation for the Operating Room
par: Wang, Tony Danjun, et autres
Publié: (2026)
par: Wang, Tony Danjun, et autres
Publié: (2026)
LTGS: Long-Term Gaussian Scene Chronology From Sparse View Updates
par: Kim, Minkwan, et autres
Publié: (2025)
par: Kim, Minkwan, et autres
Publié: (2025)
Towards Generalizable Scene Change Detection
par: Kim, Jaewoo, et autres
Publié: (2024)
par: Kim, Jaewoo, et autres
Publié: (2024)
SA-ResGS: Self-Augmented Residual 3D Gaussian Splatting for Next Best View Selection
par: Jun-Seong, Kim, et autres
Publié: (2026)
par: Jun-Seong, Kim, et autres
Publié: (2026)
SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding
par: Kim, Seongyong, et autres
Publié: (2025)
par: Kim, Seongyong, et autres
Publié: (2025)
HiCM$^2$: Hierarchical Compact Memory Modeling for Dense Video Captioning
par: Kim, Minkuk, et autres
Publié: (2024)
par: Kim, Minkuk, et autres
Publié: (2024)
Do You Remember? Dense Video Captioning with Cross-Modal Memory Retrieval
par: Kim, Minkuk, et autres
Publié: (2024)
par: Kim, Minkuk, et autres
Publié: (2024)
Documents similaires
-
SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?
par: Shin, Jongmin, et autres
Publié: (2026) -
SurgX: Neuron-Concept Association for Explainable Surgical Phase Recognition
par: Kim, Ka Young, et autres
Publié: (2025) -
CurConMix+: A Unified Spatio-Temporal Framework for Hierarchical Surgical Workflow Understanding
par: Jeon, Yongjun, et autres
Publié: (2026) -
A generalizable foundation model for intraoperative understanding across surgical procedures
par: Park, Kanggil, et autres
Publié: (2026) -
HDR-NSFF: High Dynamic Range Neural Scene Flow Fields
par: Dong-Yeon, Shin, et autres
Publié: (2026)