Surgical Visual Understanding (SurgVU) Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zia, Aneeq, Berniker, Max, Nespolo, Rogerio, Zhang, Xiaorui, Perreault, Conor, Wang, Ziheng, Mueller, Benjamin, Schmidt, Ryan, Bhattacharyya, Kiran, Liu, Xi, Jarc, Anthony |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025
von: Zia, Aneeq, et al.
Veröffentlicht: (2023)
von: Zia, Aneeq, et al.
Veröffentlicht: (2023)
MetaphorVU: Towards Metaphorical Video Understanding
von: Li, Zhuoqun, et al.
Veröffentlicht: (2026)
von: Li, Zhuoqun, et al.
Veröffentlicht: (2026)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
von: Choi, Tae-Min, et al.
Veröffentlicht: (2025)
von: Choi, Tae-Min, et al.
Veröffentlicht: (2025)
The Legacy of the Global Financial Crisis, edited by YoussefCassis and Jean‐JacquesvanHelten (2023), 219 pages + xiv front matter
von: Aneeq Sarwar
Veröffentlicht: (2024)
von: Aneeq Sarwar
Veröffentlicht: (2024)
arg-VU: Affordance Reasoning with Physics-Aware 3D Geometry for Visual Understanding in Robotic Surgery
von: Xiao, Nan, et al.
Veröffentlicht: (2026)
von: Xiao, Nan, et al.
Veröffentlicht: (2026)
Martín Kohan: Pedagogía, narrativa y deporte
von: Jimena Néspolo
Veröffentlicht: (2009)
von: Jimena Néspolo
Veröffentlicht: (2009)
VANTAGENS COMPETITIVAS DE PEQUENAS E MÉDIAS EMPRESAS COM A PARTICIPAÇÃO EM REDES DE COOPERAÇÃO: O CASO DO MERCADO TUTTO
von: Daniele Nespolo
Veröffentlicht: (2014)
von: Daniele Nespolo
Veröffentlicht: (2014)
Comportamento do consumidor: fatores que influenciam o consumo virtual nas redes sociais
von: Daniele Nespolo
Veröffentlicht: (2015)
von: Daniele Nespolo
Veröffentlicht: (2015)
Consumo consciente, meio ambiente e desenvolvimento sustentável: análise da tomada de decisão com base nas heurísticas
von: Daniele Nespolo
Veröffentlicht: (2016)
von: Daniele Nespolo
Veröffentlicht: (2016)
LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2024)
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2024)
Digital literacy in VU Libraries
von: Cazembe Kennedy, et al.
Veröffentlicht: (2024)
von: Cazembe Kennedy, et al.
Veröffentlicht: (2024)
Digital Literacy in VU Libraries
von: Cazembe Kennedy, et al.
Veröffentlicht: (2024)
von: Cazembe Kennedy, et al.
Veröffentlicht: (2024)
Bridging Vision and Language for Robust Context-Aware Surgical Point Tracking: The VL-SurgPT Dataset and Benchmark
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
SurgPub-Video: A Comprehensive Surgical Video Dataset for Enhanced Surgical Intelligence in Vision-Language Model
von: Li, Yaoqian, et al.
Veröffentlicht: (2025)
von: Li, Yaoqian, et al.
Veröffentlicht: (2025)
SurgVisAgent: Multimodal Agentic Model for Versatile Surgical Visual Enhancement
von: Lei, Zeyu, et al.
Veröffentlicht: (2025)
von: Lei, Zeyu, et al.
Veröffentlicht: (2025)
SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding
von: Drago, Mauro Orazio, et al.
Veröffentlicht: (2025)
von: Drago, Mauro Orazio, et al.
Veröffentlicht: (2025)
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos
von: Wu, Jinlin, et al.
Veröffentlicht: (2026)
von: Wu, Jinlin, et al.
Veröffentlicht: (2026)
SurgFed: Language-guided Multi-Task Federated Learning for Surgical Video Understanding
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
La "Frontera" Bonaerense en el siglo XVIII un espacio políticamente concertado: fuertes, vecinos, milicias y autoridades civilesmilitares
von: Eugenia Alicia Néspolo
Veröffentlicht: (2006)
von: Eugenia Alicia Néspolo
Veröffentlicht: (2006)
Pontos de Cultura: contribuições para a Educação Popular em Saúde na perspectiva de seus coordenadores
von: Gabriela Fabian Nespolo
Veröffentlicht: (2014)
von: Gabriela Fabian Nespolo
Veröffentlicht: (2014)
H2VU-Benchmark: A Comprehensive Benchmark for Hierarchical Holistic Video Understanding
von: Wu, Qi, et al.
Veröffentlicht: (2025)
von: Wu, Qi, et al.
Veröffentlicht: (2025)
SurgSync: Time-Synchronized Multi-Modal Data Collection Framework and Dataset for Surgical Robotics
von: Zhou, Haoying, et al.
Veröffentlicht: (2026)
von: Zhou, Haoying, et al.
Veröffentlicht: (2026)
SurgLaVi: Large-Scale Hierarchical Dataset for Surgical Vision-Language Representation Learning
von: Perez, Alejandra, et al.
Veröffentlicht: (2025)
von: Perez, Alejandra, et al.
Veröffentlicht: (2025)
SurgVidLM: Towards Multi-grained Surgical Video Understanding with Large Language Model
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
Seeing Through Smoke: Surgical Desmoking for Improved Visual Perception
von: Lu, Jingpei, et al.
Veröffentlicht: (2026)
von: Lu, Jingpei, et al.
Veröffentlicht: (2026)
VU-IVM/honeybees: 1.0.7
von: Jens de Bruijn, et al.
Veröffentlicht: (2025)
von: Jens de Bruijn, et al.
Veröffentlicht: (2025)
VU-IVM/honeybees: 1.0.8
von: Jens de Bruijn, et al.
Veröffentlicht: (2025)
von: Jens de Bruijn, et al.
Veröffentlicht: (2025)
RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering
von: Zhang, Chengyi, et al.
Veröffentlicht: (2026)
von: Zhang, Chengyi, et al.
Veröffentlicht: (2026)
SurgCUT3R: Surgical Scene-Aware Continuous Understanding of Temporal 3D Representation
von: Xu, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Xu, Kaiyuan, et al.
Veröffentlicht: (2026)
SurgLLM: A Versatile Large Multimodal Model with Spatial Focus and Temporal Awareness for Surgical Video Understanding
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
SurgWound-Bench: A Benchmark for Surgical Wound Diagnosis
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
SurgTEMP: Temporal-Aware Surgical Video Question Answering with Text-guided Visual Memory for Laparoscopic Cholecystectomy
von: Li, Shi, et al.
Veröffentlicht: (2026)
von: Li, Shi, et al.
Veröffentlicht: (2026)
SurgPETL: Parameter-Efficient Image-to-Surgical-Video Transfer Learning for Surgical Phase Recognition
von: Yang, Shu, et al.
Veröffentlicht: (2024)
von: Yang, Shu, et al.
Veröffentlicht: (2024)
LLaVA-Surg: Towards Multimodal Surgical Assistant via Structured Surgical Video Learning
von: Li, Jiajie, et al.
Veröffentlicht: (2024)
von: Li, Jiajie, et al.
Veröffentlicht: (2024)
HieraSurg: Hierarchy-Aware Diffusion Model for Surgical Video Generation
von: Biagini, Diego, et al.
Veröffentlicht: (2025)
von: Biagini, Diego, et al.
Veröffentlicht: (2025)
SurgLQA: Scalable Long-Horizon Surgical Video Question Answering
von: Guo, Diandian, et al.
Veröffentlicht: (2026)
von: Guo, Diandian, et al.
Veröffentlicht: (2026)
SurgX: Neuron-Concept Association for Explainable Surgical Phase Recognition
von: Kim, Ka Young, et al.
Veröffentlicht: (2025)
von: Kim, Ka Young, et al.
Veröffentlicht: (2025)
SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary
von: He, Jingyi, et al.
Veröffentlicht: (2026)
von: He, Jingyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025
von: Zia, Aneeq, et al.
Veröffentlicht: (2023) -
MetaphorVU: Towards Metaphorical Video Understanding
von: Li, Zhuoqun, et al.
Veröffentlicht: (2026) -
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
von: Wu, Zijian, et al.
Veröffentlicht: (2025) -
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
von: Choi, Tae-Min, et al.
Veröffentlicht: (2025) -
The Legacy of the Global Financial Crisis, edited by YoussefCassis and Jean‐JacquesvanHelten (2023), 219 pages + xiv front matter
von: Aneeq Sarwar
Veröffentlicht: (2024)