Identifying Surgical Instruments in Pedagogical Cataract Surgery Videos through an Optimized Aggregation Network
Fuente:
arXiv
Saved in:
| Main Authors: | Sinha, Sanya, Balazia, Michal, Bremond, Francois |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets
by: Agrawal, Tanay, et al.
Published: (2025)
by: Agrawal, Tanay, et al.
Published: (2025)
MVP: Multimodal Emotion Recognition based on Video and Physiological Signals
by: Strizhkova, Valeriya, et al.
Published: (2025)
by: Strizhkova, Valeriya, et al.
Published: (2025)
Swish-T : Enhancing Swish Activation with Tanh Bias for Improved Neural Network Performance
by: Seo, Youngmin, et al.
Published: (2024)
by: Seo, Youngmin, et al.
Published: (2024)
Revisiting Energy-Based Model for Out-of-Distribution Detection
by: Wu, Yifan, et al.
Published: (2024)
by: Wu, Yifan, et al.
Published: (2024)
Transforming faces into video stories -- VideoFace2.0
by: Brkljač, Branko, et al.
Published: (2025)
by: Brkljač, Branko, et al.
Published: (2025)
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
by: Ko, Hyun-kyu, et al.
Published: (2024)
by: Ko, Hyun-kyu, et al.
Published: (2024)
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
by: Lentsch, Ted, et al.
Published: (2026)
by: Lentsch, Ted, et al.
Published: (2026)
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
by: Lentsch, Ted, et al.
Published: (2024)
by: Lentsch, Ted, et al.
Published: (2024)
Enhancing Rotation-Invariant 3D Learning with Global Pose Awareness and Attention Mechanisms
by: Guo, Jiaxun, et al.
Published: (2025)
by: Guo, Jiaxun, et al.
Published: (2025)
Banana Ripeness Level Classification using a Simple CNN Model Trained with Real and Synthetic Datasets
by: Chuquimarca, Luis, et al.
Published: (2025)
by: Chuquimarca, Luis, et al.
Published: (2025)
An M-Health Algorithmic Approach to Identify and Assess Physiotherapy Exercises in Real Time
by: Kandylakis, Stylianos, et al.
Published: (2025)
by: Kandylakis, Stylianos, et al.
Published: (2025)
Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
by: Kang, Xueyang, et al.
Published: (2026)
by: Kang, Xueyang, et al.
Published: (2026)
See What You Need: Query-Aware Visual Intelligence through Reasoning-Perception Loops
by: Dong, Zixuan, et al.
Published: (2025)
by: Dong, Zixuan, et al.
Published: (2025)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
by: Zarei, Mohammad, et al.
Published: (2025)
by: Zarei, Mohammad, et al.
Published: (2025)
Archival Faces: Detection of Faces in Digitized Historical Documents
by: Vaško, Marek, et al.
Published: (2025)
by: Vaško, Marek, et al.
Published: (2025)
On the Equivalence of Regression and Classification
by: Jayadeva, et al.
Published: (2025)
by: Jayadeva, et al.
Published: (2025)
A deep learning approach to track eye movements based on events
by: Seth, Chirag, et al.
Published: (2025)
by: Seth, Chirag, et al.
Published: (2025)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
by: Zhang, Junbin, et al.
Published: (2022)
by: Zhang, Junbin, et al.
Published: (2022)
WSCIF: A Weakly-Supervised Color Intelligence Framework for Tactical Anomaly Detection in Surveillance Keyframes
by: Meng, Wei
Published: (2025)
by: Meng, Wei
Published: (2025)
HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level and Fidelity-Rich Conditions in Diffusion Models
by: Zhang, Shengkai, et al.
Published: (2024)
by: Zhang, Shengkai, et al.
Published: (2024)
Polygonizing Roof Segments from High-Resolution Aerial Images Using Yolov8-Based Edge Detection
by: Mei, Qipeng, et al.
Published: (2025)
by: Mei, Qipeng, et al.
Published: (2025)
Training-free Zero-shot Composed Image Retrieval via Weighted Modality Fusion and Similarity
by: Wu, Ren-Di, et al.
Published: (2024)
by: Wu, Ren-Di, et al.
Published: (2024)
Spectral Integrated Gradients for Coarse-to-Fine Feature Attribution
by: Kim, Soyeon, et al.
Published: (2026)
by: Kim, Soyeon, et al.
Published: (2026)
Manifold-Aligned Guided Integrated Gradients for Reliable Feature Attribution
by: Kim, Soyeon, et al.
Published: (2026)
by: Kim, Soyeon, et al.
Published: (2026)
A Landmark-Aware Visual Navigation Dataset
by: Johnson, Faith, et al.
Published: (2024)
by: Johnson, Faith, et al.
Published: (2024)
Position-Prior-Guided Network for System Matrix Super-Resolution in Magnetic Particle Imaging
by: Geng, Xuqing, et al.
Published: (2025)
by: Geng, Xuqing, et al.
Published: (2025)
DRL: Discriminative Representation Learning with Parallel Adapters for Class Incremental Learning
by: Zhan, Jiawei, et al.
Published: (2025)
by: Zhan, Jiawei, et al.
Published: (2025)
Interpretable label-free self-guided subspace clustering
by: Kopriva, Ivica
Published: (2024)
by: Kopriva, Ivica
Published: (2024)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
Fast 3D point clouds retrieval for Large-scale 3D Place Recognition
by: Zede, Chahine-Nicolas, et al.
Published: (2025)
by: Zede, Chahine-Nicolas, et al.
Published: (2025)
Extraction Of Cumulative Blobs From Dynamic Gestures
by: Naulakha, Rishabh, et al.
Published: (2025)
by: Naulakha, Rishabh, et al.
Published: (2025)
Deep Spectral Meshes: Multi-Frequency Facial Mesh Processing with Graph Neural Networks
by: Kosk, Robert, et al.
Published: (2024)
by: Kosk, Robert, et al.
Published: (2024)
Skullptor: High Fidelity 3D Head Reconstruction in Seconds with Multi-View Normal Prediction
by: Artru, Noé, et al.
Published: (2026)
by: Artru, Noé, et al.
Published: (2026)
Deep Feature Optimization for Enhanced Fish Freshness Assessment
by: Hoang, Phi-Hung, et al.
Published: (2025)
by: Hoang, Phi-Hung, et al.
Published: (2025)
HOSC: A Periodic Activation with Saturation Control for High-Fidelity Implicit Neural Representations
by: Wlodarczyk, Michal Jan, et al.
Published: (2026)
by: Wlodarczyk, Michal Jan, et al.
Published: (2026)
Dual-Teacher Ensemble Models with Double-Copy-Paste for 3D Semi-Supervised Medical Image Segmentation
by: Fa, Zhan, et al.
Published: (2024)
by: Fa, Zhan, et al.
Published: (2024)
TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection
by: Siddiqui, Yousuf Ahmed, et al.
Published: (2025)
by: Siddiqui, Yousuf Ahmed, et al.
Published: (2025)
EZ-Sort: Efficient Pairwise Comparison via Zero-Shot CLIP-Based Pre-Ordering and Human-in-the-Loop Sorting
by: Park, Yujin, et al.
Published: (2025)
by: Park, Yujin, et al.
Published: (2025)
OpenFusion++: An Open-vocabulary Real-time Scene Understanding System
by: Jin, Xiaofeng, et al.
Published: (2025)
by: Jin, Xiaofeng, et al.
Published: (2025)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
by: Zafeiri, Maria, et al.
Published: (2024)
by: Zafeiri, Maria, et al.
Published: (2024)
Similar Items
-
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets
by: Agrawal, Tanay, et al.
Published: (2025) -
MVP: Multimodal Emotion Recognition based on Video and Physiological Signals
by: Strizhkova, Valeriya, et al.
Published: (2025) -
Swish-T : Enhancing Swish Activation with Tanh Bias for Improved Neural Network Performance
by: Seo, Youngmin, et al.
Published: (2024) -
Revisiting Energy-Based Model for Out-of-Distribution Detection
by: Wu, Yifan, et al.
Published: (2024) -
Transforming faces into video stories -- VideoFace2.0
by: Brkljač, Branko, et al.
Published: (2025)