Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Ming, Yu, Zhengdi, Tang, Feilong, Chen, Kaiwen, Li, Yulong, Razzak, Imran, He, Junjun, Birdal, Tolga, Zhou, Kaijing, Ge, Zongyuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024)
by: Yu, Zhengdi, et al.
Published: (2024)
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
by: Wang, Xinkun, et al.
Published: (2025)
by: Wang, Xinkun, et al.
Published: (2025)
ScalingNoise: Scaling Inference-Time Search for Generating Infinite Videos
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
by: Yu, Zhengdi, et al.
Published: (2023)
by: Yu, Zhengdi, et al.
Published: (2023)
Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
by: Liu, Xiwei, et al.
Published: (2026)
by: Liu, Xiwei, et al.
Published: (2026)
Towards Efficient Medical Reasoning with Minimal Fine-Tuning Data
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
MSWAL: 3D Multi-class Segmentation of Whole Abdominal Lesions Dataset
by: Wu, Zhaodong, et al.
Published: (2025)
by: Wu, Zhaodong, et al.
Published: (2025)
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
by: Zhong, Yiheng, et al.
Published: (2025)
by: Zhong, Yiheng, et al.
Published: (2025)
Discriminating retinal microvascular and neuronal differences related to migraines: Deep Learning based Crossectional Study
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
TAGS: A Test-Time Generalist-Specialist Framework with Retrieval-Augmented Reasoning and Verification
by: Wu, Jianghao, et al.
Published: (2025)
by: Wu, Jianghao, et al.
Published: (2025)
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
by: Xue, Haochen, et al.
Published: (2025)
by: Xue, Haochen, et al.
Published: (2025)
Rhythm of Opinion: A Hawkes-Graph Framework for Dynamic Propagation Analysis
by: Li, Yulong, et al.
Published: (2025)
by: Li, Yulong, et al.
Published: (2025)
OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language Pretraining
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
HOG-Diff: Higher-Order Guided Diffusion for Graph Generation
by: Huang, Yiming, et al.
Published: (2025)
by: Huang, Yiming, et al.
Published: (2025)
Geometric Neural Distance Fields for Learning Human Motion Priors
by: Yu, Zhengdi, et al.
Published: (2025)
by: Yu, Zhengdi, et al.
Published: (2025)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
by: Tang, Feilong, et al.
Published: (2025)
by: Tang, Feilong, et al.
Published: (2025)
Phenome-Wide Multi-Omics Integration Uncovers Distinct Archetypes of Human Aging
by: Li, Huifa, et al.
Published: (2025)
by: Li, Huifa, et al.
Published: (2025)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
On the Interaction of Compressibility and Adversarial Robustness
by: Barsbey, Melih, et al.
Published: (2025)
by: Barsbey, Melih, et al.
Published: (2025)
Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)
by: Bastian, Lennart, et al.
Published: (2025)
by: Bastian, Lennart, et al.
Published: (2025)
A Machine Learning Approach to Predict Biological Age and its Longitudinal Drivers
by: Dunbayeva, Nazira, et al.
Published: (2025)
by: Dunbayeva, Nazira, et al.
Published: (2025)
UV-free Texture Generation with Denoising and Geodesic Heat Diffusions
by: Foti, Simone, et al.
Published: (2024)
by: Foti, Simone, et al.
Published: (2024)
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP
by: Xu, Zhongxing, et al.
Published: (2024)
by: Xu, Zhongxing, et al.
Published: (2024)
SAM-DCE: Addressing Token Uniformity and Semantic Over-Smoothing in Medical Segmentation
by: Hu, Yingzhen, et al.
Published: (2025)
by: Hu, Yingzhen, et al.
Published: (2025)
Diffusion Model Driven Test-Time Image Adaptation for Robust Skin Lesion Classification
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making
by: Liu, Yize, et al.
Published: (2026)
by: Liu, Yize, et al.
Published: (2026)
COMPOSE: Hypergraph Cover Optimization for Multi-view 3D Human Pose Estimation
by: Wang, Tony Danjun, et al.
Published: (2026)
by: Wang, Tony Danjun, et al.
Published: (2026)
Fun with Flags: Robust Principal Directions via Flag Manifolds
by: Mankovich, Nathan, et al.
Published: (2024)
by: Mankovich, Nathan, et al.
Published: (2024)
DeepChest: Dynamic Gradient-Free Task Weighting for Effective Multi-Task Learning in Chest X-ray Classification
by: Mohamed, Youssef, et al.
Published: (2025)
by: Mohamed, Youssef, et al.
Published: (2025)
Parallelised Differentiable Straightest Geodesics for 3D Meshes
by: Verninas, Hippolyte, et al.
Published: (2026)
by: Verninas, Hippolyte, et al.
Published: (2026)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025)
by: Liu, Xiwei, et al.
Published: (2025)
Beyond Words: AuralLLM and SignMST-C for Sign Language Production and Bidirectional Accessibility
by: Li, Yulong, et al.
Published: (2025)
by: Li, Yulong, et al.
Published: (2025)
SAM-aware Test-time Adaptation for Universal Medical Image Segmentation
by: Wu, Jianghao, et al.
Published: (2025)
by: Wu, Jianghao, et al.
Published: (2025)
Generalization at the Edge of Stability
by: Tuci, Mario, et al.
Published: (2026)
by: Tuci, Mario, et al.
Published: (2026)
Similar Items
-
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024) -
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
by: Wang, Xinkun, et al.
Published: (2025) -
ScalingNoise: Scaling Inference-Time Search for Generating Infinite Videos
by: Yang, Haolin, et al.
Published: (2025) -
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
by: Yu, Zhengdi, et al.
Published: (2023) -
Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
by: Li, Wei, et al.
Published: (2025)