Holistic Surgical Phase Recognition with Hierarchical Input Dependent State Space Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Haoyang, Wang, Tsun-Hsuan, Lechner, Mathias, Hasani, Ramin, Eckhoff, Jennifer A., Pak, Paul, Meireles, Ozanan R., Rosman, Guy, Ban, Yutong, Rus, Daniela |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hypergraph-Transformer (HGT) for Interactive Event Prediction in Laparoscopic and Robotic Surgery
by: Yin, Lianhao, et al.
Published: (2024)
by: Yin, Lianhao, et al.
Published: (2024)
Surgical Foundation Model Leveraging Compression and Entropy Maximization for Image-Guided Surgical Assistance
by: Yin, Lianhao, et al.
Published: (2025)
by: Yin, Lianhao, et al.
Published: (2025)
Towards Cooperative Flight Control Using Visual-Attention
by: Yin, Lianhao, et al.
Published: (2022)
by: Yin, Lianhao, et al.
Published: (2022)
Learning with Chemical versus Electrical Synapses -- Does it Make a Difference?
by: Farsang, Mónika, et al.
Published: (2023)
by: Farsang, Mónika, et al.
Published: (2023)
Looped SSMs: Depth-Recurrence and Input Reshaping for Time Series Classification
by: Farsang, Mónika, et al.
Published: (2026)
by: Farsang, Mónika, et al.
Published: (2026)
Probing Multimodal LLMs as World Models for Driving
by: Sreeram, Shiva, et al.
Published: (2024)
by: Sreeram, Shiva, et al.
Published: (2024)
Parallelization of Non-linear State-Space Models: Scaling Up Liquid-Resistance Liquid-Capacitance Networks for Efficient Sequence Modeling
by: Farsang, Mónika, et al.
Published: (2025)
by: Farsang, Mónika, et al.
Published: (2025)
Human Insights Driven Latent Space for Different Driving Perspectives: A Unified Encoder for Efficient Multi-Task Inference
by: Nguyen, Huy-Dung, et al.
Published: (2024)
by: Nguyen, Huy-Dung, et al.
Published: (2024)
ReGen: Generative Robot Simulation via Inverse Design
by: Nguyen, Phat, et al.
Published: (2025)
by: Nguyen, Phat, et al.
Published: (2025)
ABNet: Attention BarrierNet for Safe and Scalable Robot Learning
by: Xiao, Wei, et al.
Published: (2024)
by: Xiao, Wei, et al.
Published: (2024)
Learning autonomous driving from aerial imagery
by: Murali, Varun, et al.
Published: (2024)
by: Murali, Varun, et al.
Published: (2024)
SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space
by: Liu, Huanrong, et al.
Published: (2026)
by: Liu, Huanrong, et al.
Published: (2026)
Optical Flow Matters: an Empirical Comparative Study on Fusing Monocular Extracted Modalities for Better Steering
by: Makiyeh, Fouad, et al.
Published: (2024)
by: Makiyeh, Fouad, et al.
Published: (2024)
Gaussian Splatting to Real World Flight Navigation Transfer with Liquid Networks
by: Quach, Alex, et al.
Published: (2024)
by: Quach, Alex, et al.
Published: (2024)
Surg-R1: A Hierarchical Reasoning Foundation Model for Scalable and Interpretable Surgical Decision Support with Multi-Center Clinical Validation
by: Jiang, Jian, et al.
Published: (2026)
by: Jiang, Jian, et al.
Published: (2026)
Surgical Scene Segmentation by Transformer With Asymmetric Feature Enhancement
by: Yuan, Cheng, et al.
Published: (2024)
by: Yuan, Cheng, et al.
Published: (2024)
Temporal Propagation of Asymmetric Feature Pyramid for Surgical Scene Segmentation
by: Yuan, Cheng, et al.
Published: (2025)
by: Yuan, Cheng, et al.
Published: (2025)
Adaptive Control in Autonomous Driving via Real-Time Recurrent RL
by: Lemmel, Julian, et al.
Published: (2026)
by: Lemmel, Julian, et al.
Published: (2026)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
by: Chahine, Makram, et al.
Published: (2024)
by: Chahine, Makram, et al.
Published: (2024)
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition
by: Yang, Shu, et al.
Published: (2024)
by: Yang, Shu, et al.
Published: (2024)
Generating Out-Of-Distribution Scenarios Using Language Models
by: Aasi, Erfan, et al.
Published: (2024)
by: Aasi, Erfan, et al.
Published: (2024)
Enhancing Steering Estimation with Semantic-Aware GNNs
by: Makiyeh, Fouad, et al.
Published: (2025)
by: Makiyeh, Fouad, et al.
Published: (2025)
Text-to-Drive: Diverse Driving Behavior Synthesis via Large Language Models
by: Nguyen, Phat, et al.
Published: (2024)
by: Nguyen, Phat, et al.
Published: (2024)
Pixel-Wise Recognition for Holistic Surgical Scene Understanding
by: Ayobi, Nicolás, et al.
Published: (2024)
by: Ayobi, Nicolás, et al.
Published: (2024)
MoSFormer: Augmenting Temporal Context with Memory of Surgery for Surgical Phase Recognition
by: Ding, Hao, et al.
Published: (2025)
by: Ding, Hao, et al.
Published: (2025)
See Less, Drive Better: Generalizable End-to-End Autonomous Driving via Foundation Models Stochastic Patch Selection
by: Mallak, Amir, et al.
Published: (2026)
by: Mallak, Amir, et al.
Published: (2026)
State-Free Inference of State-Space Models: The Transfer Function Approach
by: Parnichkun, Rom N., et al.
Published: (2024)
by: Parnichkun, Rom N., et al.
Published: (2024)
SAFe-Copilot: Unified Shared Autonomy Framework
by: Nguyen, Phat, et al.
Published: (2025)
by: Nguyen, Phat, et al.
Published: (2025)
Comentário Emmanuel Levinas, leitor de Martin Buber
by: Ozanan Vicente Carrara
Published: (2020)
by: Ozanan Vicente Carrara
Published: (2020)
Comentário a “Para uma ethical turn da tecnologia: por que Hans Jonas não é um tecnofóbico” - Simondon: um outro ponto de vista sobre a técnica
by: Ozanan Vicente Carrara
Published: (2022)
by: Ozanan Vicente Carrara
Published: (2022)
Methodological Insights and Trends: Participatory Positioning of Children and Youth in Visual Art Education Research
by: Angela Eckhoff
Published: (2025)
by: Angela Eckhoff
Published: (2025)
Neural Finite-State Machines for Surgical Phase Recognition
by: Ding, Hao, et al.
Published: (2024)
by: Ding, Hao, et al.
Published: (2024)
Understanding and Reducing the Class-Dependent Effects of Data Augmentation with A Two-Player Game Approach
by: Jiang, Yunpeng, et al.
Published: (2024)
by: Jiang, Yunpeng, et al.
Published: (2024)
See in Depth: Training-Free Surgical Scene Segmentation with Monocular Depth Priors
by: Yang, Kunyi, et al.
Published: (2025)
by: Yang, Kunyi, et al.
Published: (2025)
Exploring Latent Pathways: Enhancing the Interpretability of Autonomous Driving with a Variational Autoencoder
by: Bairouk, Anass, et al.
Published: (2024)
by: Bairouk, Anass, et al.
Published: (2024)
SR-Mamba: Effective Surgical Phase Recognition with State Space Model
by: Cao, Rui, et al.
Published: (2024)
by: Cao, Rui, et al.
Published: (2024)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
by: Hori, Toshiaki, et al.
Published: (2025)
by: Hori, Toshiaki, et al.
Published: (2025)
The SAGES Critical View of Safety Challenge: A Global Benchmark for AI-Assisted Surgical Quality Assessment
by: Alapatt, Deepak, et al.
Published: (2025)
by: Alapatt, Deepak, et al.
Published: (2025)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
by: Rollman, John C., et al.
Published: (2025)
by: Rollman, John C., et al.
Published: (2025)
Similar Items
-
Hypergraph-Transformer (HGT) for Interactive Event Prediction in Laparoscopic and Robotic Surgery
by: Yin, Lianhao, et al.
Published: (2024) -
Surgical Foundation Model Leveraging Compression and Entropy Maximization for Image-Guided Surgical Assistance
by: Yin, Lianhao, et al.
Published: (2025) -
Towards Cooperative Flight Control Using Visual-Attention
by: Yin, Lianhao, et al.
Published: (2022) -
Learning with Chemical versus Electrical Synapses -- Does it Make a Difference?
by: Farsang, Mónika, et al.
Published: (2023) -
Looped SSMs: Depth-Recurrence and Input Reshaping for Time Series Classification
by: Farsang, Mónika, et al.
Published: (2026)