Towards Suturing World Models: Learning Predictive Models for Robotic Surgical Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Turkcan, Mehmet Kerem, Ballo, Mattia, Filicori, Filippo, Kostic, Zoran |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data-Driven Traffic Simulation for an Intersection in a Metropolis
by: Zang, Chengbo, et al.
Published: (2024)
by: Zang, Chengbo, et al.
Published: (2024)
A Real-Time Bike-Pedestrian Safety System with Wide-Angle Perception and Evaluation Testbed for Urban Intersections
by: Turkcan, Mehmet Kerem
Published: (2026)
by: Turkcan, Mehmet Kerem
Published: (2026)
Detect Anything in Real Time: From Single-Prompt Segmentation to Multi-Class Detection
by: Turkcan, Mehmet Kerem
Published: (2026)
by: Turkcan, Mehmet Kerem
Published: (2026)
A Vision-Based Analysis of Congestion Pricing in New York City
by: Turkcan, Mehmet Kerem, et al.
Published: (2026)
by: Turkcan, Mehmet Kerem, et al.
Published: (2026)
Boundless: Generating Photorealistic Synthetic Data for Object Detection in Urban Streetscapes
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
Constellation Dataset: Benchmarking High-Altitude Object Detection for an Urban Intersection
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
G-SHARP: Gaussian Surgical Hardware Accelerated Real-time Pipeline
by: Nath, Vishwesh, et al.
Published: (2025)
by: Nath, Vishwesh, et al.
Published: (2025)
The Streetscape Application Services Stack (SASS): Towards a Distributed Sensing Architecture for Urban Applications
by: Pargoo, Navid Salami, et al.
Published: (2024)
by: Pargoo, Navid Salami, et al.
Published: (2024)
Cosmos-H-Surgical: Learning Surgical Robot Policies from Videos via World Modeling
by: He, Yufan, et al.
Published: (2025)
by: He, Yufan, et al.
Published: (2025)
WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens
by: Wang, Xiaofeng, et al.
Published: (2024)
by: Wang, Xiaofeng, et al.
Published: (2024)
A Step Toward World Models: A Survey on Robotic Manipulation
by: Zhang, Peng-Fei, et al.
Published: (2025)
by: Zhang, Peng-Fei, et al.
Published: (2025)
Benchmarking and Enhancing Surgical Phase Recognition Models for Robotic-Assisted Esophagectomy
by: Li, Yiping, et al.
Published: (2024)
by: Li, Yiping, et al.
Published: (2024)
World Model for Robot Learning: A Comprehensive Survey
by: Hou, Bohan, et al.
Published: (2026)
by: Hou, Bohan, et al.
Published: (2026)
Occupancy World Model for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
Research on World Models Is Not Merely Injecting World Knowledge into Specific Tasks
by: Zeng, Bohan, et al.
Published: (2026)
by: Zeng, Bohan, et al.
Published: (2026)
DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
by: Deng, Yueci, et al.
Published: (2026)
by: Deng, Yueci, et al.
Published: (2026)
Causal World Modeling for Robot Control
by: Li, Lin, et al.
Published: (2026)
by: Li, Lin, et al.
Published: (2026)
Toward Zero-Shot Learning for Visual Dehazing of Urological Surgical Robots
by: Wu, Renkai, et al.
Published: (2024)
by: Wu, Renkai, et al.
Published: (2024)
WorldSimBench: Towards Video Generation Models as World Simulators
by: Qin, Yiran, et al.
Published: (2024)
by: Qin, Yiran, et al.
Published: (2024)
Surgical Vision World Model
by: Koju, Saurabh, et al.
Published: (2025)
by: Koju, Saurabh, et al.
Published: (2025)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
by: Barcellona, Leonardo, et al.
Published: (2024)
by: Barcellona, Leonardo, et al.
Published: (2024)
SWEET: Sparse World Modeling with Image Editing for Embodied Task Execution
by: Song, Yiren, et al.
Published: (2026)
by: Song, Yiren, et al.
Published: (2026)
Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks
by: Ren, Tianhe, et al.
Published: (2024)
by: Ren, Tianhe, et al.
Published: (2024)
SurgFed: Language-guided Multi-Task Federated Learning for Surgical Video Understanding
by: Fang, Zheng, et al.
Published: (2026)
by: Fang, Zheng, et al.
Published: (2026)
TurboTrain: Towards Efficient and Balanced Multi-Task Learning for Multi-Agent Perception and Prediction
by: Zhou, Zewei, et al.
Published: (2025)
by: Zhou, Zewei, et al.
Published: (2025)
Latent Video Prediction Learns Better World Models
by: Alrasheed, Ali J, et al.
Published: (2026)
by: Alrasheed, Ali J, et al.
Published: (2026)
Robot Learning from a Physical World Model
by: Mao, Jiageng, et al.
Published: (2025)
by: Mao, Jiageng, et al.
Published: (2025)
Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models
by: Jin, Juseong, et al.
Published: (2024)
by: Jin, Juseong, et al.
Published: (2024)
Benchmarking CNN- and Transformer-Based Models for Surgical Instrument Segmentation in Robotic-Assisted Surgery
by: Ameli, Sara
Published: (2026)
by: Ameli, Sara
Published: (2026)
TeleWorld: Towards Dynamic Multimodal Synthesis with a 4D World Model
by: Chen, Yabo, et al.
Published: (2025)
by: Chen, Yabo, et al.
Published: (2025)
Omni-WorldBench: Towards a Comprehensive Interaction-Centric Evaluation for World Models
by: Wu, Meiqi, et al.
Published: (2026)
by: Wu, Meiqi, et al.
Published: (2026)
Persistent Robot World Models: Stabilizing Multi-Step Rollouts via Reinforcement Learning
by: Bardhan, Jai, et al.
Published: (2026)
by: Bardhan, Jai, et al.
Published: (2026)
SCOUT+: Towards Practical Task-Driven Drivers' Gaze Prediction
by: Kotseruba, Iuliia, et al.
Published: (2024)
by: Kotseruba, Iuliia, et al.
Published: (2024)
Towards Holistic Surgical Scene Graph
by: Shin, Jongmin, et al.
Published: (2025)
by: Shin, Jongmin, et al.
Published: (2025)
Surgical Depth Anything: Depth Estimation for Surgical Scenes using Foundation Models
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Benchmarking Vision, Language, & Action Models on Robotic Learning Tasks
by: Guruprasad, Pranav, et al.
Published: (2024)
by: Guruprasad, Pranav, et al.
Published: (2024)
WorldCompass: Reinforcement Learning for Long-Horizon World Models
by: Wang, Zehan, et al.
Published: (2026)
by: Wang, Zehan, et al.
Published: (2026)
Efficient Surgical Robotic Instrument Pose Reconstruction in Real World Conditions Using Unified Feature Detection
by: Liang, Zekai, et al.
Published: (2025)
by: Liang, Zekai, et al.
Published: (2025)
MuDreamer: Learning Predictive World Models without Reconstruction
by: Burchi, Maxime, et al.
Published: (2024)
by: Burchi, Maxime, et al.
Published: (2024)
Inference-Time Enhancement of Generative Robot Policies via Predictive World Modeling
by: Qi, Han, et al.
Published: (2025)
by: Qi, Han, et al.
Published: (2025)
Similar Items
-
Data-Driven Traffic Simulation for an Intersection in a Metropolis
by: Zang, Chengbo, et al.
Published: (2024) -
A Real-Time Bike-Pedestrian Safety System with Wide-Angle Perception and Evaluation Testbed for Urban Intersections
by: Turkcan, Mehmet Kerem
Published: (2026) -
Detect Anything in Real Time: From Single-Prompt Segmentation to Multi-Class Detection
by: Turkcan, Mehmet Kerem
Published: (2026) -
A Vision-Based Analysis of Congestion Pricing in New York City
by: Turkcan, Mehmet Kerem, et al.
Published: (2026) -
Boundless: Generating Photorealistic Synthetic Data for Object Detection in Urban Streetscapes
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)