Future Frame Prediction for Robot-Assisted Surgery

机译：未来机器人辅助手术的帧预测

获取原文

页面导航

摘要
著录项
相似文献
相关主题

摘要

Predicting future frames for robotic surgical video is an interesting, important yet extremely challenging problem, given that the operative tasks may have complex dynamics. Existing approaches on future prediction of natural videos were based on either deterministic models or stochastic models, including deep recurrent neural networks, optical flow, and latent space modeling. However, the potential in predicting meaningful movements of robots with dual arms in surgical scenarios has not been tapped so far, which is typically more challenging than forecasting independent motions of one arm robots in natural scenarios. In this paper, we propose a ternary prior guided variational autoen-coder (TPG-VAE) model for future frame prediction in robotic surgical video sequences. Besides content distribution, our model learns motion distribution, which is novel to handle the small movements of surgical tools. Furthermore, we add the invariant prior information from the gesture class into the generation process to constrain the latent space of our model. To our best knowledge, this is the first time that the future frames of dual arm robots are predicted considering their unique characteristics relative to general robotic videos. Experiments demonstrate that our model gains more stable and realistic future frame prediction scenes with the suturing task on the public JIGSAWS dataset.

机译：考虑到操作任务可能具有复杂的动态，预测机器人外科视频的未来框架是一个有趣的，重要但极具挑战性的问题。现有的未来预测自然视频的方法是基于确定性模型或随机模型，包括深度经常性神经网络，光学流量和潜空间建模。然而，到目前为止，目前尚未立即预测具有双臂的机器人有意义的机器人的潜力，这通常比预测自然情景中的一个手臂机器人的独立运动更具挑战性。在本文中，我们提出了一种用于组织外科视频序列的未来帧预测的三元先前引导变分自动编码器（TPG-VAE）模型。除了内容分布外，我们的模型还学习运动分布，这是处理手术工具的小型运动的新颖。此外，我们将来自手势类的不变性先前信息添加到生成过程中以限制我们模型的潜在空间。为了我们的最佳知识，这是第一次预测双臂机器人的未来框架，考虑到普通机器人视频的独特特征。实验表明，我们的模型在公共拼写数据集上与缝线任务进行了更稳定和现实的未来框架预测场景。

著录项

来源
《International conference on information processing in medical imaging》|2021年|533-544|共12页
会议地点
作者
Xiaojie Gao; Yueming Jin; Zixu Zhao; Qi Dou; Pheng-Ann Heng;
展开▼
作者单位

展开▼
会议组织
原文格式 PDF
正文语种
中图分类
关键词
Video prediction for medical robotics; Deep learning for visual perception; Medical robots and systems;

机译：医学机器人的视频预测;深入学习视觉感知;医疗机器人和系统;

相似文献

外文文献
中文文献
专利

1. Electrode placement accuracy in robot-assisted epilepsy surgery: A comparison of different referencing techniques including frame-based CT versus facial laser scan based on CT or MRI [J] . Spyrantis Andrea, Cattani Adriano, Woebbecke Tirza, Epilepsy & behavior: E&B . 2019,第期

机译：机器人辅助癫痫手术中的电极放置精度：基于CT或MRI的基于帧的CT与面部激光扫描的不同参考技术的比较
2. Techniques for Stereotactic Neurosurgery: Beyond the Frame, Toward the Intraoperative Magnetic Resonance Imaging–Guided and Robot-Assisted Approaches [J] . Ziyan Guo, Martin Chun-Wing Leong, Hao Su, World neurosurgery . 2018,第期

机译：立体定向神经外科的技术：超越框架，朝向朝内磁共振成像引导和机器人辅助方法
3. Comparative Study of Robot-Assisted versus Conventional Frame-Based Deep Brain Stimulation Stereotactic Neurosurgery [J] . Neudorfer Clemens, Hunsche Stefan, Hellmich Martin, Stereotactic and Functional Neurosurgery: Official Journal of the World Society for Stereotactic and Functional Neurosurgery . 2018,第5期

机译：机器人辅助与常规帧深脑刺激立体定向神经外科的比较研究
4. daVinciNet: Joint Prediction of Motion and Surgical State in Robot-Assisted Surgery [C] . Yidan Qin, Seyedshams Feyzabadi, Max Allan, IEEE/RSJ International Conference on Intelligent Robots and Systems . 2020

机译：Davicinnet：机器人辅助手术中的运动和外科手术状态的联合预测
5. Near-future Prediction in Videos: Applications in Video Annotation and Frame Reconstruction [D] . ?Mahmud, Tahmida B. 2019

机译：视频近期预测：视频注释和帧重建中的应用
6. Future-Frame Prediction for Fast-Moving Objects with Motion Blur [O] . Dohae Lee, Young Jin Oh, In-Kwon Lee 2020

机译：带运动模糊的快速移动物体的未来帧预测
7. Future Frame Prediction for Robot-Assisted Surgery [O] . Xiaojie Gao, Yueming Jin, Zixu Zhao, 2021

机译：未来机器人辅助手术框架预测

Future Frame Prediction for Robot-Assisted Surgery

摘要

著录项

相似文献

相关主题

期刊订阅