Scinovex
article Open AccessTop 1% cited

EndoNet: A Deep Architecture for Recognition Tasks on Laparoscopic Videos

IEEE Transactions on Medical Imaging · 2016 · Vol. 36(1) · pp. 86–97
Andru Putra TwinandaSherif ShehataDidier MutterJacques MarescauxMichel de MathelinNicolas Padoy

Abstract

Surgical workflow recognition has numerous potential medical applications, such as the automatic indexing of surgical video databases and the optimization of real-time operating room scheduling, among others. As a result, surgical phase recognition has been studied in the context of several kinds of surgeries, such as cataract, neurological, and laparoscopic surgeries. In the literature, two types of features are typically used to perform this task: visual features and tool usage signals. However, the used visual features are mostly handcrafted. Furthermore, the tool usage signals are usually collected via a manual annotation process or by using additional equipment. In this paper, we propose a novel method for phase recognition that uses a convolutional neural network (CNN) to automatically learn features from cholecystectomy videos and that relies uniquely on visual information. In previous studies, it has been shown that the tool usage signals can provide valuable information in performing the phase recognition task. Thus, we present a novel CNN architecture, called EndoNet, that is designed to carry out the phase recognition and tool presence detection tasks in a multi-task manner. To the best of our knowledge, this is the first work proposing to use a CNN for multiple recognition tasks on laparoscopic videos. Experimental comparisons to other methods show that EndoNet yields state-of-the-art results for both tasks.

Surgical Simulation and TrainingColorectal Cancer Screening and DetectionMedical Image Segmentation TechniquesComputer scienceConvolutional neural networkWorkflowArtificial intelligenceSearch engine indexingVisualizationTask (project management)Context (archaeology)Feature extractionDeep learning

MeSH terms

AlgorithmsLaparoscopyDatabases, FactualNeural Networks, Computer

Funding

  • Nvidia
  • Agence Nationale de la Recherche
Citations
1,005
FWCI
44.88
field-weighted impact
References
45
Percentile
100%
vs. same field & year
Citations per year
References
Error bounds for convolutional codes and an asymptotically optimum decoding algorithm
IEEE Transactions on Information Theory · 1967 · 6,722 citations
ImageNet Large Scale Visual Recognition Challenge
International Journal of Computer Vision · 2015 · 39,683 citations
ImageNet classification with deep convolutional neural networks
Communications of the ACM · 2017 · 75,550 citations
Object Detection with Discriminatively Trained Part-Based Models
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2009 · 9,994 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.