Abstract
Laparoscopic video tracking primarily focuses on two target types: surgical instruments and anatomy. The former could be used for skill assessment, while the latter is necessary for the projection of virtual overlays. Where instrument and anatomy tracking have often been considered two separate problems, in this article, a method is proposed for joint tracking of all structures simultaneously. Based on a single 2D monocular video clip, a neural field is trained to represent a continuous spatiotemporal scene, used to create 3D tracks of all surfaces visible in at least one frame. Due to the small size of instruments, they generally cover a small part of the image only, resulting in decreased tracking accuracy. Therefore, enhanced class weighting is proposed to improve the instrument tracks. The authors evaluate tracking on video clips from laparoscopic cholecystectomies, where they find mean tracking accuracies of 92.4% for anatomical structures and 87.4% for instruments. Additionally, the quality of depth maps obtained from the method's scene reconstructions is assessed. It is shown that these pseudo-depths have comparable quality to a state-of-the-art pre-trained depth estimator. On laparoscopic videos in the SCARED dataset, the method predicts depth with an MAE of 2.9 mm and a relative error of 9.2%. These results show the feasibility of using neural fields for monocular 3D reconstruction of laparoscopic scenes. Code is available via GitHub: https://github.com/Beerend/Surgical-OmniMotion.
| Original language | English |
|---|---|
| Pages (from-to) | 411-417 |
| Number of pages | 7 |
| Journal | Healthcare Technology Letters |
| Volume | 11 |
| Issue number | 6 |
| Early online date | 12 Dec 2024 |
| DOIs | |
| Publication status | Published - Dec 2024 |
Keywords
- Computer vision
- Endoscopes
- Image reconstruction
- Learning (artificial intelligence)
- Neural nets
- Optical tracking
- Surgery
Fingerprint
Dive into the research topics of 'Neural fields for 3D tracking of anatomy and surgical instruments in monocular laparoscopic video clips'. Together they form a unique fingerprint.Research output
- 2 Citations
- 1 Preprint
-
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
Gerats, B. G. A., Wolterink, J. M., Mol, S. P. & Broeders, I. A. M. J., 28 Mar 2024, ArXiv.org.Research output: Working paper › Preprint › Academic
Open AccessFile85 Downloads (Pure)
Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver