REVIEW 2 cited by
Self-Supervised Siamese Learning on Stereo Image Pairs for Depth Estimation in Robotic Surgery
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Robotic surgery has become a powerful tool for performing minimally invasive procedures, providing advantages in dexterity, precision, and 3D vision, over traditional surgery. One popular robotic system is the da Vinci surgical platform, which allows preoperative information to be incorporated into live procedures using Augmented Reality (AR). Scene depth estimation is a prerequisite for AR, as accurate registration requires 3D correspondences between preoperative and intraoperative organ models. In the past decade, there has been much progress on depth estimation for surgical scenes, such as using monocular or binocular laparoscopes [1,2]. More recently, advances in deep learning have enabled depth estimation via Convolutional Neural Networks (CNNs) [3], but training requires a large image dataset with ground truth depths. Inspired by [4], we propose a deep learning framework for surgical scene depth estimation using self-supervision for scalable data acquisition. Our framework consists of an autoencoder for depth prediction, and a differentiable spatial transformer for training the autoencoder on stereo image pairs without ground truth depths. Validation was conducted on stereo videos collected in robotic partial nephrectomy.
Forward citations
Cited by 2 Pith papers
-
Uncertainty-aware Test-Time Training (UT$^3$) for Efficient On-the-fly Domain Adaptive Dense Regression
UT3 selects keyframes via entropy of an uncertainty-aware masked-autoencoder self-supervision head, skipping test-time training on most frames and cutting inference time by about 70% with similar accuracy.
-
Med-Banana: Learning Quality-Controlled Medical Image Editing from Success-and-Failure Trajectories
Med-Banana-50K is a dataset of ~88K AI-generated medical image edits (accepts and rejects) across 23 diseases, labeled by a single commercial LLM judge with minimal expert validation.
Discussion (0). Continue with ORCID to comment.