Pith. sign in

Paper Citation Record · LEDGER

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery

As of 22 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.10764.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10764 v4

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:46.857185Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e60926cf-af69-400e-8546-520ccbdaed2c · outbound

This paper cites Almeida, R.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Almeida, R

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.444692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.690704Z digest=sha256:5c66cecdaaa2ef1f94cb49e65e354f104f8277084c3c14205478f4984d9c5c91

Observation 389e6699-ac1d-4bad-a0d9-5d4e89935510 · outbound

This paper cites Abdulbaki Alshirbaji, H.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Abdulbaki Alshirbaji, H

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.430149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.695764Z digest=sha256:d166b2d88c75adc4a306c8c60ac5dc4e726b13469fd8d5f5fd4da12994862390

Observation dc060ca3-5ddf-47d6-87da-dc81a4650236 · outbound

This paper cites Vision-based and marker-less surgical tool detection and tracking: a review of the literature.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Vision-based and marker-less surgical tool detection and tracking: a review of the literature

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.415825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.700152Z digest=sha256:093a3a32439800470c197855a09226a299f7bb7b191fb676e56632de828ff074

Observation a954791e-9dd8-4009-b88d-3e8ad88260eb · outbound

This paper cites Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.401341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.704830Z digest=sha256:5766e960a1fda94757fe4085ca5ce4d2a2a9e1de635cb40cd5ba591eeab2c2b9

Observation 3a0440d2-1a60-4d91-be8b-5226ce1ca000 · outbound

This paper cites Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.386898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.709653Z digest=sha256:1102cd1195f764868a53874a5c89ade79b57271b94e83bfc9aa4b8061afa978a

Observation 77194edf-ce30-4848-a722-7f798ad3a5ee · outbound

This paper cites an unresolved cited work.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:47.372561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.714640Z digest=sha256:954fdbb7d837f6ffda9443dc3767e3c0dee665d5be4bd70ef8c9d9daa102019b

Observation a819e996-22dd-471c-9336-6f16c5e13ce2 · outbound

This paper cites Robotic surgery.Nature Reviews Bioengineering, pages 1–14, 2025.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Robotic surgery.Nature Reviews Bioengineering, pages 1–14, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.358573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.719094Z digest=sha256:12836d44dea3581ac06c6767a6047572ed9df5acddd011a3b2538b0e43480af9

Observation fa9b85ec-45bf-489d-9bac-94fb6e6961c7 · outbound

This paper cites Dosovitskiy, L.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Dosovitskiy, L

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.344476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.723809Z digest=sha256:1a97468fd73c9199eaad43233b7423d41e1e6b5f3147262746fad6866e2144c5

Observation 553bf8d8-c5d9-427b-8b1f-3ddfcc358a2b · outbound

This paper cites A decade retrospective of medical robotics research from 2010 to 2020.Science robotics, 6(60):eabi8017, 2021.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery A decade retrospective of medical robotics research from 2010 to 2020.Science robotics, 6(60):eabi8017, 2021

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.331021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.728329Z digest=sha256:dfc808f951534622b2a731085c3799beac4a60c6c0b7787ac4ad7ec2631883c6

Observation 183d11b3-0f93-45e0-92c5-59a258bf0897 · outbound

This paper cites Toolnet: holistically-nested real-time segmentation of robotic surgical tools.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Toolnet: holistically-nested real-time segmentation of robotic surgical tools

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.316972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.732767Z digest=sha256:8c0b139a7feb11f7d4c8bdd1e98b6b7c05c0c980509cf1c2c3f13e1fbbbe9214

Observation b2fda3f3-f2f3-4503-9550-e3822174cb38 · outbound

This paper cites Goyal, Z.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Goyal, Z

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.302874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.737435Z digest=sha256:aa346618a946fc75673fc0e8417c89f761635c833a9712d3470272d9678199e9

Observation 4f062208-5671-4dd0-a4b8-0476c0514ded · outbound

This paper cites an unresolved cited work.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:47.288017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.741974Z digest=sha256:af4a4cfac26a2dc3dd44c49a3d6e9eefeb74652790c1409a9afbb12071e9362b

Observation 458c409e-8b37-4913-9d8f-f20016f604c6 · outbound

This paper cites an unresolved cited work.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:47.274102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.746137Z digest=sha256:f260e06060b9d264988f08c24d147178ee32c7d4f55dad9c7a628d686e94ac9c

Observation 608db306-805d-4326-a4e6-71739dd52c19 · outbound

This paper cites Blip: Bootstrapping language-image pre- training for unified vision-language understanding and generation.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Blip: Bootstrapping language-image pre- training for unified vision-language understanding and generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.259851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.750389Z digest=sha256:d8c3f4ab03bebec59c19f44462aa6c20a894148fdee5261ecb33b038e8429dbb

Observation f2a217b0-a9d3-4a92-a33f-ffd31c07dac5 · outbound

This paper cites Lc-gan: Image- to-image translation based on generative adversarial network for endoscopic images.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Lc-gan: Image- to-image translation based on generative adversarial network for endoscopic images

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.245084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.754753Z digest=sha256:7881859a2ab7a0fad4f2d8c3760390c106ff11e8809e3cd0ec8c4b8f08d4949f

Observation 09a1d9fd-ebef-4da9-a0c9-e5a47dd2fd50 · outbound

This paper cites Multi-frame feature aggregation for real-time instrument seg- mentation in endoscopic video.IEEE Robotics and Automation Letters, 6(4):6773–6780, 2021.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Multi-frame feature aggregation for real-time instrument seg- mentation in endoscopic video.IEEE Robotics and Automation Letters, 6(4):6773–6780, 2021

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.229141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.758889Z digest=sha256:8b6668d6dce1e872d271cf46f4b91d487fd4bc54ff70b017b0980d8c27eef4a8

Observation a94bf806-2bdd-4a2a-9bfd-4597ab9f23d3 · outbound

This paper cites Visual instruction tuning.Ad- vances in Neural Information Processing Systems, 36:34892–34916, 2023.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Visual instruction tuning.Ad- vances in Neural Information Processing Systems, 36:34892–34916, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.212229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.762745Z digest=sha256:276ab02cd05e10fffeaa661803fa5d31e272d8b2fb0d34a12b63d895e844cea3

Observation da9bf049-ce0c-4f48-8008-298c0c4a0f37 · outbound

This paper cites Attention-guided lightweight network for real-time segmentation of robotic surgical instruments.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Attention-guided lightweight network for real-time segmentation of robotic surgical instruments

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.197362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.766780Z digest=sha256:87b447f0894a64a471e7fd9ab4991c7ef4420593cb5e89bf6168e31fbf8ec111

Observation b1a766a8-42b1-4ca1-ac1f-cdfe3fc0de06 · outbound

This paper cites an unresolved cited work.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:47.181937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.771115Z digest=sha256:61f7bff903a3ff5697d3744b2bc607c613d66c9c0d06204eb215fd0f3c28d35a

Observation 4b285379-a563-43d9-b861-90f1ed800187 · outbound

This paper cites Evaluation: from precision, recall and F-measure to ROC, informedness, markedness and correlation.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Evaluation: from precision, recall and F-measure to ROC, informedness, markedness and correlation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.775660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.775660Z digest=sha256:c3c5edd678e56df6685b9bb30494578eeec6ad3be3cd4790817b3ca8b340bcd5

Observation 89539047-cb6c-4671-aa2f-456c272ea106 · outbound

This paper cites Learning transferable visual models from natural language supervision.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Learning transferable visual models from natural language supervision

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.166701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.780120Z digest=sha256:cff9c5952b0193eee1cebfeae07b681825568a17d3c7b960569dfe3abe98ff0b

Observation 9915387b-9a07-495b-8aef-a00823df3d23 · outbound

This paper cites Systematic Evaluation of Large Vision-Language Models for Surgical Artificial Intelligence.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Systematic Evaluation of Large Vision-Language Models for Surgical Artificial Intelligence

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:07:46.973743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.784102Z digest=sha256:2cdc362392ae2c2ee6e74e3641a93dcbf10759a4d72dd58322cbe36807f1d4e3

Observation 70d5e674-dab5-4169-b54d-13606aaf62db · outbound

This paper cites an unresolved cited work.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:47.151974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.788310Z digest=sha256:dafa93a0048594d031873edb1827c26d07726d38d5e911b70a5345a9149c10cf

Observation 2dfb96d3-d373-4420-9005-6c7e06b41c7c · outbound

This paper cites Comparative validation of multi-instance instrument segmentation in endoscopy: results of the robust-mis 2019 challenge.Medical image analysis, 70:101920, 2021.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Comparative validation of multi-instance instrument segmentation in endoscopy: results of the robust-mis 2019 challenge.Medical image analysis, 70:101920, 2021

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.792484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.792484Z digest=sha256:7a689d47e2bfda825c7651812228d975170b90952ddce7044eac4b41e92d1193

Observation 9db1583e-184a-4bfc-9cba-4b171939d7f3 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.796405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.796405Z digest=sha256:ac1d19b3040165e6174402f3fae2efde3437966747894c58da623fc462d94e3c

Observation 5b9899bb-4eb7-4022-a506-2cbb82a0ec4e · outbound

This paper cites LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.800582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.800582Z digest=sha256:983b499f33defae706210a21fc602dd9796d85990888ea2cc5a4af151a8f82cf

Observation 42582ce1-4f11-41e5-9fbf-1fa5c293a479 · outbound

This paper cites Teed and J.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Teed and J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.118857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.805596Z digest=sha256:124856c172f72f19ad09e694644712d5a9efe09ad03eda81e181cfd60379c11a

Observation db790e7c-0331-4f1a-907e-9f49c8243820 · outbound

This paper cites Endonet: A deep architecture for recognition tasks on laparoscopic videos.IEEE Transactions on Medical Imaging, 36(1):86–97, 2016.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Endonet: A deep architecture for recognition tasks on laparoscopic videos.IEEE Transactions on Medical Imaging, 36(1):86–97, 2016

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.104054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.810596Z digest=sha256:59c09fc70b348264bf56b4429dacf176aea5fe63af5f1040b89868ecf8c6a16f

Observation 19021599-9a95-4c6b-945f-c72ad9611898 · outbound

This paper cites Data Splits and Metrics for Method Benchmarking on Surgical Action Triplet Datasets.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Data Splits and Metrics for Method Benchmarking on Surgical Action Triplet Datasets

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.814658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.814658Z digest=sha256:ba81c8b0bab5d93f228e5b893ebff9a7e091ea071a643e1aa4108b721792b4fb

Observation a13641bd-9548-40b1-8d0f-42f0479fd553 · outbound

This paper cites Yuille, and Wei Shen.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Yuille, and Wei Shen

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.087995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.819296Z digest=sha256:597e72876246c2f065d047aaf6d5d599bf67ec44042504e452d48382aa88f096

Observation 285878ed-c13d-4dca-8202-265a432a30cf · outbound

This paper cites The robot will see you now: Foundation models are the path forward for autonomous robotic surgery.Science Robotics, 10(104):eadt0684, 2025.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery The robot will see you now: Foundation models are the path forward for autonomous robotic surgery.Science Robotics, 10(104):eadt0684, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.073857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.823301Z digest=sha256:42b2ad05a3227ff1d7cccaf2368e0a355ac35fffa2db4592b7f3410ebf0b0104

Observation 2b2789ea-e37b-45d5-a90a-6cd2f679fb13 · outbound

This paper cites HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.827584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.827584Z digest=sha256:23107ec1e87d887d5171c72cdc39aa853ba1ef3c682bc28c21ac23045e444e76

Observation eeeb5bef-b636-4432-aead-4b0dd1315b2e · outbound

This paper cites Procedure-aware surgical video-language pretraining with hierarchical knowledge augmenta- tion.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Procedure-aware surgical video-language pretraining with hierarchical knowledge augmenta- tion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.058805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.831940Z digest=sha256:0e44c795249fcf0d5fb1c4e569fd0138cfb7ba76b5629f34affd8caccad608a3

Observation a28e53fd-cf19-4854-8bed-a00dceddfaeb · outbound

This paper cites Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.835946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.835946Z digest=sha256:da254e407a57513ff08cca388390750e7ce6a03a2a0571aacc77f972d528edec

Observation 821f38d1-7a0c-4ee2-907d-db88ade4234d · outbound

This paper cites Zablocki, H.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Zablocki, H

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.043450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.840469Z digest=sha256:a2d9f81dd79296d563122ad4f1a3550cd0fa5f77641e366b90ce6406f5549884

Observation 8f331546-61d5-4cbd-8b90-abc5d7bebe77 · outbound

This paper cites Zeiler and Rob Fergus.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Zeiler and Rob Fergus

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.028914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.844819Z digest=sha256:9ac5f3bed92d7e89e5d21c1b33a1060c0adf8d3293cbe8834c234bdcbee16f7c

Observation bfe2cdf9-2e4a-4420-ac75-f46915e17466 · outbound

This paper cites Vision-language models for vision tasks: A survey.IEEE transactions on pattern analysis and machine intelligence, 46(8):5625–5644, 2024.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery Vision-language models for vision tasks: A survey.IEEE transactions on pattern analysis and machine intelligence, 46(8):5625–5644, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.848532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.848532Z digest=sha256:334684528fd85cc9a46ef5635dec0306e49b0dc36c2a1b3bd895065bf4d1b5ce

Observation e9b8ccd5-ff8a-4537-a8ef-63e08937d4ab · outbound

This paper cites From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:46.852698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:46.852698Z digest=sha256:3888050b292f0348e837b4a477cfb4908ead5e19cb2f0ed95f39af0e56aed8fb

Observation 83e1d77b-2aca-4948-bfc3-9c9ed7703f8b · outbound

This paper cites What surgical tools do you see? Choose from: Grasper, Bipolar, Hook, Scissors, Clipper, Irrigator, Bag.

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery What surgical tools do you see? Choose from: Grasper, Bipolar, Hook, Scissors, Clipper, Irrigator, Bag

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:47.004475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:07:46.857185Z digest=sha256:7113f9bd4095b93bbdfa589c66dc81d68240e708a0225c5d70f1f462540ca2a9

Pith citing papers

No inbound Pith citation observations are available.