Pith. sign in

Paper Citation Record · LEDGER

Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.09750.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.09750 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:51.489661Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:15.548001Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 61fdff65-6464-410d-9327-c5c0eda4d477 · inbound

ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models cites this paper.

ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T14:51:13.235289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:51:13.235289Z digest=sha256:c24c1f1a9e4f4bfb1032b647e0e9e84e312bd22cfb6a164cbe7ac3b93513c6ab

Observation 556cd2fa-a95f-4620-adb7-79ebe2802848 · inbound

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery cites this paper.

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T18:24:41.948056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:24:41.948056Z digest=sha256:36fe9765e4e422f9c22220d8c10805b70d0c8dfe81d2ab19e5756689c3e4a56c

Observation 445db077-f930-4e18-9523-cd54e4e81cc0 · inbound

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding cites this paper.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.489661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.489661Z digest=sha256:3e0bd11153bbce5e34da680faacbcd1073a1e939b9f4a7fc66a7ed4a0823a684

Observation 71f0ffbc-740c-47b1-ba44-c3de66dbfb82 · inbound

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation cites this paper.

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.784277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T03:15:40.944411Z digest=sha256:6a686324368a09880856b691abfbda4c153a4eec0e01589ac0bb592df8e2c0a5

Observation 94f56968-5d9a-4cc0-b202-6135877875ac · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 266

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.549626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:2238481f723637963270bb4377c84108525f23712d7d6a40de23535cd5ae84a8

Observation 0da2ed4b-95fd-479b-865a-c3da79e130ab · inbound

Fine-tuning a multimodal large language model for clinician-grade autism behavioral scoring from short home videos cites this paper.

Fine-tuning a multimodal large language model for clinician-grade autism behavioral scoring from short home videos Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-07-01T18:15:58.620703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T02:13:48.020651Z digest=sha256:75d058ce6eca52afe3e6ac77dd6695b74b76c3a23e7f3d838a40107609c2d9c6