Pith. sign in

Paper Citation Record · LEDGER

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning

As of 19 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2506.17645.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17645 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:32:35.305415Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8cc23c65-a2ca-44da-9458-d6c17bb96daa · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Bottom-up and top-down attention for image captioning and visual question answering

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.580028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.224862Z digest=sha256:f2f72e784ae57ad5deebbccae3771b09a75ef93fb6273be5416b1e3d562a9ad1

Observation 92a41822-b70e-43f3-9ee0-3f738d7de9b3 · outbound

This paper cites Improving diagnostic accuracy through feedback: The diagnosis learning cycle.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Improving diagnostic accuracy through feedback: The diagnosis learning cycle

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.565101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.231800Z digest=sha256:ef055302412214b6726c431b35299c1836e3c13f61398ace8d1dfaea083fc6a4

Observation 36faa479-fbb2-4991-a684-b397d6576d12 · outbound

This paper cites an unresolved cited work.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:32:35.550041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.238497Z digest=sha256:d55394fef29fd891e3bd63e202aebaeb9bd26a11b82fb9247cc4b6f464e17a28

Observation 4ef30ea5-f1dc-4879-8658-fe0b25ccc5b6 · outbound

This paper cites Wsicaption: Multiple instance generation of pathology reports for gigapixel whole-slide images.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Wsicaption: Multiple instance generation of pathology reports for gigapixel whole-slide images

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.534492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.243951Z digest=sha256:8755868c99c2d78199ae204805995681627975090d569efa8d8229bee22c075c

Observation 5a33e654-c73a-47e6-bd2e-6f1040da5214 · outbound

This paper cites Generating radiology reports via memory-driven transformer.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Generating radiology reports via memory-driven transformer

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.518570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.249315Z digest=sha256:adb341ac5cf51d3853f5a8c71a2265bcb999d72b79d1ede3364f7634630b55ee

Observation 8d710a77-3c1d-42d0-84df-e4241bc0daef · outbound

This paper cites Cross-modal memory networks for radiology report generation.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Cross-modal memory networks for radiology report generation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.503695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.254714Z digest=sha256:4460333801f874813af4e2808731a419ca817342a83e345b3243f4a59a8793c2

Observation e256619f-cfea-4e5f-9dc6-919addd5f14d · outbound

This paper cites Meshed-memory transformer for image captioning.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Meshed-memory transformer for image captioning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.487915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.259937Z digest=sha256:0044230e5b0b0229c09b985fa1c146482b7efa179b0162ad01fa9deffba6eb09

Observation 06a2b681-89ad-4a71-8eee-380cc1e5160f · outbound

This paper cites Anatomic pathology quality assurance through peer review.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Anatomic pathology quality assurance through peer review

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.472092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.264991Z digest=sha256:d8392fc03990b026d2c2b1a419f81b2dca5e2b6ae6e575f12001f0922c0e71fc

Observation ead99c06-f324-4be9-8ef3-d22252b89824 · outbound

This paper cites Histgen: A local-global encoding framework for pathology report generation.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Histgen: A local-global encoding framework for pathology report generation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.457182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.270208Z digest=sha256:44089ba74276a20728d625876b094200b7b52ecf1994ef3bfa013845e7687109

Observation 079348f2-6b39-4a8d-bc02-2f20f3a05237 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Lora: Low-rank adaptation of large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.441528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.275250Z digest=sha256:a30f21b0c18bab7ea7dc946c8ef2187e3bd209512b9ac58b240cda836db0c88d

Observation 22c56e16-71e5-47be-ad87-3b61b9a2f6df · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Retrieval-augmented generation for knowledge-intensive nlp tasks

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.424255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.280246Z digest=sha256:9037f042c23761364dddf222e5488e38666a3b3fbac124d076f569eccccfd19a

Observation b68315e8-bcbd-494f-af58-e4f1f0f60c03 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.408117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.284848Z digest=sha256:eb11823cb2102a4acae38faeb919f9c9d6de3fff17ae1104ac7107253402ff2c

Observation 761e1ef5-6bd5-4d2f-85dc-1d8f8116607c · outbound

This paper cites Improving factual completeness and consistency of image-to-text radiology report generation.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Improving factual completeness and consistency of image-to-text radiology report generation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.393332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.290054Z digest=sha256:56becc5ede2c034856c9ac309101c8f03d91d4855e4391317a66aaf21066f36d

Observation de8d5563-737d-4acf-b3df-d12cecef4cdd · outbound

This paper cites Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:32:35.294994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:32:35.294994Z digest=sha256:508c3a950052c6c77e6b0792c223888ffddbca4e4e5c4ff148c89391d40fc74a

Observation 61be891b-0c9a-4df0-afc1-15bf991767a0 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.376829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.300180Z digest=sha256:96dc87fb4c949b4e28547c7959f5e4a567bb3dd3ddde1c3e164726b21c2ede09

Observation 6c652973-3b90-4eae-b8c8-6cfe09c24fe4 · outbound

This paper cites Show and tell: A neural image caption generator.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Show and tell: A neural image caption generator

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:32:35.360454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T23:32:35.305415Z digest=sha256:f9e4da50ca3728a3edd10acb13420a614f0733679396688d3ab0fe518521abe6

Pith citing papers

No inbound Pith citation observations are available.